Multi-modal AI
Coursera · beginner · 3h
Learn to build production applications by combining visual and textual inputs with AI coding tools. You will explore multi-modal programming where screenshots, images, and text serve as inputs for AI-assisted code generation, and set up development environments configured for visual AI workflows. The course covers prompt engineering with visual context to improve code generation accuracy, and hands-on development with GitHub Copilot in VS Code for inline suggestions and chat-based interactions.
Skills covered
Disclaimer
Suggestions only — review each course yourself to judge whether it meets the role's requirements. Completing a course doesn't guarantee proficiency or that you'll qualify; hiring standards vary by employer.
We may earn a commission through some course links.