Share this link via
Or copy link
The "Multimodal Generative AI: Vision, Speech, and Assistants" course introduces learners to the exciting frontier of AI systems that understand and generate content across multiple modalities—such as text, images, audio, and video. This course is ideal for those interested in how cutting-edge AI powers tools like virtual assistants, image captioning systems, voice generators, and AI chatbots that integrate vision and speech. You will explore the foundations of multimodal AI, including how it combines data from different sources and uses models like...
Codio via Coursera
14 hours 25 minutes
Paid Certificate Available
English
On-Demand
Beginner
Kevin Noelsaint
No reviews yet. Be the first to review!
You must be logged in to submit a review.