Skip to content
CourseAsk.
Multimodal Generative AI: Vision, Speech, and Assistants
Coursera MOOC / Non-credit all levels

Multimodal Generative AI: Vision, Speech, and Assistants

About this course

We are introducing a new course to replace the "Coding with ChatGPT" course in the Generative AI specialization. This updated course will cover materials, models, and content released in 2024. Some of the new additions include material on using AI for image-to-text (vision), text-to-speech, speech-to-text, and the Assistant API. All these topics come with new labs, lessons, and exercises.

$49.00

Price shown by Coursera — confirm on their site.

Enroll on Coursera

You'll be redirected to Coursera to complete enrollment.

  • Listed & compared by CourseAsk
  • English · All Levels