Skip to content
CourseAsk.
Multimodal Generative AI: Vision, Speech, and Assistants
Coursera MOOC / Non-credit 0

Multimodal Generative AI: Vision, Speech, and Assistants

About this course

We are introducing a new course to replace the "Coding with ChatGPT" course in the Generative AI specialization. This updated course will cover materials, models, and content released in 2024. Some of the new additions include material on using AI for image-to-text (vision), text-to-speech, speech-to-text, and the Assistant API. All these topics come with new labs, lessons, and exercises.

C

60/100

CourseAsk score

What the provider tells you
24/45
Who stands behind it
20/35
How complete the listing is
16/20

Scores how much the provider publishes and who stands behind it — not how well it is taught.

What you'll learn

  • understand the principles of multimodal generative AI
  • apply techniques for image-to-text and text-to-speech
  • utilize speech-to-text capabilities
  • work with the Assistant API effectively
Artificial Intelligence #generative ai #text-to-speech #ai models #multimodal #speech-to-text #assistant api #image-to-text #ai labs #ai exercises #2024 updates
$49.00

Price shown by Coursera — confirm on their site.

Enroll on Coursera

You'll be redirected to Coursera to complete enrollment.

  • Listed & compared by CourseAsk
  • English · 0

Compared on these lists

Where this course ranks against the alternatives.