Coursera
MOOC / Non-credit
0
Multimodal Generative AI: Vision, Speech, and Assistants
About this course
We are introducing a new course to replace the "Coding with ChatGPT" course in the Generative AI specialization. This updated course will cover materials, models, and content released in 2024. Some of the new additions include material on using AI for image-to-text (vision), text-to-speech, speech-to-text, and the Assistant API. All these topics come with new labs, lessons, and exercises.
C
60/100
CourseAsk score
- What the provider tells you
- 24/45
- Who stands behind it
- 20/35
- How complete the listing is
- 16/20
Scores how much the provider publishes and who stands behind it — not how well it is taught.
What you'll learn
- understand the principles of multimodal generative AI
- apply techniques for image-to-text and text-to-speech
- utilize speech-to-text capabilities
- work with the Assistant API effectively
Artificial Intelligence
#generative ai
#text-to-speech
#ai models
#multimodal
#speech-to-text
#assistant api
#image-to-text
#ai labs
#ai exercises
#2024 updates
$49.00
Price shown by Coursera — confirm on their site.
Enroll on CourseraYou'll be redirected to Coursera to complete enrollment.
- Listed & compared by CourseAsk
- English · 0
Compared on these lists
Where this course ranks against the alternatives.
edX
Coursera