Chevron Left
返回到 Build Multimodal Generative AI Applications

学生对 IBM 提供的 Build Multimodal Generative AI Applications 的评价和反馈

4.7
34 个评分

课程概述

Ready to level up your GenAI skills? Step into the exciting world of multimodal AI, where language, images, and speech come together to build smarter, more interactive applications. In this hands-on course, you’ll learn how to build systems that work across multiple modalities, from creating AI-powered storytellers and meeting assistants to developing image captioning tools and video generation apps. You’ll gain experience with real-world tools like IBM’s Granite, OpenAI’s Whisper, Sora and DALL·E, Meta’s Llama, Mistral’s Mixtral, and Gradio. Plus, you'll explore multimodal search, question answering, and retrieval systems that combine text, speech, and visual data. By the end of the course, you’ll be able to design and build full-stack multimodal AI solutions using Python and frameworks like Flask and Gradio. If you’re looking to gain in-demand skills for building the next generation of AI applications, enroll today and power up your AI career!...

热门审阅

筛选依据:

1 - Build Multimodal Generative AI Applications 的 3 个评论(共 3 个)

创建者 Muhammad A H

Oct 27, 2025

Wow, It was next Level Experience to learn the Multimodal Gen AI Development. Truly Amazing.

创建者 Mansib M

Oct 15, 2025

Well-structured, easy to digest.

创建者 Sajjan M

Sep 22, 2025

Not that useful in creating AI applications