Building Multimodal AI Applications: Integrating Text, Vision, and Audio
Learn how to combine text, image, and speech models to build intelligent, cross-modal applications using modern AI frameworks and vector databases.
-
💬
ผู้สอน AI
ถามเกี่ยวกับบทเรียนใดก็ได้ แล้วรับคำตอบที่ชัดเจนทันที ทุกเมื่อ -
🕐
เริ่มเมื่อไรก็ได้
ไม่มีตารางหรือเดดไลน์ — เรียนตามจังหวะของคุณ เมื่อไรก็ได้ -
🌐
เป็นภาษาไทย
บทเรียน แบบฝึกหัด และใบรับรอง — ทั้งหมดเป็นภาษาของคุณอย่างครบถ้วน
เกี่ยวกับคอร์สนี้
AI is evolving beyond simple text-based interactions to understand the world through multiple senses. To build modern, intelligent applications, you need to know how to connect vision, audio, and language models seamlessly. This text-only course guides you through the foundational concepts and practical architectures of multimodal and cross-modal AI. You will understand how different data modalities are represented, aligned, and integrated to solve complex, real-world problems. What you'll learn: Understand the core principles of multimodal AI, including embeddings, joint representations, and cross-modal alignment; Process and combine text, image, and audio inputs using modern machine learning frameworks; Implement cross-modal retrieval systems using vector databases to search images with text queries; Apply prompt engineering techniques tailored specifically for multimodal models; Design robust AI architectures that can seamlessly transition between different data modalities. You will start with key terminology and foundational definitions of data modalities before moving into practical code-based architectures and integration strategies. Designed for developers and AI enthusiasts new to multimodal systems, this course requires no advanced prerequisites beyond basic programming familiarity. Start reading today to unlock the potential of multi-sensory artificial intelligence.
สิ่งที่คุณจะได้รับ
-
📜
ใบประกาศนียบัตร
เพิ่มในโปรไฟล์ LinkedIn ของคุณ -
💬
ติวเตอร์ AI ส่วนตัว
ติดขัดในบทเรียน? ถามติวเตอร์ในตัวของคุณได้ทุกอย่าง ทุกเวลา -
🎧
รวมเวอร์ชันเสียง
เรียนได้ทุกที่ ไม่ต้องดูจอ -
♾️
เข้าถึงตลอดชีพ
กลับมาเรียนได้ตลอด ไม่มีหมดอายุ -
📱
โทรศัพท์หรือคอมพิวเตอร์
ใช้งานได้ทุกที่ ทุกอุปกรณ์ -
💸
คืนเงิน 14 วัน
ไม่ต้องอธิบาย -
⚡
กระชับและตรงประเด็น
2 ชม. 30 นาที เนื้อหาเชิงปฏิบัติ
รีวิว
ยังไม่มีรีวิว — เป็นคนแรกที่แชร์ประสบการณ์
ผู้เรียนคนอื่นเรียน
🎓 มีใบรับรอง
AI ส่วนตัวด้วย LLMs แบบโอเพนซอร์ส: การติดตั้งในเครื่อง, RAG, และเอเจนต์
ใบรับรอง
ลงมือทำ
$14.99
→
💼 พร้อมสำหรับงาน
🎓 มีใบรับรอง
การปรับแต่งโมเดล OpenAI (Fine-Tuning): ปรับแต่ง LLMs ด้วยข้อมูลของคุณเอง
ใบรับรอง
ลงมือทำ
$14.99
→
🏆 ยอดนิยมมากที่สุด
🎓 มีใบรับรอง
การพัฒนาระบบ RAG ด้วย Azure OpenAI และ Azure AI Search
ใบรับรอง
ลงมือทำ
$14.99
→
💼 พร้อมสำหรับงาน
🎓 มีใบรับรอง
การพัฒนาแอปพลิเคชัน AI ด้วย LangChain
ใบรับรอง
ลงมือทำ
$14.99
→
คำถามที่พบบ่อย
ฉันต้องใช้อะไรในการเรียนคอร์สนี้? +
แค่โทรศัพท์หรือคอมพิวเตอร์ที่มีอินเทอร์เน็ต ไม่ต้องติดตั้งหรือใช้อุปกรณ์พิเศษ
ฉันชำระเงินอย่างไร? +
ผ่านบัตรด้วย Stripe เราไม่เก็บข้อมูลบัตร — Stripe จัดการอย่างปลอดภัย
ฉันขอคืนเงินได้ไหม? +
ใช่ — คืนเงินเต็มจำนวนใน 14 วัน ไม่ต้องอธิบาย
ฉันมีสิทธิ์เข้าถึงนานเท่าไร? +
ตลอดไป เมื่อซื้อแล้วคอร์สเป็นของคุณ กลับมาเรียนได้ตลอด
ฉันจะได้ใบประกาศนียบัตรไหม? +
ได้ เมื่อเรียนจบจะได้รับใบประกาศนียบัตรที่เพิ่มในโปรไฟล์ LinkedIn ได้
ออกแบบสำหรับผู้เรียนใน
เทคโนโลยี
ดีไซน์
การเงิน
การตลาด
สาธารณสุข
การศึกษา
ธุรกิจการบริการ
อุตสาหกรรม