Nội dung series
4 phần · 14 bài học
1.Phần 1: Nền tảng Multimodal Learning
3 bài2.Phần 2: Vision-Language Models (VLMs)
4 bài- Bài 6: LLaVA & Open-source VLMs180 phút
3.Phần 3: Document AI & Video Understanding
3 bài4.Phần 4: Advanced Multimodal & Production
4 bàiTác giả
DUY TRAN
Pursuing an AI-first mindset and intelligent system architecture. I build solutions by combining technology, creativity, and the ability to see structure in chaos — the foundation for becoming a Solution Architect.
