aa6d95544d
- Add Use/Build/Understand journey pages and topic pages with a 101/201/301 catalog, surfaced through a Journey x Level grid in the README - Redistribute the free-course and notebook lists into the new navigation; add 2025-2026 courses and remove paid or dead entries - Backfill monthly best-papers lists from March 2025 through June 2026 - Extend the RAG, AI evaluation, and agentic search research tables to mid-2026 - Archive the 2024 course and paper material with banners - Fix the citation block and drop stale calls to action
1.3 KiB
1.3 KiB
Topic: Multimodal
Models that work across text, images, audio, and video: diffusion models, vision-language models, and multimodal application patterns. A thin subject in the repo today, covered mainly by external courses plus a 2024 guide.
Tags: Format 📖 Read / 🎥 Video / 🛠️ Notebook / 📝 Practice · Source ⭐ LevelUp original / 🌐 External · year.
🏗️ Build 101
- Multimodal LLMs Guide ⭐ 📖 (2024) 🔴 stale: an introduction to multimodal models (pre native-multimodal-model era).
- How Diffusion Models Work 🌐 🎥 by DeepLearning.AI
- How to Use Midjourney, AI Art and ChatGPT to Create an Amazing Website 🌐 🎥 by Brad Hussey
🏗️ Build 201
- 11-777: Multimodal Machine Learning 🌐 🎥 by Carnegie Mellon University
- Prompt Engineering for Vision Models 🌐 🎥 by DeepLearning.AI
Related topics: Foundations · Prompting. Journeys: Build. Back to the repository index.