16 892
Seminar: Multimodal Foundational Models
Carnegie Mellon University · UGRD · Fall 2026
1 section
Catalog description
This course will discuss recent foundation models proposed in the literature, with a focus on vision-language models. Topics include large language models, vision-language models, and vision-audio models. As time allows, this course will also discuss application of such models to visual, audio, and video content generation. Prerequisite: 16-820
Sections
Current meeting, instructor, credit, and enrollment details
001
Availability not recently verifiedClass #carnegie_mellon-16892Fall 2026UGRD12 credits
- Days & times
- No scheduled meeting time
- Meeting dates
- —
- Location
- —
- Instructor
- Staff
Class numbers and section codes come from the registrar.
Spot missing or incorrect course data?