CS 6674
Multimodal Computer Vision
Cornell University · UGRD · Fall 2026
1 section
Catalog description
Multimodal representations are reshaping computer vision, driving advances in both understanding and generation across a wide range of perceptual tasks. This research-oriented course explores computer vision techniques that integrate images with additional modalities such as language and 3D geometry for addressing challenges in both analysis and synthesis tasks. Possible topics include visual grounding, multimodal alignment, and text-guided generation and editing over multiple 2D and 3D representations.
Sections
Current meeting, instructor, credit, and enrollment details
001
Availability not recently verifiedClass #cornell_2-CS6674Fall 2026UGRD3 credits
- Days & times
- No scheduled meeting time
- Meeting dates
- —
- Location
- —
- Instructor
- Staff
Class numbers and section codes come from the registrar.
Spot missing or incorrect course data?