16 824

Visual Learning and Recognition

Carnegie Mellon University · UGRD · Fall 2026

1 section
Add to a schedule

Catalog description

This graduate-level computer vision course explores representation and reasoning for large-scale data, such as images, videos, 3D data, and text, toward understanding the visual world surrounding us. Students will engage with a diverse selection of classic and recent research papers covering mid-level vision (grouping, segmentation), object and scene recognition, 3D scene understanding, action recognition, multimodal perception, vision-language models, multimodal deep generative models, efficient deep learning, and more. We will explore state-of-the-art neural architectures, including CNNs and transformers, and a wide range of supervised, semi-supervised, self-supervised, and unsupervised approaches for each topic above. Prerequisites: 16-720 Min. grade B or 16-722 Min. grade B or 10-701 Min. grade B or 16-385 Min. grade B or 15-781 Min. grade B Course Website: https://visual-learning.cs.cmu.edu/

Sections

Current meeting, instructor, credit, and enrollment details

Updated 4 hours ago

001

Availability not recently verified
Class #carnegie_mellon-16824Fall 2026UGRD12 credits
Days & times
No scheduled meeting time
Meeting dates
Location
Instructor
Staff
Class numbers and section codes come from the registrar.
Spot missing or incorrect course data?