STOR 743
Reinforcement Learning and Markov Decision Processes.
University of North Carolina at Chapel Hill · UGRD · Fall 2026
1 section
Catalog description
Markov decision processes (stochastic dynamic programming): finite horizon, infinite horizon, discounted and average-cost criteria; reinforcement learning(RL): design and analysis of model-free, model-based, value-based, and policy-based RL algorithms, RL algorithms in continuous and discrete state and action space, and RL with functional approximation. These algorithms include but are not limited to (deep) Q-learning, asynchronous advantage actor-critic, soft actor-critic, and proximal policy optimization.
Sections
Current meeting, instructor, credit, and enrollment details
001
Availability not recently verifiedClass #north_carolina_chapel_hill-10355Fall 2026UGRD3 credits
- Days & times
- No scheduled meeting time
- Meeting dates
- —
- Location
- —
- Instructor
- Staff
Class numbers and section codes come from the registrar.
Spot missing or incorrect course data?