Lecture
Mediaspace scheduled maintenance: Aug 25, 2026 07:00 - 12:00 AM. During this time, videos will be temporarily unavailable. Check status updates.
This lecture covers the application of reinforcement learning to teach Pacman to play autonomously, focusing on policy gradient methods and Markov decision processes. It discusses the challenges faced, such as the large parameter space, and proposes solutions like log linear parametrization and vectorization.