Lecture
Mediaspace scheduled maintenance: Aug 25, 2026 07:00 - 12:00 AM. During this time, videos will be temporarily unavailable. Check status updates.
This lecture covers the importance of mini-batches in Deep Reinforcement Learning, explaining how to avoid data correlation by using replay buffers or multiple actors. It discusses on-policy and off-policy methods, such as Q-Learning and Advantage Actor-Critic, and the pros and cons of each approach.
Network: Computation in Neural Systems', Journal of Computational Neuroscience', and `Science'.