Lecture
Mediaspace scheduled maintenance: Aug 25, 2026 07:00 - 12:00 AM. During this time, videos will be temporarily unavailable. Check status updates.
This lecture provides an overview of reinforcement learning, focusing on the BackUp diagram as a graphic representation of the steps an RL algorithm remembers. Topics covered include deep reinforcement learning, neural networks, policy branching probabilities, total expected reward, Bellman equation, SARSA algorithm, and the application of SARSA for estimating Q values.
Network: Computation in Neural Systems', Journal of Computational Neuroscience', and `Science'.