Explores learning strategies in robotics, including reward systems, fixture usage, and training generalization from simulation to real-world applications.
Provides an overview of policy gradient methods in reinforcement learning, focusing on the log-likelihood trick and the transition from batch to online learning.
Introduces reinforcement learning, covering its definitions, applications, and theoretical foundations, while outlining the course structure and objectives.