Explores training robots through reinforcement learning and learning from demonstration, highlighting challenges in human-robot interaction and data collection.
Delves into Reinforcement Learning with Human Feedback, discussing convergence of estimators and introducing a pessimistic approach for improved performance.