Skip to main content
Graph
Search
fr
en
Login
Search
All
Categories
Concepts
Courses
Lectures
MOOCs
People
Quizes
Exercises
Publications
Startups
Units
Show all results for
Home
Lecture
Monte-Carlo Methods for Reinforcement Learning
Graph Chatbot
Related lectures (29)
Biased Monte Carlo Markov Chain
Log in to Mediaspace to watch this video
Explores Biased Monte Carlo Markov Chain, including Bayes-optimal estimation and Metropolis-Hastings algorithm.
Bayes Estimator, Simulated Annealing and EM
Log in to Mediaspace to watch this video
Covers Bayes estimator, Simulated Annealing, and EM for parameter estimation.
Interactive Lecture: Reinforcement Learning
Log in to Mediaspace to watch this video
Explores advanced reinforcement learning topics, including policies, value functions, Bellman recursion, and on-policy TD control.
Introduction to Reinforcement Learning: Concepts and Applications
Log in to Mediaspace to watch this video
Introduces reinforcement learning, covering its concepts, applications, and key algorithms.
Policy Gradient Methods in Reinforcement Learning
Log in to Mediaspace to watch this video
Covers policy gradient methods in reinforcement learning, focusing on optimization techniques and practical applications like the cartpole problem.
Safe Learning and Control
Log in to Mediaspace to watch this video
Explores safe learning, control, multi-agent coordination, and Nash equilibrium convergence in intelligent systems.
Discrepancy Function: Estimation and Applications
Log in to Mediaspace to watch this video
Explores estimating discrepancy functions and their practical applications in generating sets with low discrepancy.
Model Selection: Evaluation and Generalization
Log in to Mediaspace to watch this video
Explores model selection, evaluation, and generalization in machine learning, emphasizing unbiased performance estimation and the risks of over-learning.
Stochastic Simulation: Monte Carlo Method
Log in to Mediaspace to watch this video
Covers the properties and error estimates of the Monte Carlo method in stochastic simulation.
Previous
Page 2 of 2
Next