Skip to main content
Graph
Search
fr
en
Login
Search
All
Categories
Concepts
Courses
Lectures
MOOCs
People
Quizes
Exercises
Publications
Startups
Units
Show all results for
Home
Lecture
Actor-Critic Architecture and Advantage-Actor-Critic
Graph Chatbot
Related lectures (29)
Reinforcement Learning: Policy Gradient and Actor-Critic Methods
Provides an overview of reinforcement learning, focusing on policy gradient and actor-critic methods for deep artificial neural networks.
Reinforcement Learning: Q-Learning
Covers Q-Learning in reinforcement learning, exploring action values, policies, and the societal impact of algorithms.
Principled Reinforcement Learning with Human Feedback
Delves into Reinforcement Learning with Human Feedback, discussing convergence of estimators and introducing a pessimistic approach for improved performance.
TD Learning: Temporal Difference Learning
Covers Temporal Difference Learning, V-values, state-values, and TD methods in reinforcement learning.
Reinforcement Learning: TD Learning and SARSA Variants
Discusses reinforcement learning, focusing on temporal difference learning and SARSA algorithm variations.
Reinforcement Learning: SARSA Algorithm
Explores the SARSA algorithm for reinforcement learning, focusing on updating Q-values and the importance of exploration in learning by rewards.
Deep Reinforcement Learning: Mini-Batches and Policy Methods
Discusses deep reinforcement learning methods, focusing on mini-batches and the implications of on-policy and off-policy training techniques.
Policy Gradient Methods: Binary Actor Example
Introduces policy gradient methods using a simple example of a single neuron with binary output.
Risk Minimization from Adaptively Collected Data
Explores risk minimization from adaptively collected data with guarantees for policy learning and the importance of exploration strategies.
Policy Gradient Methods: Single Neuron Example
Covers policy gradient methods using a single neuron with binary output.
Memory-Efficient Adaptive Optimization
Explores memory-efficient adaptive optimization for large-scale learning and the challenges of memory overhead in training large models.
Mathematics of Data: Models and Learning
Explores models, learning paradigms, and applications in Mathematics of Data.
Stochastic Softmax Tricks
Explores stochastic softmax tricks, reparametrization, and argmax, addressing challenges in expectation estimation and gradient variance.
Reinforcement Learning: BackUp Diagrams
Introduces the BackUp diagram as a key graphic representation in reinforcement learning.
Structures in Non-Convex Optimization
Covers non-convex optimization, deep learning training problems, stochastic gradient descent, adaptive methods, and neural network architectures.
Reinforcement Learning Fundamentals
Delves into the fundamentals of reinforcement learning, discussing states, actions, rewards, policies, and neural network applications.
Introduction to Reinforcement Learning: Concepts and Applications
Log in to Mediaspace to watch this video
Introduces reinforcement learning, covering its concepts, applications, and key algorithms.
Continuous Reinforcement Learning: Advanced Machine Learning
Log in to Mediaspace to watch this video
Explores continuous-state reinforcement learning challenges, value function estimation, policy gradients, and Policy learning by Weighted Exploration.
Deep Learning Agents: Reinforcement Learning
Log in to Mediaspace to watch this video
Explores Deep Learning Agents in Reinforcement Learning, emphasizing neural network approximations and challenges in training multiagent systems.
Linear Programming Techniques in Reinforcement Learning
Log in to Mediaspace to watch this video
Covers the linear programming approach to reinforcement learning, focusing on its applications and advantages in solving Markov decision processes.
Previous
Page 1 of 2
Next