Q-Learning - Model Free Reinforcement Learning and Temporal Difference Learning
Steve Brunton via YouTube
Learn Backend Development Part-Time, Online
Learn Excel and Financial Modeling the Way Finance Teams Actually Use Them
Overview
Google, IBM & Meta Certificates – 40% Off
One Coursera Plus subscription covers most Professional Certificates on Coursera.
Unlock All Certificates
This lecture explains model-free, gradient-free reinforcement learning through Monte Carlo learning, temporal-difference learning, SARSA, and Q-learning. It also discusses value functions, off-policy exploration, and connections between TD errors and biological learning.
Syllabus
Introduction
Recap
Monte Carlo Learning
Temporal Difference Learning
QLearning
SARSA
Off Policy
Conclusion
Taught by
Steve Brunton