Machine Learning DesignPRO
Part 18: Reinforcement Learning — From Rewards to Deep Q-Learning
A practical guide to reinforcement learning, policy search, policy gradients, MDPs, Q-Learning, Deep Q-Networks, DQN variants, and TF-Agents.
Deepak Mishra34 min read
Continue reading
This is a Pro article. Sign in and subscribe to Pro to read the full article.