Courses/Reinforcement Learning

Reinforcement Learning

Policy gradient, Q-learning, and actor-critic methods built from the Bellman equation up. Train agents in MuJoCo and Atari.

Intermediate14 weeks · 18 lessons
Your progress00 / 18 · 0%
4 lessons here are on the Guided Foundations Path· Steps 11-14