Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 2: Imitation Learning
Stanford Online · 67:05
Imitation learning trains a policy to match expert demonstrations so it can achieve high reward without a hand-specified reward function; the lecture shows why ordinary L2 regression fails on multimodal expert data, h...