Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 2: Imitation Learning

Stanford Online · 67:05

Imitation learning trains a policy to match expert demonstrations so it can achieve high reward without a hand-specified reward function; the lecture shows why ordinary L2 regression fails on multimodal expert data, h...

Read the full summary on tuber

Redirecting...