Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 17: Advancing Robot Intelligence
Stanford Online · 49:48
Successful on-policy reinforcement learning—paired with well-specified rewards and evaluation that scales with compute—has produced superhuman Go play, long-horizon LLM reasoning, and real-world robot dexterity that i...