Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 18: Frontiers
Stanford Online · 70:49
Deep RL now works well where rewards are cheap to verify (games, math, code), but the hard remaining problems are how we define objectives, how methods use prior knowledge and scale beyond short-horizon online loops,...