Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 12: Multi-Task RL

Stanford Online · 70:29

This lecture finishes model-based RL by showing how a learned dynamics model can generate short synthetic rollouts that augment real data and train a reusable policy, then covers multitask imitation and RL: condition...

Read the full summary on tuber

Redirecting...