Stanford CS221 | Autumn 2025 | Lecture 7: Markov Decision Processes

Stanford Online · 80:19

Markov decision processes generalize last week’s search problems by letting each action produce a distribution over next states instead of one deterministic successor, so a solution is a policy (state → action) whose...

Read the full summary on tuber

Redirecting...