Stanford CS221 | Autumn 2025 | Lecture 7: Markov Decision Processes
Stanford Online · 80:19
Markov decision processes generalize last week’s search problems by letting each action produce a distribution over next states instead of one deterministic successor, so a solution is a policy (state → action) whose...