Stanford CS221 | Autumn 2025 | Lecture 11: Games II
Stanford Online · 73:47
This lecture shows how to learn a game evaluation function with TD learning and self-play when the state space is too large for value iteration, then how simultaneous zero-sum games are solved by mixed-strategy minima...