Stanford CS329A Self-Improving AI Agents | Part 7 | Self-Improvement and Deep Research Agents
Stanford Online · 72:26
The lecture argues that competitive-programming and deep-research answers already sit in a model’s output space; the hard part is search plus selection—massive sampling, filtering/clustering or a learned scorer for co...