NEURAL NETWORKS ARE WEIRD! - Neel Nanda (DeepMind)

Machine Learning Street Talk · 222:37

Neel Nanda, who leads DeepMind's mechanistic interpretability team, spends 3+ hours laying out what mech interp actually is, why sparse autoencoders (SAEs) became its central tool, and where they fall short — his core...

Read the full summary on tuber

Redirecting...