Lec 07. Scaling Rules for Optimization

MIT OpenCourseWare · 80:55

This lecture argues that modern deep learning has largely treated approximation and generalization as “solved by scale,” so the remaining bottleneck is optimization: finding weights that minimize a composite, high-dim...

Read the full summary on tuber

Redirecting...