Lec 13. Representation Learning: Theory
MIT OpenCourseWare · 75:20
A randomly initialized neural-network architecture already encodes which inputs it treats as similar, and that implicit similarity becomes a covariance kernel in the infinite-width limit: the network is then a Gaussia...