If an AI Model Can Cheat, It Will | Turing CEO on Reward Hacking
Sourcery with Molly O'Shea · 58:31
Jonathan Siddharth (Turing) argues that AI has shifted from "mastering tests" to "mastering real work". He says the key lever is now realistic reinforcement learning (RL) environments, and that both frontier and open-...