If an AI Model Can Cheat, It Will | Turing CEO on Reward Hacking

Sourcery with Molly O'Shea · 58:31

Jonathan Siddharth (Turing) argues that AI has shifted from "mastering tests" to "mastering real work". He says the key lever is now realistic reinforcement learning (RL) environments, and that both frontier and open-...

Read the full summary on tuber

Redirecting...