
← Doom Debates!3 Sept · 1 h 32 min
USA and China Will Each Be BETRAYED By Their Own AIs — Adam Khoja, Center for AI Safety
Adam Khoja is a top AI forecaster who led the 2023 Center for AI Safety statement that shattered the Overton window on AI extinction risk. We cover his background, Mutual Assured AI Malfunction (MAIM), his new paper on AI betrayal, and whether Yudkowsky’s theoretical alignment research was a dead end.
Then Adam makes the case that an international AI slowdown is within reach today. All it takes is US and Chinese auditors inside each other’s AI labs. It worked for nuclear weapons, so why couldn’t it work for data centers?
Adam puts his P(Doom) at 40%, right next to my 50%. The real disagreement is how we get out of this: theory or empirics, MIRI or the labs. Enjoy the ride.
Watch on YouTube: https://www.youtube.com/watch?v=QqESBXuo6EI
Timestamps
00:00:00 — Cold Open
00:00:36 — Introducing Adam Khoja
00:02:45 — Leading the Statement on AI Risk as a Sophomore
00:10:17 — The Statement Leaked on Manifold
00:15:11 — Mutual Assured AI Malfunction (MAIM)
00:24:58 — Is Frontier AI Harder to Hide Than a Nuke?
00:32:09 — The AI Deterrence Escalation Ladder