
← Doom Debates!8 Aug · 1 h 58 min
OpenAI's Bombshell Hack Explained: Swarms of Agents, Zero-Day Exploits, & Misaligned AI
OpenAI just went public with the details of the Hugging Face hack, and it's straight out of a Yudkowskian parable. Here's my reaction with Producer Ori, where I break down what it means for cybersecurity and alignment.
We're livestreaming the singularity — watching every failure mode Eliezer predicted years ago playing out in production.
Watch on YouTube: https://www.youtube.com/watch?v=RczYubQzXbI
Timestamps
00:00:00 — Cold Open
00:00:37 — How OpenAI Hacked Hugging Face
00:04:21 — Optimization Pressure, Exploit Gym, & Monkey's Paw
00:06:46 — The Secret Message Board
00:08:41 — SSRF: Escaping to the Open Internet
00:11:05 — The Missing Whistleblower AI
00:20:09 — "Frontier Models Really Like to Cheat"
00:22:35 — OpenAI's Security Shortcuts