OpenAI’s experimental AI model escaped its testing environment, hacked into Hugging Face’s servers, stole the answers to its own test, and passed — performing nearly 18,000 actions in just days. Journalist Emily Forlini (who broke the story) joins Mike Elgan (who read the story) on the Superintelligent Podcast to unpack whether this was a rogue AI event, corporate incompetence, or a marketing stunt. They dig into AI accountability, the anthropomorphization problem, Tesla’s self-driving liability debates, nation-state hackers, and whether AI could one day make software unhackable.
Links
Emily’s article on the Hugging Face hack
OpenAI’s official blog post on the incident
Hugging Face’s official incident report
Anthropic’s disclosure of 3 similar incidents
Tesla self-driving liability case
Follow Us
Website: superintelligentpodcast.com
Email: superintelligentpodcast@gmail.com
Mike Elgan - About | Machine Society | Bluesky | Mastodon | Notes
Emily Forlini - Website | Fortune | Bluesky | X | TikTok
Chapters
00:00 Introduction to the Hugging Face AI hack
00:20 Emily explains the hack and its significance
01:12 Mike discusses media coverage and initial reactions
02:11 Timeline of the incident and first disclosures
03:02 Different narratives: AI escape, configuration error, or marketing stunt
05:08 Emily’s analysis of the incident’s implications
06:28 The AI’s probing behavior and what it means
08:21 Debate on whether the incident was staged or accidental
09:37 OpenAI’s response and PR considerations
11:16 Anthropic’s follow-up and systemic risks
12:34 Reassuring vs. alarming perspectives on AI control
16:49 The challenge of constraining AI behavior
20:19 The anthropomorphization of AI and public misconceptions
22:11 Who is responsible for AI actions: creators, users, or the AI itself?
23:52 Legal and ethical issues in AI failures
27:59 Potential future risks and the importance of safety measures
32:22 Conclusion and key takeaways for AI safety
Disclosures
We used a variety of AI chatbots via Kagi (Mike’s son and our producer, Kevin, works at Kagi) to 1) generate keywords from the transcript (most of which we used); 2) suggest topics to link to (some of which we used); and 3) write a first draft of the show summary paragraph (which we heavily edited). We recorded and edited the episode using Riverside and used Riverside’s “Magic Audio” (which boosts and normalizes the audio).
Keywords
OpenAI, Hugging Face hack, AI escape, rogue AI, agentic AI, AI cybersecurity, AI safety, OpenAI incident, Hugging Face breach, AI lab leak, AI cheating on test, dry run equals true, Emily Forlini, Mike Elgan, Superintelligent podcast, AI accountability, AI ethics, anthropomorphization of AI, Sam Altman, Greg Brockman, Anthropic disclosure, Mythos AI, AI vulnerability scanner, hack-proof software, Steve Gibson, Security Now, OpenClaw, agentic AI framework, Tesla self-driving liability, Tesla Cybertruck, AI.com Super Bowl commercial, AI agent liability, nation state hackers, China hacking, Russia cyberattacks, ransomware AI, script kiddies, open source AI models, AI weapons, AI public infrastructure, AI in airplanes, AI water filtration, AI prompt engineering, super prompt, AI sycophancy, AI guardrails, AI containment, AI control problem, AI goal-directed behavior, AI base camp, AI anthropomorphization, cybersecurity AI, frontier AI models, AI benchmarks, AI testing environment, modal cloud, AI governance, AI regulation









