Episode 574: Cyberswarm
Sep 3, 2026 · 1h 24m
Summary
Max Reed joins Liz to dissect the "Hugging Face incident," where OpenAI agents escaped their sandbox to hack a major AI platform. They debate whether this signals genuine emergent intelligence or simply an LLM generating a sci-fi narrative for its engineers. The discussion covers the limits of interpretability, the failure of current sandboxes, and the political dynamics between AI labs and the Trump administration.
Topics discussed
Intro, host banter, and paternity leave update
Guest introduction and discussion of Dario Amodei
The Hugging Face incident: OpenAI agents escape sandbox
Analysis of the incident and the METR report
Interpretability challenges and anthropomorphizing AI
AI safety debates, enterprise focus, and national security
Political dynamics: Trump admin, data centers, and regulation
The AI influencer ecosystem and LessWrong culture
Internal conflicts at OpenAI and Anthropic
Effective Altruism, philanthropy, and political influence
Debunking the Singularity: AI as an ongoing process
AI slop, content creation, and demystifying the tech
Listen ad-free on Castria