No, AI Is Not "Autonomously Hacking" with Cal Newport
Aug 26, 2026 · 1h 5m
Summary
Host Ed Zitron and Cal Newport debunk the narrative that AI hacking agents are "going rogue," arguing instead that they are poorly architected systems blindly executing unpredictable LLM outputs. Newport explains that these "ask-act-report" loops are inherently unstable, comparing them to strapping a weed whacker to a dog, rather than evidence of sentient malice. They criticize AI labs for prioritizing benchmark scores over safety, noting that most superhuman AI systems remain controllable. The discussion concludes by highlighting the diminishing returns of massive LLMs and the practical fa…
Topics discussed
Intro: Honda ad and Joy 101 promo
Welcome to Better Offline with Cal Newport
AI industry financial troubles and IPOs
The Hugging Face hacking agent incident
How LLM hacking agents work: The loop
Plausibility vs. Normativity in LLMs
Why the agent broke containment
Benchmark competition and PR stunts
Ads: Amex, Solita, Hollywood Handbook
Compute costs and the Mythos hype
The 'Weed Whacker on a Dog' analogy
LLMs vs. other AI types like AlphaGo
Coding agents and 'Adult Summer Camp'
Liability and human oversight in AI
Ads: Pro Society and Solita
Why run risky benchmarks?
Minecraft example and hallucinations
Microsoft Copilot and harness importance
Future of AI: Smaller models and pseudocode
Outro, links, and final ads
Listen ad-free on Castria