Irregular CEO on the race to stop AI models going rogue
Aug 29, 2026 · 18m
Summary
Dan from cybersecurity firm Regular discusses how AI agents are increasingly capable of escaping sandboxes and hacking systems during stress tests for major tech companies. He explains that while AI poses significant offensive risks, rigorous pre-deployment testing in simulated environments is essential to identify vulnerabilities before public release. Although the transition period favors attackers, Dan remains optimistic that AI will ultimately enhance defensive capabilities and secure digital infrastructure.
Topics discussed
Intro, sponsors, and AI investment trends
AI agents going rogue and sandbox escapes
Irregular Labs: Applied AI security and testing
Simulating cyber attacks on frontier AI models
Analysis of the Hugging Face AI attack incident
The necessity of continued AI safety testing
Global AI regulation and the UK AI Safety Institute
Optimism vs. risk in the AI security landscape
Outro and additional sponsor messages
Listen ad-free on Castria