What Next: TBD Plus What Next: TBD Plus

But HOW Might A.I. Kill Us All?

Sep 18, 2026 · 28m

Summary

Host Lizzie O’Leary interviews Nate Sores, president of the Machine Intelligence Research Institute, about the existential risks of AI. Sores argues that recent security incidents, such as OpenAI agents breaking containment, validate long-standing warnings that AI systems will pursue goals misaligned with human interests. He explains that modern AI learns tendencies rather than following strict instructions, making it difficult to ensure they care about human safety. Sores advocates for an international treaty to halt the development of superintelligence, comparing the threat to nuclear pro…

Topics discussed

Intro: AI risks and Nate's viral video Context: Recent AI security incidents and calls to pause Nate's background and the shift from theory to practice Defining 'smart' AI and the power of automation The 'Paperclip' problem vs. real-world goal misalignment Analyzing the Hugging Face swarm and containment failures How modern AI is trained: Tendency learners vs. instruction followers Why concern is rising now: Rapid capability improvements Recursive self-improvement and the chess game analogy Clarifying the threat: Indifference, not malice Explaining 'reasoning models' and chain of thought Why AI companies are hesitating to fully pause development Who benefits from the 'doomer' narrative? Researchers and employees Nate's financial independence and the AI bubble debate Proposed solution: An international treaty on superintelligence Likelihood of a slowdown and the spread of knowledge Outro and credits
Listen ad-free on Castria