Last Week in AI Last Week in AI

#243 - GPT 5.5, DeepSeek V4, AI safety sabotage

May 3, 2026

Summary

Hosts Andrei Krenkov and Jeremy Harris discuss the launch of GPT-5.5, its performance in coding, and safety concerns regarding chain-of-thought faithfulness. They cover XAI’s Grok Voice Think Fast 1.0, which claims significant leads in real-time conversation benchmarks, and DeepSeek’s open-source DeepSeek V4, featuring a massive 1.6-trillion parameter model with a 1-million-token context window. The episode also touches on the Elon Musk vs. Sam Altman trial and Anthropic’s new integrations with creative tools like Photoshop.

Topics discussed

Intro, hosts, and sponsor segments GPT-5.5 release, benchmarks, and safety evals xAI Grok updates and voice model performance DeepSeek V4 architecture and Tencent's Pi3 New benchmark for complex reasoning tasks Google and Anthropic compute partnership Meta's AWS Graviton chip deal for AI US restrictions on Chinese AI acquisitions OpenAI-Microsoft deal and Elon Musk trial DOJ antitrust case against OpenAI Google Gemini private appliance for enterprise David Silver's new AI startup raises $5.1B AI sabotage risks and document editing errors Temporal sparse autoencoders for interpretability Policy memos and AI chatbot dating concerns AI economic impact and government security access Deep neural lesions and model fragility research
Listen ad-free on Castria