#208 - Claude Integrations, ChatGPT Sycophancy, Leaderboard Cheats
May 8, 2025
Summary
This episode covers Anthropic’s new app integrations for Claude and OpenAI’s controversial “sycophantic” model update. It also discusses Baidu’s new Ernie models, Adobe’s Firefly updates, and Mia Maruti’s $2 billion raise for Thinking Machines Lab. Finally, the hosts analyze Huawei’s chip progress, Chinese stockpiling of Nvidia GPUs, and Elon Musk’s plans for a massive new supercomputer.
Topics discussed
Sponsors: Box AI agents and Outshift infrastructure
Intro: 988 hotline experience and episode overview
Anthropic's Computer Use and Agentic capabilities
OpenAI's persuasion update and safety concerns
Chinese AI models: Ernie X1 and Qwen 2.5
Adobe Firefly Ultra and Google's AI competition
OpenAI board drama and Mira Murati's role
Huawei chips and US export controls on China
Elon Musk's Colossus 2 and AI compute scaling
DeepSeek R1 and the DeLoco training method
Meta's Llama 4 and RL training efficiency
Reasoning models and multi-path exploration
AI espionage risks and cybersecurity assessments
Jailbreaking, alignment failures, and geopolitical ties
Outro and final thoughts on AI progress
Listen ad-free on Castria