Last Week in AI Last Week in AI

#208 - Claude Integrations, ChatGPT Sycophancy, Leaderboard Cheats

May 8, 2025

Summary

This episode covers Anthropic’s new app integrations for Claude and OpenAI’s controversial “sycophantic” model update. It also discusses Baidu’s new Ernie models, Adobe’s Firefly updates, and Mia Maruti’s $2 billion raise for Thinking Machines Lab. Finally, the hosts analyze Huawei’s chip progress, Chinese stockpiling of Nvidia GPUs, and Elon Musk’s plans for a massive new supercomputer.

Topics discussed

Sponsors: Box AI agents and Outshift infrastructure Intro: 988 hotline experience and episode overview Anthropic's Computer Use and Agentic capabilities OpenAI's persuasion update and safety concerns Chinese AI models: Ernie X1 and Qwen 2.5 Adobe Firefly Ultra and Google's AI competition OpenAI board drama and Mira Murati's role Huawei chips and US export controls on China Elon Musk's Colossus 2 and AI compute scaling DeepSeek R1 and the DeLoco training method Meta's Llama 4 and RL training efficiency Reasoning models and multi-path exploration AI espionage risks and cybersecurity assessments Jailbreaking, alignment failures, and geopolitical ties Outro and final thoughts on AI progress
Listen ad-free on Castria