#238 - GPT 5.4 mini, OpenAI Pivot, Mamba 3, Attention Residuals
Mar 26, 2026
Summary
This episode covers OpenAI’s release of efficient GPT-4.5 Mini and Nano models, highlighting their speed and token efficiency despite higher per-token costs. Mistral introduces its open-source Small 4 family, consolidating reasoning, coding, and multimodal capabilities into a single sparse model. The discussion also explores the "land grab" for local AI agent runtimes, featuring Meta’s Manus and NVIDIA’s Nemo Claw, alongside NVIDIA’s DLSS 5 graphics update. Finally, the hosts analyze OpenAI’s strategic pivot toward enterprise productivity and the controversial plans for an adult mode in Cha…
Topics discussed
Sponsors: Box, OutShift, and Grammarly
Intro: Hosts, AI news preview, and research focus
OpenAI GPT-5 Mini and Nano: Pricing and Performance
Meta's Small Models and Multimodal Consolidation
NVIDIA GTC: OpenClaw, Nemo, and AI Infrastructure
NVIDIA DLSS 5 and AI-Generated Graphics
OpenAI's Shift to Adult Content and Internal Strategy
OpenAI vs. Anthropic: Market Share and Focus
NVIDIA Blackwell, Grok LPU, and Memory Architecture
Cohere, Enterprise AI, and Data Sovereignty
ByteDance AI Clusters and Export Controls
Meta's Superintelligence Labs and Hiring
Google Gemini Restructuring and Microsoft AI Leadership
Research: Detecting Secret LLM Communication
Research: Early Exiting and Reasoning Confidence
Research: Defenses Against Fine-Tuning Misalignment
Research: AI Agents in Cyber Ranges and Benchmark Gaming
Anthropic's Automated Evaluation and Opus 4.6 Safety
Policy: Congressional Push for AI Export Control Transparency
Deep Dive: Residual Attention and Memory Optimization
Deep Dive: Mamba 3 State Space Models and Complex Numbers
Listen ad-free on Castria