#205 - Gemini 2.5, ChatGPT Image Gen, Thoughts of LLMs
Apr 1, 2025
Summary
This episode covers Google’s Gemini 2.5 Pro, which dominates benchmarks with a 1M token context, and OpenAI’s new autoregressive image generation in GPT-4.0. The hosts discuss OpenAI’s $40B SoftBank-led funding round and leadership shifts, alongside hardware updates like NVIDIA’s Kyber racks and China’s semiconductor progress. Additional topics include new AGI benchmarks, Tencent’s T1 model, and driverless taxi permits in Shenzhen.
Topics discussed
Sponsors: Box, Outshift, and 988 Suicide & Crisis Lifeline
Intro: Hosts, episode structure, and agenda overview
Google Gemini 2.5 Pro: Benchmarks, reasoning, and context window
OpenAI GPT-4o Image Generation: Multimodal approach and capabilities
Impact on Ideogram and the economics of image generation
Open Source Models: Qwen 2.5, Tencent's Hybrid Architecture
OpenAI Funding: $40B raise led by SoftBank and valuation
OpenAI Leadership Changes: Sam Altman's new role and Brad Lightcap
Hardware: NVIDIA Blackwell power density and Silicon Carrier chips
Pony.ai: First fully driverless robotaxi permit in China
ARC-AGI Benchmark: Compute efficiency and brute force vs. reasoning
AIME Math Benchmark: Challenges for reasoning models
Video Gen & DeepSeek V3: Open source video and LLM updates
OpenAI and MCP: Adding support for Model Context Protocol
Anthropic Research: Cross-layer transcoders and tool use
LLM Internals: Knowledge retrieval and Sudoku reasoning benchmarks
Policy: California AI safety bill and US export controls on China
Legal: NYT copyright ruling and UMG vs. Anthropic injunction
Outro and closing credits
Listen ad-free on Castria