AppStories AppStories

Local AI Gets Serious on the M5 Ultra Mac Studio

Sep 21, 2026 · 39m

Summary

Federico Vitici reviews the new M5 Pro Mac mini and M5 Ultra Mac Studio, focusing on their local AI performance. He benchmarks models like Kimi K3 and Qwen 3.8, highlighting massive improvements in token pre-fill and generation speeds over the M3 Ultra. The episode also compares the Mac Studio to an NVIDIA 5090 PC, noting the Mac's advantages in unified memory, size, and quiet operation despite the GPU's raw speed.

Topics discussed

Intro: Mac Mini and Mac Studio reviews Hardware specs and review unit details Focus on local AI performance testing Mac Mini M5 Pro vs M4 Pro benchmarks John's M4 Pro Mac Mini setup workflow Sponsor: Astroad Workbench remote desktop Mac Studio M5 Ultra specs and GPU architecture Running local agents for 99 days on M3 Ultra M5 Ultra memory bandwidth and prefill speeds Benchmark visualizations and interactive charts Running Qwen 3.8 Flash Next locally Quantization levels and RAM usage analysis Future potential: 512GB RAM and Thunderbolt 5 Daily driver apps and voice model integration Sponsor: Claude from Anthropic Comparison with NVIDIA RTX 5090 GPU Mac Studio vs PC: Size, noise, and preference Outro and App Stories Plus teaser
Listen ad-free on Castria