How I AI How I AI

Claude Opus 5 review: this model is brilliant (but annoying)

Jul 24, 2026 · 24m

Summary

The host reviews Anthropic’s Opus 5, highlighting its high capability but "neurotic" and overly cautious personality compared to GPT-5.6. Despite frustration with verbose outputs, Opus 5 tops the live "How I AI" benchmark for design and coding tasks. The episode contrasts model personalities, concluding that while Opus 5 is excellent for autonomous work, its conversational style remains exasperating for direct interaction.

Topics discussed

Intro: Fatigue with rapid model releases and intelligence overhang Opus 5 introduction and plan to analyze model personality Opus 5's neurotic and timid behavior in coding tasks Comparing Opus 5 and GPT personalities via direct interviews Testing trust dynamics and 'Claude Slop' verbosity issues Summary of Opus 5's neuroticism vs high-quality output How I AI benchmark methodology and setup Benchmark results: Opus 5 and GPT-5.6 Soul lead the pack Analysis of design outputs and final verdict on Opus 5 Conclusion, TLDR, and call to action
Listen ad-free on Castria