Summary
Matt Wolfe breaks down the week’s major AI model releases, arguing that some have been significantly overhyped while others have flown under the radar. He focuses on Anthropic’s Fable 5.1, Google DeepMind’s Gemini 3.8 Flash, and the teased OpenAI Astra announcement, sharing both benchmark analysis and hands-on testing results.
Fable 5.1 draws mixed praise: Wolfe acknowledges it leads on agentic scientific research and coding benchmarks, and notes that Anthropic has reduced refusal rates by roughly 60% and introduced anti-distillation mechanisms to prevent other labs from training on Claude’s reasoning traces. However, the model’s pricing — unchanged from Fable 5 at $10 per million input tokens and $50 per million output tokens — led to a $100 bill during one of his own coding tests, which he considers prohibitive for most users.
The video’s central argument is that Gemini 3.8 Flash is the week’s most underappreciated release. On the Deep Sweep agentic benchmark, it scores 73.7%, putting it on par with Claude Opus 5, but at an average cost per task of $2.36 versus $11.84 for Opus 5 and $21.63 for Fable 5. Wolfe contends that Google’s long absence from the conversation has caused commentators to overlook a model that outperforms frontier competitors on coding tasks at a fraction of the price.
📺 Source: Matt Wolfe · Published September 03, 2026
🏷️ Format: News Analysis







