39:23 Benchmarks6 days ago Agentic Engineering Benchmarks: How I RANK Astra, Fable 5.1, and Open-Weights IndyDevDan presents a practical framework for selecting AI benchmarks tailored to agentic engineering, arguing that broad composite i... 0 comments 4.5K views
14:30 Benchmarks1 week ago MiniCPM5-2B: The Best Sub-Agent Model Yet? Sam Witteveen examines MiniCPM 2B (technically 2.5B), the latest small model from OpenBMB, which claims to beat 4B-class models on fu... 0 comments 4.1K views
30:39 News & Opinion1 week ago AI Model Month is Already Delivering Big Gains The AI Daily Brief covers a dense opening to September 2026, led by OpenAI's claim to have solved the Navier-Stokes equations — one o... 0 comments 0.9K views
19:32 Reviews & Comparisons2 weeks ago GPT-6 Astra Is Finally Here (And It’s REALLY Good) OpenAI launched GPT-6 Astra on September 3rd, 2026, and Matt Wolfe secured early access to test it before the public rollout to ChatG... 0 comments 98.4K views
35:31 Reviews & Comparisons2 weeks ago Claude Fable 5.1 is savage The AI Search channel runs Claude Fable 5.1 through an extensive battery of demanding tests using Claude Code in Ultra mode, aiming t... 0 comments 30.6K views
16:51 Reviews & Comparisons2 weeks ago GOOGLE IS BACK! (Gemini 3.8 Flash) Matthew Berman breaks down Google's newly released Gemini 3.8 Flash, a lightweight model drawing attention for strong performance on... 0 comments 8.5K views
18:58 Reviews & Comparisons3 weeks ago Cancel your subscriptions, Ox-Alpha is here! (GLM 5.3 Flash) Matthew Berman covers GLM 5.3 Flash, the latest open-weights model from Chinese AI lab ZAI that briefly appeared on OpenRouter under... 0 comments 78.8K views
13:09 News & Opinion1 month ago AI News: ChatGPT Ultrafast, Grok 4.6, 3 New Open-Source Models, and more! Matthew Berman's August 14, 2026 weekly roundup covers one of the busiest product weeks in recent AI history, with four model release... 0 comments 20.5K views
17:15 Reviews & Comparisons1 month ago xAI actually did it… (Grok 4.6) Matthew Berman covers the release of Grok 4.6, xAI's latest model and a significant step up from its predecessor Grok 4.5. The video... 0 comments 54.6K views
20:02 News & Opinion1 month ago Improving Agents is a Data Mining Problem — Vivek Trivedy, LangChain Vivek Trivedy, who leads applied research at LangChain, presented a systematic framework for continuously improving AI agents by trea... 0 comments 1.4K views