02:09:37 Interviews3 weeks ago Watching & Learning from Agents with CEOs of Arthur & Datacamp The Cognitive Revolution hosts Nathan and Pash open with a deep dive into leaked intelligence around Anthropic's internal "Model 2" —... 0 comments 138 views
06:08 News & Opinion2 months ago Don’t Let the LLM Drive – Ornella Bahidika & Joel Allou, Microsoft Microsoft engineers Ornella Bahidika and Joel Allou make the case that agent reliability is a control-flow problem, not a prompting p... 0 comments 596 views
05:45 News & Opinion2 months ago Your Voice Agent Doesn’t Need a Frontier Model – Joel Allou & Ornella Bahidika, Microsoft Microsoft engineers Ornella Bahidika and Joel Allou present the architecture behind Ace, a live AI voice tutor they built to run reli... 0 comments 414 views
01:00:35 Tutorials2 months ago Build an AI Agent That Runs 24/7 With Tank Creator Magic host Mike walks through building AI agents that run continuously around the clock using Tank, an open-source AI coding... 0 comments 665 views
10:50 Reviews & Comparisons2 months ago Laguna XS 2.1: Poolside’s Local Coding Agent Tested – Nine Languages Fahd Mirza puts Poolside's newly released Laguna XS 2.1 through a live evaluation using the Hermit agentic framework. The model is a... 0 comments 1.2K views
26:44 News & Opinion2 months ago Fable Is Back: Here’s What You Should Try First The AI Daily Brief covers three major stories from July 1–2, 2026. The lead story examines OpenAI's reported inference optimization b... 0 comments 10.5K views
10:41 News & Opinion3 months ago The First Real LLM Breakthrough Is Here… SubQ (1000x Less Compute) TheAIGRID covers the release of SubQ 1.1 Small, which its developers claim is the first large language model built on a fully sub-qua... 0 comments 55.7K views
30:17 News & Opinion3 months ago AI News: Microsoft Finally Reveals Their Plan! Matt Wolfe reports firsthand from Microsoft Build in San Francisco and Nvidia's Computex event in Taiwan, covering a dense week of AI... 0 comments 10.3K views
17:03 Benchmarks3 months ago Finally a good benchmark (DeepSWE) Matthew Berman breaks down DeepSWE, a new long-horizon software engineering benchmark released by data-curve.ai that claims to fix th... 0 comments 14.5K views
24:02 Deep Dives4 months ago The thinking lever Anthropic product manager Matt Bleifer delivers a detailed technical explanation of how Claude uses test-time compute — also called i... 0 comments 855 views