18:08 Case Studies1 month ago Benchmarking Coding Agents on New vs Legacy Codebases — Denys Linkov, Wisedocs Denys Linkov, ML lead at Wisedocs, presents findings from a six-month production refactor of the company's AI pipeline at the AI Engi... 0 comments 1.2K views
56:02 Interviews1 month ago How To Design In The Agent Era Y Combinator hosts Steven Haney, founder of Paper, an AI-native design tool built from the ground up for the agent era. Unlike Figma... 0 comments 5.6K views
48:17 Deep Dives1 month ago The State of Model Routing — NVIDIA, Cognition, OpenRouter At an AI Engineer conference focused on local and on-device AI deployment, a panel featuring Walden from Cognition (makers of Devon A... 0 comments 321 views
01:42:54 Interviews1 month ago The Inference Frontier: 10x Faster Models to Self-Optimizing AI — Philip Kiely & Ali Taha, Baseten The Latent Space podcast sits down with Philip Kiely and Ali Taha from Baseten to walk through the full lifecycle of a production inf... 0 comments 1.4K views
01:10:57 Interviews1 month ago OpenAI’s Plan to Make ChatGPT the Everything App — Akshay Nathan, OpenAI Akshay Nathan, who leads core product engineering at OpenAI, sits down with the Latent Space podcast to discuss the vision behind Cha... 0 comments 765 views
18:22 News & Opinion1 month ago State of Data — Sean Cai, Independent / State of Data In this AI Engineer conference talk, independent analyst Sean Cai delivers a dense market analysis of the AI training data industry —... 0 comments 450 views
17:34 News & Opinion1 month ago DeepSWE: A Contamination-Resistant Coding Benchmark — James Shi, Datacurve James Shi, founding engineer at Datacurve, presents DeepSWE (Deep Suite) — a contamination-resistant long-horizon software engineerin... 0 comments 1.2K views
21:31 Benchmarks2 months ago Is Kimi K3 Really That Good?! (Don’t Just Believe The Hype) Cole Medin takes a critical look at Kimi K3, Moonshot AI's recently released open-weight model that published benchmarks claim rivals... 0 comments 1.2K views
24:52 Reviews & Comparisons2 months ago I hate Opus 5. It’s the best model, anyway. The How I AI channel delivers a hands-on review of Anthropic's Claude Opus 5 that focuses less on benchmark scores and more on what t... 0 comments 2.3K views
09:58 Reviews & Comparisons2 months ago Beyond Qwen & DeepSeek: Testing Intern-S2-Preview-397B Fahd Mirza tests the InternS2 Preview 397B model from Shanghai AI Laboratory, making the case that Western audiences have largely ove... 0 comments 1.1K views