19:04 Foundation Models3 months ago Evals Are Broken, Use Them Anyway — Ara Khan, Cline Ara Khan, an engineer on the Cline team, delivers a pointed critique of how the AI industry uses evaluation benchmarks — and why most... 0 comments 1.4K views
12:40 Foundation Models3 months ago What Lies Beneath the API — Benjamin Cowen, Modal Benjamin Cowen, a forward-deployed machine learning engineer at Modal, delivers a conference talk examining one of the most consequen... 0 comments 373 views
01:09:33 Interviews3 months ago Devin’s 80% Moment: Background Agents, 7x PRs, & End of Hand-Held Coding — Walden Yan & Cole Murray Walden Yan, co-founder and CPO of Cognition (the company behind Devin), and Cole Murray, creator of Open Inspect, join the Latent Spa... 0 comments 210 views
18:02 Coding & Dev Tools3 months ago Fast Models Need Slow Developers — Sarah Chieng, Cerebras Sarah Chang, Head of Developer Experience at Cerebras, argues that the arrival of ultra-fast AI coding models demands a fundamental r... 0 comments 647 views
24:37 Foundation Models3 months ago AI Dev 26 x SF | Ara Khan: Evals Are Broken Use Them Anyway Ara Khan, speaking at the AI Dev 26 x SF event hosted by DeepLearningAI, argues that most developers are fundamentally wrong about AI... 0 comments 449 views
29:04 Business & Strategy3 months ago How to get to production faster with Claude Managed Agents Michael and Harrison, both members of technical staff at Anthropic, deliver a keynote-style deep dive into Claude Managed Agents — th... 0 comments 2.1K views
27:23 Coding & Dev Tools3 months ago Build a production-ready agent with Claude Managed Agents An Anthropic engineer leads a live coding workshop at a launch event, walking developers through building a production-ready multi-ag... 0 comments 3.3K views
46:27 Business & Strategy3 months ago Code with Claude London 2026: Opening Keynote Anthropic held its first international Code with Claude event in London in May 2026, with the opening keynote covering both the broad... 0 comments 20K views
21:48 Research & Benchmarks3 months ago I Tested 3 Ways to Deploy Claude Agents (Here’s When to Use Each) Nate Herk walks through three practical methods for deploying Claude Code agents so they run autonomously on a schedule — a question... 0 comments 4.5K views
37:01 Tutorials4 months ago Hermes Agent: The Self-Improving AI Agent (Explained) Hermes Agent is an open-source personal AI agent gaining traction as a self-improving alternative to OpenClaw, featuring a built-in m... 0 comments 11.9K views