53:53 Tutorials2 months ago How to Build Better AI Evals with Claude Code in 5 Steps | Shreya & Hamel Peter Yang hosts Hamel Husain and Shreya Shankar — the instructors behind one of the most widely taken AI evaluation courses online,... 0 comments 863 views
47:11 Interviews2 months ago Exo: Harnesses should see their own code and logs — Alex Krentsel Alex Krentsel, a UC Berkeley PhD student and co-founder of Exo (backed by a16z partners Martin Casado and Anker Goel), joins the Late... 0 comments 388 views
02:06:11 Interviews2 months ago Lindy Teammate: Flo Crivello on Multiplayer Agents, Memory & Why He’d Ban the Chinese Models He Uses Nathan Labenz of the Cognitive Revolution interviews Flo Crivello, founder and CEO of Lindy, on the occasion of launching Lindy Teamm... 0 comments 8.2K views
40:12 Interviews4 months ago How this startup uses AI agents to eliminate bugs and optimize infrastructure Claire of How I AI interviews Anker Goyel, CEO of BrainTrust — a startup specializing in AI evaluation and observability tooling — ab... 0 comments 1.3K views
41:54 Case Studies4 months ago I Ranked Cloudflare’s Software Factory and Wow… S TIER TOKENOMICS IndyDevDan delivers a detailed engineering tier-list analysis of Cloudflare's AI-powered code review system, dissecting the architect... 0 comments 4.5K views
13:03 Deep Dives4 months ago Spec-Driven Testing for Agents With A Brain the Size of A Planet — Steven Willmott, SafeIntelligence Steven Willmott, CEO of SafeIntelligence, delivers a conference talk at AI Engineer on spec-driven testing for AI agents — arguing th... 0 comments 754 views
20:43 Deep Dives4 months ago How agent o11y differs from traditional o11y — Phil Hetzel, Braintrust Phil Hetzel, head of solutions engineering at Braintrust, delivers a conference talk breaking down exactly why agent observability is... 0 comments 780 views
18:34 Deep Dives4 months ago The maturity phases of running evals — Phil Hetzel, Braintrust Phil Hetzel, head of solutions engineering at Braintrust, walks through the maturity phases teams go through as they build out evalua... 0 comments 1.9K views
18:54 News & Opinion5 months ago Does GenAI “belong” to data scientists? — Phil Hetzel, Braintrust Phil Hetzel, head of solutions engineering at Braintrust—an agent quality platform built around evals and observability—makes a delib... 0 comments 1K views
31:40 Deep Dives5 months ago Picking the right model Lucas from Anthropicʼs applied AI team delivers a practical framework for selecting the right Claude model in production — addressing... 0 comments 727 views