53:53 Tutorials15 hours ago How to Build Better AI Evals with Claude Code in 5 Steps | Shreya & Hamel Peter Yang hosts Hamel Husain and Shreya Shankar — the instructors behind one of the most widely taken AI evaluation courses online,... 0 comments 778 views
47:11 Interviews1 week ago Exo: Harnesses should see their own code and logs — Alex Krentsel Alex Krentsel, a UC Berkeley PhD student and co-founder of Exo (backed by a16z partners Martin Casado and Anker Goel), joins the Late... 0 comments 319 views
40:12 Interviews2 months ago How this startup uses AI agents to eliminate bugs and optimize infrastructure Claire of How I AI interviews Anker Goyel, CEO of BrainTrust — a startup specializing in AI evaluation and observability tooling — ab... 0 comments 1.3K views
41:54 Agents & Automation3 months ago I Ranked Cloudflare’s Software Factory and Wow… S TIER TOKENOMICS IndyDevDan delivers a detailed engineering tier-list analysis of Cloudflare's AI-powered code review system, dissecting the architect... 0 comments 4.5K views
13:03 Foundation Models3 months ago Spec-Driven Testing for Agents With A Brain the Size of A Planet — Steven Willmott, SafeIntelligence Steven Willmott, CEO of SafeIntelligence, delivers a conference talk at AI Engineer on spec-driven testing for AI agents — arguing th... 0 comments 701 views
20:43 Foundation Models3 months ago How agent o11y differs from traditional o11y — Phil Hetzel, Braintrust Phil Hetzel, head of solutions engineering at Braintrust, delivers a conference talk breaking down exactly why agent observability is... 0 comments 732 views
18:34 Foundation Models3 months ago The maturity phases of running evals — Phil Hetzel, Braintrust Phil Hetzel, head of solutions engineering at Braintrust, walks through the maturity phases teams go through as they build out evalua... 0 comments 1.9K views
18:54 Business & Strategy3 months ago Does GenAI “belong” to data scientists? — Phil Hetzel, Braintrust Phil Hetzel, head of solutions engineering at Braintrust—an agent quality platform built around evals and observability—makes a delib... 0 comments 0.9K views
31:40 Foundation Models3 months ago Picking the right model Lucas from Anthropicʼs applied AI team delivers a practical framework for selecting the right Claude model in production — addressing... 0 comments 654 views
20:20 Foundation Models3 months ago These 5 Companies Secretly Control AI Nate B Jones argues that the real power brokers in the AI agent economy are not model companies like OpenAI or Anthropic, but a layer... 0 comments 14.5K views