20:56 Foundation Models2 months ago Stop Making Models Bigger, Make Them Behave — Kobie Crawdord, Snorkel Kobie Crawford, developer advocate at Snorkel AI, presented research conducted in partnership with UC Berkeley's RLLM lab showing tha... 0 comments 739 views
34:00 Foundation Models2 months ago Claude Fable 5 – Full 319 page Breakdown AI Explained dedicates a full video to breaking down Anthropic's 319-page Claude Fable 5 system card — reading it cover to cover and... 0 comments 90.6K views
11:13 Foundation Models3 months ago RAG is dead, right?? — Kuba Rogut, Turbopuffer Kuba Rogut, a deployed engineer at Turbopuffer, pushes back on the \"RAG is dead\" wave that swept AI social media in late 2025 and e... 0 comments 1.3K views
15:50 Foundation Models3 months ago Road to 5 Million Tokens: Breaking Barriers in Long Context Training — Max Ryabinin, Together AI Max Ryabinin, VP of Research and Development at Together AI, presents the company's research project on extending transformer trainin... 0 comments 726 views
24:51 Foundation Models3 months ago Why Eval++ Is the Next Great Compute Primitive — Sunil Pai & Matt Carrie, Cloudflare Cloudflare engineers Sunil Pai and Matt Carrie take the stage at the AI Engineer conference to explain how Cloudflare's Durable Objec... 0 comments 1.7K views
26:27 Foundation Models3 months ago Why More Context Makes Your Agent Dumber and What to Do About It — Nupur Sharma, Qodo Nupur Sharma, an engineer at Qodo with a background in DevSecOps, presents hard-won lessons from deploying production agentic code re... 0 comments 1.4K views
16:32 Foundation Models3 months ago LLM Observability, Evaluation, Experimentation Platform — Dat Ngo, Arize Dat Ngo, AI architect at Arize AI, presents a structured framework for making LLM systems observable, evaluable, and experimentally i... 0 comments 473 views
19:04 Foundation Models3 months ago Evals Are Broken, Use Them Anyway — Ara Khan, Cline Ara Khan, an engineer on the Cline team, delivers a pointed critique of how the AI industry uses evaluation benchmarks — and why most... 0 comments 1.4K views
06:49 Foundation Models3 months ago Nanowhale-100m: Fascinating Implemention of DeepSeek-V4 Architecture Fahd Mirza walks through Nanowhale-100M, a 110 million parameter language model built entirely from scratch—no borrowed weights—that... 0 comments 1.1K views
25:20 Foundation Models3 months ago Beyond Transcription: Building Voice AI That Understands Conversations — Hervé Bredin, pyannoteAI Hervé Bredin, chief science officer and co-founder of pyannoteAI, presents a conference talk exploring what becomes possible when voi... 0 comments 598 views
22:38 Foundation Models3 months ago Building Agent Interfaces: Lessons from Chrome DevTools (MCP) for Agents — Michael Hablich, Google Michael Hablich, Product Manager for Chrome DevTools at Google, shares four engineering lessons from building Chrome DevTools for Age... 0 comments 628 views
44:53 Foundation Models3 months ago It’s starting… Matthew Berman breaks down Anthropic's newly published paper on recursive self-improvement, which traces the evolution of AI developm... 0 comments 21.3K views
16:30 Foundation Models3 months ago SWE-rebench: Lessons from Evaluating Coding Agents — Ibragim Badertdinov, Nebius Ibragim Badertdinov, an AI researcher at Nebius with an unconventional background—a trained dentist turned NeurIPS and ICML author—pr... 0 comments 881 views
23:25 Foundation Models3 months ago The Art & Science of Benchmarking Agents — Vincent Chen, Snorkel AI Vincent Chen, research fellow and co-founder at Snorkel AI, took the stage at AI Engineer to share meta-level lessons on what separat... 0 comments 402 views
15:59 Foundation Models3 months ago Nemotron 3 Ultra NVIDIA’s 550B Open Model NVIDIA has released Nemotron 3 Ultra, a 550 billion parameter mixture-of-experts model built specifically for agentic workloads, and... 0 comments 1.9K views
28:03 Foundation Models3 months ago Text Diffusion — Brendon Dillon, Google DeepMind Brendan Dillon, a research scientist at Google DeepMind, delivered a technically rigorous presentation at AI Engineer on text diffusi... 0 comments 462 views