23:35 Foundation Models2 weeks ago Stop Evaluating Models Like It’s the 50s – Alejandro Vidal, Mindmakers Alejandro Vidal, founder of Mind Makers, argues that the AI industry's standard approach to LLM benchmarking—summing correct answers... 0 comments 287 views
23:00 Foundation Models3 weeks ago Stop Evaluating Models Like It’s the 50s – Alejandro Vidal, Mindmakers Alejandro Vidal of Mindmakers presents a conference talk arguing that LLM benchmarks are fundamentally flawed because they treat ever... 0 comments 880 views
01:58:57 Interviews3 months ago 🔴LIVE: ChatGPT 5.5 is here. Does it beat Claude Opus 4.7? Alex Finn hosts a live stream testing ChatGPT 5.5 immediately after its release, pitting it against Claude Opus 4.7 across multiple t... 0 comments 13.3K views
01:51:43 Interviews3 months ago LIVE: Opus 4.7 is incredible, new Codex automated my life, Claude Design is MWAH Alex Finn runs an extended live stream centered on evaluating Claude Opus 4.7, opening with a pointed defense of Anthropic's models a... 0 comments 29 views
13:12 Tutorials8 months ago Anthropic Just Added These Features to Claude Code Ray Amjad walks through a batch of new Claude Code updates released in late 2025, covering changes that affect how developers work wi... 0 comments 19.2K views
32:13 Tutorials8 months ago Claude Opus 4.5: The Engineers Model IndyDevDan benchmarks Claude Opus 4.5 in a live multi-agent engineering session, arguing that the model's most underrated capability... 0 comments 27.7K views
33:50 Tutorials8 months ago Be a 10x Vibe Coder (Claude Code + Cursor + MCP) Developer and app builder Chris joins Greg Isenberg's podcast to share his complete AI coding workflow, focused on getting maximum ou... 0 comments 148.7K views