05:15 Business & Strategy3 months ago Google just casually disrupted the open-source AI narrative… Fireship's Code Report covers Google's release of Gemma 4, a large language model launched under the Apache 2.0 license — making it o... 0 comments 73.1K views
25:59 Foundation Models4 months ago They solved AI’s memory problem! The AI Search channel breaks down a new research paper from the Kimi team — a Chinese AI lab frequently compared to DeepSeek — titled... 0 comments 51.1K views
01:46:33 Interviews4 months ago Success without Dignity? Nathan finds Hope Amidst Chaos, from The Intelligence Horizon Podcast Nathan Labenz of the Cognitive Revolution podcast appears as a guest on the Intelligence Horizon podcast, hosted by Yale seniors Owen... 0 comments 407 views
22:52 Foundation Models4 months ago Google’s TurboQuant Crashed the AI Chip Market Google has released TurboQuant, a new KV cache compression algorithm that delivers a 6x reduction in KV cache memory usage and an 8x... 0 comments 42.8K views
03:02:16 Interviews4 months ago Aravind Srinivas: Perplexity CEO on Future of AI, Search Aravind Srinivas, CEO and co-founder of Perplexity AI, joins Lex Fridman to explain in technical depth how Perplexity works and where... 0 comments 0.9M views
No Image Available Interviews4 months ago Four CEOs on the Future of AI: CoreWeave, Perplexity, Mistral, and IREN Recorded live at Nvidia's annual GTC conference, this All-In Podcast session brings together four AI company CEOs for back-to-back in... 0 comments 62 views
02:30:45 Interviews4 months ago Dylan Patel — The Single Biggest Bottleneck to Scaling AI Compute Dylan Patel, CEO of semiconductor research firm SemiAnalysis, joins Dwarkesh Patel to deliver one of the most data-dense analyses ava... 0 comments 184.6K views
09:13 Research & Benchmarks4 months ago NVIDIA Launches Nemotron 3 Super: 120B LatentMoE Explained & Tested Fahd Mirza covers the launch of NVIDIA's Nemotron 3 Super, a 120-billion-parameter language model built on a novel architecture calle... 0 comments 3.5K views
01:26:00 Interviews5 months ago Agent Inference at the “Speed of Light” — How NVIDIA moves like a $4.3 Trillion Startup This Latent Space podcast episode features Netter and Kyle from NVIDIA—engineering leaders on the Dynamo inference system—in a techni... 0 comments 5.4K views
02:25:43 Interviews5 months ago Approaching the AI Event Horizon? Part 2, w/ Abhi Mahajan, Helen Toner, Jeremie Harris, @8teAPi The second half of a Cognitive Revolution live event brings together three guests to examine AI from the intersecting perspectives of... 0 comments 866 views