01:00:11 News & Opinion17 hours ago Self-Improving Harnesses, Local Personal AI And YC’s Agent For Work | YC Paper Club Y Combinator's Paper Club devotes a full session to agent harnesses — the scaffolding, context engineering, and orchestration layers... 0 comments 12.2K views
01:04:40 Interviews3 days ago How to Build Your Own Data Center & Why Every Startup Should Do It Speechify founder and CEO Cliff Whitesman joins Harry Stebbings on 20VC to explain why the text-to-speech startup has invested tens o... 0 comments 2K views
08:11 Benchmarks7 days ago JetSpec Locally: Breaking the Speed Ceiling of LLM Inference – Up to 9x Fahd Mirza installs JetSpec — a speculative decoding framework from UCSD — and benchmarks it locally on an H100 GPU running Qwen 3 8B... 0 comments 7.4K views
08:25 Tutorials2 weeks ago Qwen3.8 Just Landed in SIE — Running It Locally On Day One Fahd Mirza walks through a day-one setup of Alibaba's Qwen 3.8 27B model running locally via SIE — the Superlinked Inference Engine —... 0 comments 5K views
21:48 Deep Dives2 weeks ago KV Cache-Aware Routing and P/D Disaggregation on Kubernetes — Yuchen Fama & Ashish Kamra, Red Hat At the AI Engineer conference, Red Hat's Yuchen Fama (vLLM contributor and product manager) and Ashish Kamra (senior manager of perfo... 0 comments 3.4K views
30:00 News & Opinion2 weeks ago Can LLMs Write Fast Multi-GPU Kernels? — Simran Arora, Together AI Simran Arora, principal scientist at Together AI and incoming Caltech professor, presents original research asking a pointed question... 0 comments 174 views
08:39 News & Opinion4 weeks ago Why You Shouldn’t Run Qwen3.8 2.4T Locally Fahd Mirza delivers a frank assessment of Qwen's newly released Qwen3 2.4-trillion-parameter open-weight model, making the case that... 0 comments 1.3K views
02:29:32 Interviews4 weeks ago Sergey Brin Retakes Gemini, 4 Labs Lose Containment, Compute Trades at NYSE w/ Kush Bavaria | EP 278 Episode 278 of Peter Diamandis's Moonshots podcast gathers regulars AWG, DB2, and SEM Ismael alongside special guest Kush Bavaria — C... 0 comments 58.9K views
05:15 Interviews4 weeks ago RUM Group’s $3 Billion AI Opportunity Rumble's parent company has rebranded as Rum Group and reported a record Q2 2026 with revenue of $40.4 million, up 61% year-over-year... 0 comments 450 views
09:58 Tutorials4 weeks ago MiniMax H3 Locally with ComfyUI and ClipProj on 1 GPU Fahd Mirza demonstrates running MiniMax H3, a large multimodal video generation model, completely locally using ComfyUI on a single H... 0 comments 2K views