01:24:35 Interviews3 months ago ARC-AGI-3 Explained by the Team That’s Winning It Machine Learning Street Talk convenes several members of a top-performing ARC-AGI-3 competition team — Jon Kotar, Stephano, D. Smith,... 0 comments 1.9K views
16:35 Reviews & Comparisons3 months ago Introducing Ornith 1.0 Sam Witteveen introduces Ornith 1.0, a new family of open-weight models from Deep Reinforce that takes a fundamentally different appr... 0 comments 9.3K views
08:51 Tutorials3 months ago OpenJarvis + Ollama: Local AI Agent That Tracks Every Watt Fahd Mirza walks through the installation and hands-on testing of Open Jarvis, a newly released local-first personal AI framework dev... 0 comments 2.2K views
12:18 Tutorials4 months ago Unlimited OCR from Baidu: One-shot Long-horizon Parsing: Run Locally Baidu's Unlimited OCR introduces a fundamentally different approach to document parsing: processing entire multi-page documents in a... 0 comments 1.9K views
09:36 Benchmarks4 months ago Qwen3.6 (REAP 90pct GGUF): The Brain-Damaged Model Fahd Mirza takes a deep look at an aggressively pruned variant of Qwen 3.6 — a 35-billion-parameter mixture-of-experts model — compre... 0 comments 2.8K views
18:17 Benchmarks4 months ago VibeThinker 3B – Taking on Giant Models Sam Witteveen digs into VibeThinker 3B, a small language model from Waybo AI Lab — the AI research arm of the Chinese social network... 0 comments 4.1K views
09:42 Tutorials4 months ago Luce Spark: Run a 35B Model Under 16GB VRAM Locally Fahd Mirza demonstrates LuceSpark, a memory management technique that allows a 35-billion-parameter mixture-of-experts model to run w... 0 comments 3.6K views
09:03 Tutorials4 months ago Gemma 4 Was Broken for Agents – Google Just Fixed It Google's Gemma 4 12B model contained a subtle but impactful bug in its official Jinja chat template that was silently breaking multi-... 0 comments 3.8K views
13:08 Tutorials4 months ago Gemma 4 12B QAT + MTP on llama.cpp Locally – Twice the Speed, Same Quality? This video by Fahd Mirza walks through running Google's newly released Gemma 4 12B QAT (Quantization-Aware Training) model alongside... 0 comments 2.9K views
06:49 Deep Dives4 months ago Nanowhale-100m: Fascinating Implemention of DeepSeek-V4 Architecture Fahd Mirza walks through Nanowhale-100M, a 110 million parameter language model built entirely from scratch—no borrowed weights—that... 0 comments 1.2K views