09:36 Benchmarks4 months ago Qwen3.6 (REAP 90pct GGUF): The Brain-Damaged Model Fahd Mirza takes a deep look at an aggressively pruned variant of Qwen 3.6 — a 35-billion-parameter mixture-of-experts model — compre... 0 comments 2.8K views
14:42 Tutorials4 months ago Qwen3.6 27B (Pi-Reasoning GGUF) – Fine-Tuned for Local Heavy AI Agent Fahd Mirza tests Pi-Reasoning, a community fine-tune of Qwen 3.6 27B built specifically for agentic coding — tasks like reading files... 0 comments 3.9K views
08:12 Reviews & Comparisons4 months ago GLM 5.2 – Why Everyone is Loving It? And How to Run It Locally Fahd Mirza covers GLM-5.2, the 744 billion parameter mixture-of-experts model from Chinese AI lab Zhipu AI that has become one of the... 0 comments 479 views
08:46 News & Opinion4 months ago Weekly AI Recap – Claude Fable 5, DiffusionGemma, US Export Controls | June 2026 Fahd Mirza delivers a fast-paced weekly AI recap covering the most consequential developments from the first two weeks of June 2026.... 0 comments 607 views
16:47 Reviews & Comparisons4 months ago Google QAT vs Unsloth QAT + MTP – Which Gemma 4 12B Is Actually Better? This video pits two quantized versions of Google's Gemma 4 12B against each other in a practical, locally-run benchmark: Google's own... 0 comments 3.2K views
09:03 Tutorials4 months ago Gemma 4 Was Broken for Agents – Google Just Fixed It Google's Gemma 4 12B model contained a subtle but impactful bug in its official Jinja chat template that was silently breaking multi-... 0 comments 3.8K views
13:08 Tutorials4 months ago Gemma 4 12B QAT + MTP on llama.cpp Locally – Twice the Speed, Same Quality? This video by Fahd Mirza walks through running Google's newly released Gemma 4 12B QAT (Quantization-Aware Training) model alongside... 0 comments 2.9K views
06:01 Coding & Builds4 months ago Run Google’s newest 12B AI on a phone? Yes, it’s possible! The Alphastack channel walks through a custom cross-platform app that runs Google's Gemma 4 12B multimodal model entirely on-device —... 0 comments 125 views
09:07 Tutorials4 months ago DwarfStar: Run DeepSeek V4 Locally with DS4 at 34 tok/s Fahd Mirza covers DwarfStar, a brand-new inference engine built specifically for DeepSeek V4 Flash (DS4) by the creator of Radius. Un... 0 comments 2.7K views
32:57 Tutorials4 months ago Unsloth Studio is insane… fine-tune any AI model locally Unsloth Studio is a free, open-source desktop application that brings LLM fine-tuning to consumer hardware — and this video by David... 0 comments 8.3K views