13:43 Tutorials4 months ago Gemma4 12B in Quantization-Aware Training (QAT) with Ollama – Full Testing Google's Gemma 4 12B model now has Quantization-Aware Training (QAT) checkpoints, and Fahd Mirza puts them through a full workout in... 0 comments 1.7K views
12:13 Tutorials4 months ago Gemma 4 12B – Google’s Unified Multimodal Model Running Locally Fahd Mirza walks through a complete local installation and multi-modal evaluation of Gemma 4 12B, Google's newest open-weight model,... 0 comments 6.1K views
13:13 Tutorials4 months ago Adaptive PFlash + Hermes Agent – Self-Tuning Prefill on a Single GPU Locally Fahd Mirza demonstrates the newly shipped adaptive compression feature in PFlash, the prefill-acceleration component of the open-sour... 0 comments 2.2K views
08:32 Tutorials4 months ago WeasyPrint with Ollama Tutorial: HTML to PDF with AI Integration Fahd Mirza walks through a complete integration of WeasyPrint — a lightweight Python-based HTML-to-PDF rendering engine — with Ollama... 0 comments 1.8K views
10:07 Tutorials4 months ago Dolphin X1 Trinity Nano: The Model That Never Says No: Run and Test Locally Fahd Mirza walks through the local deployment and testing of Dolphin X1 Trinity Nano, the first model trained entirely within a custo... 0 comments 592 views
08:38 Tutorials4 months ago HRM-Text-1B: A 1B Model That Beats 7B Models for $1,500: Test Locally HRM-Text-1B is a 1 billion parameter pre-trained language model that claims to match or outperform models in the 2–7 billion paramete... 0 comments 2.1K views
08:53 Coding & Builds4 months ago LFM2.5-8B-A1B: Local Agentic AI with Multilingual Support Tested LFM2.5-8B-A1B is Liquid AI's latest open-weight model — an 8.3 billion parameter mixture-of-experts architecture that activates only... 0 comments 1.7K views
20:40 Tutorials4 months ago Stable Audio 3: Created Music From 20+ Countries Locally Fahd Mirza walks through a complete local installation and stress test of Stable Audio 3, Stability AI's latest open-weights audio ge... 0 comments 746 views
10:06 Coding & Builds5 months ago DFlash Leaves Qwen Territory – Gemma 4 31B Now Runs 5x Faster with Speculative Decoding Fahd Mirza demonstrates the first end-to-end deployment of Llama Box DFlash with Google's Gemma 4 31B model, following the merge of P... 0 comments 3.5K views
12:12 Coding & Builds5 months ago Microsoft Lens: Impressive on Paper, But Does It Deliver Locally? Let’s Test Fahd Mirza installs and tests Microsoft Lens, a new 3.8-billion-parameter text-to-image model quietly pushed to Hugging Face, running... 0 comments 894 views