12:46 Coding & Dev Tools2 months ago VibeThinker-3B: 3B Model That Challenges Claude Opus? Test Locally Fahd Mirza installs and tests VibeThinker-3B — a reasoning model released by Weibo, the Chinese social media giant — directly on an N... 0 comments 2.7K views
10:03 Coding & Dev Tools2 months ago GLM-5.2: Anthropic Got Banned. China Shipped Same Night: Hands-on Testing Fahd Mirza delivers a hands-on test of GLM 5.2, the latest model from Chinese lab ZAI (Zhipu AI), released the same evening Anthropic... 0 comments 2.9K views
13:19 Coding & Dev Tools2 months ago Kimi K2.7 Code + Hermes Agent – Clinically Certified to Be Insane Fahd Mirza puts Kimi K2.7 Code — Moonshot AI's latest open-weight coding model — through a demanding real-world test using the Hermes... 0 comments 3.9K views
09:42 Tutorials2 months ago Luce Spark: Run a 35B Model Under 16GB VRAM Locally Fahd Mirza demonstrates LuceSpark, a memory management technique that allows a 35-billion-parameter mixture-of-experts model to run w... 0 comments 3.6K views
12:04 Tutorials2 months ago NVIDIA Ships Nemotron 3.5 ASR Streaming 0.6b: Run Locally on CPU Fahd Mirza provides a hands-on walkthrough of NVIDIA's newly released NeMo-Tron 3.5 ASR — a 600-million-parameter streaming speech re... 0 comments 2.2K views
14:01 Tutorials2 months ago DiffusionGemma GGUF: Run Google’s Fastest Model Locally on Any GPU Fahd Mirza demonstrates how to run DiffusionGemma — Google's new diffusion-based text generation model — locally using a quantized GG... 0 comments 4.6K views
09:03 Tutorials3 months ago Gemma 4 Was Broken for Agents – Google Just Fixed It Google's Gemma 4 12B model contained a subtle but impactful bug in its official Jinja chat template that was silently breaking multi-... 0 comments 3.8K views
11:01 Tutorials3 months ago Higgs Audio v3 TTS: This Model Does Not Read, It Talks in Your Language Fahd Mirza demonstrates Higgs Audio V3, a multilingual text-to-speech model from Boson AI, running entirely locally on an Nvidia RTX... 0 comments 1.1K views
08:21 Coding & Dev Tools3 months ago SimpleMem + Ollama: Local AI Memory That Actually Gets Smarter SimpleMem is an open-source AI memory framework that challenges the conventional approach taken by tools like Mem Zero and MemoryBear... 0 comments 1.2K views
08:56 Coding & Dev Tools3 months ago BLS-Mini-Code-1.0: Testing Cohere’s Secret Coding Model Locally Fahd Mirza walks through a same-day local installation and test of BLS-Mini-Code-1.0, Cohere's first dedicated coding model released... 0 comments 1.5K views