08:05 Tutorials4 months ago Dots.TTS SOAR – State of the Art Speaker Similarity, Runs Fully Local Fahd Mirza walks through the installation and hands-on testing of Dots.TTS SOAR, a fully open-source zero-shot voice cloning model th... 0 comments 1.8K views
09:07 Coding & Builds4 months ago Luce KVFlash: Finding a Needle in 256K Tokens with Low VRAM Fahd Mirza runs a controlled needle-in-a-haystack experiment to test whether Luce KVFlash can reliably retrieve a specific fact from... 0 comments 1.7K views
12:46 Coding & Builds4 months ago VibeThinker-3B: 3B Model That Challenges Claude Opus? Test Locally Fahd Mirza installs and tests VibeThinker-3B — a reasoning model released by Weibo, the Chinese social media giant — directly on an N... 0 comments 2.8K views
10:03 Coding & Builds4 months ago GLM-5.2: Anthropic Got Banned. China Shipped Same Night: Hands-on Testing Fahd Mirza delivers a hands-on test of GLM 5.2, the latest model from Chinese lab ZAI (Zhipu AI), released the same evening Anthropic... 0 comments 2.9K views
13:19 Coding & Builds4 months ago Kimi K2.7 Code + Hermes Agent – Clinically Certified to Be Insane Fahd Mirza puts Kimi K2.7 Code — Moonshot AI's latest open-weight coding model — through a demanding real-world test using the Hermes... 0 comments 3.9K views
12:04 Tutorials4 months ago NVIDIA Ships Nemotron 3.5 ASR Streaming 0.6b: Run Locally on CPU Fahd Mirza provides a hands-on walkthrough of NVIDIA's newly released NeMo-Tron 3.5 ASR — a 600-million-parameter streaming speech re... 0 comments 2.3K views
09:42 Tutorials4 months ago Luce Spark: Run a 35B Model Under 16GB VRAM Locally Fahd Mirza demonstrates LuceSpark, a memory management technique that allows a 35-billion-parameter mixture-of-experts model to run w... 0 comments 3.6K views
14:01 Tutorials4 months ago DiffusionGemma GGUF: Run Google’s Fastest Model Locally on Any GPU Fahd Mirza demonstrates how to run DiffusionGemma — Google's new diffusion-based text generation model — locally using a quantized GG... 0 comments 4.7K views
09:03 Tutorials4 months ago Gemma 4 Was Broken for Agents – Google Just Fixed It Google's Gemma 4 12B model contained a subtle but impactful bug in its official Jinja chat template that was silently breaking multi-... 0 comments 3.8K views
11:01 Tutorials4 months ago Higgs Audio v3 TTS: This Model Does Not Read, It Talks in Your Language Fahd Mirza demonstrates Higgs Audio V3, a multilingual text-to-speech model from Boson AI, running entirely locally on an Nvidia RTX... 0 comments 1.1K views