10:30 Benchmarks4 days ago Ornith-1.5-9B: Great on Paper, Struggled in My Tests Locally Fahd Mirza tests the Ornith 1.5 9B model — the smallest and densest member of the Ornith 1.5 family — on a single NVIDIA A100 with 80... 0 comments 1.5K views
09:20 Benchmarks5 days ago DFlash 2: Qwen3.8-27B at 2× Speed – Live Benchmark Locally Fahd Mirza benchmarks DFlash 2, a community-developed speculative decoding enhancement for the Qwen3.8-27B model, running live on an... 0 comments 6.5K views
11:35 Coding & Dev Tools5 days ago Ornith 1.5 35B-A3B: Self-Improving Model Goes General-Purpose: Run Locally Fahd Mirza walks through deploying Ornith 1.5 35B-A3B — the latest release from Deep Reinforce — on a single Nvidia A100 80GB GPU usi... 0 comments 4.6K views
18:00 Research & Benchmarks6 days ago Qwen3.8-27B & How to Serve it Fast Sam Witteveen takes a close look at the newly released Qwen 3.8 27B model from Alibaba's Qwen team, positioning it as the natural suc... 0 comments 15.2K views
09:42 Coding & Dev Tools7 days ago Qwen3.8-9B: Community Distillation of a Frontier Model: Run Locally Fahd Mirza walks through installing and testing Qwen3.8-9B, a community-built 9-billion-parameter distillation of Alibaba's 27-billio... 0 comments 4.1K views
06:18 Coding & Dev Tools7 days ago Antares-1B: Vulnerability Localization with AI Locally Fahd Mirza demonstrates Antares 1B, a 1-billion parameter model released by Cisco's Foundation AI team specifically for vulnerability... 0 comments 883 views
13:13 Coding & Dev Tools1 week ago Qwen3.8-27B Locally: Does It Live Up to the Hype? Fahd Mirza pulls the Qwen3.8-27B weights and runs the model locally on a single Nvidia A100 80GB GPU using vLLM, providing a practica... 0 comments 1.4K views
08:39 Business & Strategy2 weeks ago Why You Shouldn’t Run Qwen3.8 2.4T Locally Fahd Mirza delivers a frank assessment of Qwen's newly released Qwen3 2.4-trillion-parameter open-weight model, making the case that... 0 comments 1.3K views
10:42 Tutorials2 weeks ago NeMo Switchyard: NVIDIA’s Model Router, Tested Locally Fahd Mirza installs and demonstrates NeMo Switchyard, a newly released NVIDIA tool written in Rust that acts as an intelligent model... 0 comments 1.1K views
12:51 Coding & Dev Tools2 weeks ago Nemotron 3.5 Lightning: Specialized Local AI for Long-Running Agents Fahd Mirza walks through a complete local installation and live test of NVIDIA's newly released Nemotron 3.5 Lightning — a 30-billion... 0 comments 1.6K views