46:01 Deep Dives2 months ago Compression at the Edge — NVIDIA, Unsloth, HuggingFace, Ollama A conference panel titled "Compression at the Edge" brings together engineers from four of the most active organizations in local AI:... 0 comments 1K views
28:10 News & Opinion2 months ago New Deepseek, Seedance 2.5, Minimax H3, Gemini Robotics, AMD models: AI NEWS AI Search delivers a dense weekly roundup covering an unusually active period across multiple AI verticals. The biggest model news: B... 0 comments 74.6K views
25:28 Reviews & Comparisons3 months ago AMD Ryzen AI Halo – 100% Local AI Sam Witteveen reviews the AMD Ryzen AI Halo, a new workstation built around the Ryzen AI Max Plus 395 chip with 128 GB of unified LPD... 0 comments 8.7K views
35:15 News & Opinion4 months ago New robot waifus, GLM 5.2 craze, AI spas, new world models, new science agents: AI NEWS AI Search's weekly roundup covers a dense slate of releases spanning open-source models, video generation, robotics, and AI agents. T... 0 comments 52.3K views
14:01 Tutorials4 months ago DiffusionGemma GGUF: Run Google’s Fastest Model Locally on Any GPU Fahd Mirza demonstrates how to run DiffusionGemma — Google's new diffusion-based text generation model — locally using a quantized GG... 0 comments 4.7K views
16:47 Reviews & Comparisons4 months ago Google QAT vs Unsloth QAT + MTP – Which Gemma 4 12B Is Actually Better? This video pits two quantized versions of Google's Gemma 4 12B against each other in a practical, locally-run benchmark: Google's own... 0 comments 3.2K views
15:50 Deep Dives4 months ago Road to 5 Million Tokens: Breaking Barriers in Long Context Training — Max Ryabinin, Together AI Max Ryabinin, VP of Research and Development at Together AI, presents the company's research project on extending transformer trainin... 0 comments 769 views
14:35 Benchmarks4 months ago Google QAT vs Unsloth Q4_0 – Which Gemma 4 12B Quantization Is Better? Fahd Mirza runs a controlled comparison between two 4-bit quantized versions of Google's Gemma 4 12B model: Google's own QAT (quantiz... 0 comments 3.3K views
06:01 Coding & Builds4 months ago Run Google’s newest 12B AI on a phone? Yes, it’s possible! The Alphastack channel walks through a custom cross-platform app that runs Google's Gemma 4 12B multimodal model entirely on-device —... 0 comments 125 views
32:57 Tutorials4 months ago Unsloth Studio is insane… fine-tune any AI model locally Unsloth Studio is a free, open-source desktop application that brings LLM fine-tuning to consumer hardware — and this video by David... 0 comments 8.3K views