12:27 Coding & Builds4 months ago Nemotron 3 Ultra – NVIDIA’s Most Powerful Open Model – Long Running Agents NVIDIA has released Nemotron Ultra, its largest open model to date at 550 billion total parameters—with only 55 billion active at inf... 0 comments 1.8K views
12:40 Deep Dives4 months ago What Lies Beneath the API — Benjamin Cowen, Modal Benjamin Cowen, a forward-deployed machine learning engineer at Modal, delivers a conference talk examining one of the most consequen... 0 comments 444 views
09:39 Coding & Builds4 months ago Step 3.7 Flash – 198B Open Source Model That Does Everything; Does it Really? Step 3.7 Flash is a 198 billion parameter sparse mixture-of-experts model from Step One, activating only 11 billion parameters per to... 0 comments 1.8K views
01:07:19 Deep Dives4 months ago Inference, Diffusion, World Models, and More | YC Paper Club Y Combinator's inaugural Paper Club brought together roughly 100 selected founders and researchers at Pioneer in Woodside, CA — drawi... 0 comments 8.7K views
10:10 Tutorials5 months ago Intern-S2-Preview FP8: 35B Scientific Multimodal Model Running Locally InternLM's latest release, Intern-S2-Preview, is a 35-billion-parameter scientific multimodal model that takes a different approach t... 0 comments 1.3K views
33:45 Coding & Builds5 months ago AI Dev 26 x SF | Eda Zhou & Mahdi Ghodsi: Building Personal AI Agents with Open Source Models At the AI Dev 26 conference in San Francisco, AMD engineers Eda Zhou and Mahdi Ghodsi lead a hands-on workshop teaching attendees how... 0 comments 651 views
15:56 Reviews & Comparisons5 months ago MiniCPM-V 4.6: The Agent Vision Model Sam Witteveen examines MiniCPM-V 4.6, a 1.3 billion parameter vision-language model released by OpenBMB—a joint initiative between AI... 0 comments 2.4K views
08:06 Reviews & Comparisons5 months ago MTP vs DFlash — Speculative Decoding Explained Simply This video by Fahd Mirza offers a clear, structured comparison of two speculative decoding techniques — Multi-Token Prediction (MTP)... 0 comments 1.2K views
08:43 Tutorials5 months ago DFlash Drafter for Gemma 4 26B – Official Speculative Decoding is Here: Run Locally ZLab, the UC San Diego research team that invented DFlash speculative decoding, has released the first official drafter model paired... 0 comments 607 views
16:34 Benchmarks6 months ago Is ERNIE Image Turbo Better Than FLUX? I Tested It Locally Fahd Mirza installs and tests Baidu's ERNIE Image Turbo locally, an open-weights text-to-image model built on a single-stream diffusi... 0 comments 0.9K views