12:46 Coding & Dev Tools2 months ago VibeThinker-3B: 3B Model That Challenges Claude Opus? Test Locally Fahd Mirza installs and tests VibeThinker-3B — a reasoning model released by Weibo, the Chinese social media giant — directly on an N... 0 comments 2.7K views
11:01 Tutorials3 months ago Higgs Audio v3 TTS: This Model Does Not Read, It Talks in Your Language Fahd Mirza demonstrates Higgs Audio V3, a multilingual text-to-speech model from Boson AI, running entirely locally on an Nvidia RTX... 0 comments 1.1K views
12:27 Coding & Dev Tools3 months ago Nemotron 3 Ultra – NVIDIA’s Most Powerful Open Model – Long Running Agents NVIDIA has released Nemotron Ultra, its largest open model to date at 550 billion total parameters—with only 55 billion active at inf... 0 comments 1.7K views
12:40 Foundation Models3 months ago What Lies Beneath the API — Benjamin Cowen, Modal Benjamin Cowen, a forward-deployed machine learning engineer at Modal, delivers a conference talk examining one of the most consequen... 0 comments 373 views
09:39 Coding & Dev Tools3 months ago Step 3.7 Flash – 198B Open Source Model That Does Everything; Does it Really? Step 3.7 Flash is a 198 billion parameter sparse mixture-of-experts model from Step One, activating only 11 billion parameters per to... 0 comments 1.8K views
01:07:19 Foundation Models3 months ago Inference, Diffusion, World Models, and More | YC Paper Club Y Combinator's inaugural Paper Club brought together roughly 100 selected founders and researchers at Pioneer in Woodside, CA — drawi... 0 comments 8.7K views
10:10 Tutorials3 months ago Intern-S2-Preview FP8: 35B Scientific Multimodal Model Running Locally InternLM's latest release, Intern-S2-Preview, is a 35-billion-parameter scientific multimodal model that takes a different approach t... 0 comments 1.2K views
33:45 Coding & Dev Tools3 months ago AI Dev 26 x SF | Eda Zhou & Mahdi Ghodsi: Building Personal AI Agents with Open Source Models At the AI Dev 26 conference in San Francisco, AMD engineers Eda Zhou and Mahdi Ghodsi lead a hands-on workshop teaching attendees how... 0 comments 603 views
15:56 Research & Benchmarks3 months ago MiniCPM-V 4.6: The Agent Vision Model Sam Witteveen examines MiniCPM-V 4.6, a 1.3 billion parameter vision-language model released by OpenBMB—a joint initiative between AI... 0 comments 2.4K views
08:06 Research & Benchmarks3 months ago MTP vs DFlash — Speculative Decoding Explained Simply This video by Fahd Mirza offers a clear, structured comparison of two speculative decoding techniques — Multi-Token Prediction (MTP)... 0 comments 1.1K views