19:50 Business & Strategy2 weeks ago Taking Reinforcement Learning Cross Datacenter — Nan Jiang, Modal At an AI Engineer conference, Nan Jiang from Modal presents a technical architecture for running reinforcement learning post-training... 0 comments 248 views
08:48 Benchmarks1 month ago Catmind-1.2b: A Reasoning Model that Thinks in Cat Stories Fahd Mirza tests CatMind-1.2B, a fine-tune of the LFM 2.5 1.2B thinking model in which all reasoning traces have been replaced with i... 0 comments 863 views
14:04 Tutorials1 month ago I Cut the Internet and Let AI Read the File I Could Never Upload. It Caught the Leak. Local AI models have matured to the point where sensitive documents can be analyzed entirely offline, without ever touching cloud inf... 0 comments 14.8K views
46:52 Coding & Dev Tools1 month ago The Prime Intellect Stack — Will Brown, Prime Intellect Will Brown, head of applied research at Prime Intellect, delivers a technical workshop at the AI Engineer conference on the Prime Int... 0 comments 592 views
49:44 Interviews1 month ago The AI Memory Problem: Why Long Context Isn’t Enough — Dan Biderman, Engram Co-founder & CEO Dan Biderman, co-founder and CEO of Engram — the AI memory startup that raised a $98 million seed round — joins the Latent Space podc... 0 comments 726 views
13:58 Foundation Models2 months ago Your LLM Deception Monitor Is Broken. The Fix Is in the Training Data – Sachin Kumar, LexisNexis Sachin Kumar, senior data scientist at LexisNexis, presents a peer-reviewed paper accepted at IJCNN that introduces a new technique f... 0 comments 289 views
29:08 Tutorials2 months ago Fine-Tune the biggest open-source models (even with a bad PC) David Ondrej walks through a complete tutorial for fine-tuning Kimi K2.7, a 2.7-trillion-parameter open-source model that he position... 0 comments 11.5K views
22:35 Foundation Models2 months ago Continual Learning for AI Agents: From Failures to Durable Improvements – Soheil Feizi, RELAI Soheil Feizi, founder and CEO of RELAI and associate professor of computer science at the University of Maryland, presents a rigorous... 0 comments 1.1K views
10:25 Tutorials2 months ago Krea2 Has No Good Reference Mode. LoRA Is the Fix|From Dataset to Turbo Output This tutorial walks through the complete process of training a LoRA for Krea2, the image and video generation model praised for its c... 0 comments 1.1K views
24:07 Tutorials3 months ago Hermes Agent powered by local models on the DGX Spark is basically magic Alex Finn demonstrates a complete end-to-end setup of a Hermes Agent running entirely on a locally-hosted model — specifically Qwen 3... 0 comments 8.5K views