09:07 Research & Benchmarks2 months ago Hy3 – Tencent Says its Best Chinese Model – Let’s Test Fahd Mirza puts Tencent's newly released Hunyuan 3 (Hi3) through a battery of practical tests in Tencent's AI Studio, covering physic... 0 comments 1.4K views
20:54 Research & Benchmarks2 months ago MiniCPM5 – The 1B Cognitive Core? Sam Witteveen reviews MiniCPM5, a 1-billion-parameter dense model from OpenBMB (Tsinghua NLP Lab), framing it against Andrej Karpathy... 0 comments 4.6K views
09:06 Research & Benchmarks2 months ago Leanstral 1.5 119B A6B: The Free AI That Proves Your Code Is Correct Leanstral 1.5 is a 119-billion-parameter open model from Mistral, available free via the Mistral API, built specifically to write for... 0 comments 1.9K views
14:03 Research & Benchmarks2 months ago Fable 5 is Back! Here’s the Best Way to Use It… With Fable 5 restored to public access on July 1st following government export control restrictions, The AI Advantage breaks down the... 0 comments 1K views
21:10 Research & Benchmarks2 months ago I Tested Gemini Spark: What Google’s AI Agent Can Actually Do in 21 Minutes Peter Yang offers a 21-minute live walkthrough of Gemini Spark, Google's personal AI agent embedded inside the Gemini interface and c... 0 comments 361 views
10:50 Research & Benchmarks2 months ago Laguna XS 2.1: Poolside’s Local Coding Agent Tested – Nine Languages Fahd Mirza puts Poolside's newly released Laguna XS 2.1 through a live evaluation using the Hermit agentic framework. The model is a... 0 comments 1.2K views
12:40 Research & Benchmarks2 months ago Sonnet 5 vs Ornith 35B: Can a Local Model Beat Closed-Source? Fahd Mirza runs a direct head-to-head evaluation between Claude Sonnet 5 and Ornith 1 35B — a new open-source mixture-of-experts codi... 0 comments 4.4K views
10:26 Research & Benchmarks2 months ago NotebookLM’s Brand New Feature Generates Shorts With One Click Google's NotebookLM has launched a new "short video overviews" feature that generates polished, educational short-form videos from up... 0 comments 4.7K views
28:52 Research & Benchmarks2 months ago GLM-5.2 Proves Open-Source AI is Finally Good Now! Matt Wolfe puts ZAI's GLM-5.2 through an extended hands-on evaluation, starting with a clear-eyed explanation of what 'open-weight' a... 0 comments 34.6K views
12:25 Research & Benchmarks2 months ago LongCat-2.0: China Breaks Free From Nvidia to Train a 1.6T Model Meituan — better known outside China as a food-delivery giant — has quietly released LongCat 2.0, a 1.6-trillion-parameter mixture-of... 0 comments 2.3K views
26:20 Research & Benchmarks2 months ago GLM-5.2 vs MiniMax-M3: Opus Has REAL COMPETITION (Model Stacking) IndyDevDan makes the case that the open-weight model landscape has fundamentally changed, positioning GLM-5.2 from Zhipu AI as a genu... 0 comments 3.5K views
15:41 Research & Benchmarks2 months ago New Agentic Coding Model Ornith 9B — Is It Worth Running Locally? Bart Slodyczka tests Ornith 1.0, a newly released open-source agentic coding model family from Deep Reinforce AI (San Francisco), foc... 0 comments 9.3K views
17:36 Research & Benchmarks2 months ago GLM 5.2 Is Free And Beats Claude On Most Work. So Why Can’t Companies Switch? Nate B. Jones of AI News & Strategy Daily delivers a candid, firsthand evaluation of GLM 5.2, an open-source model from Zhipu AI... 0 comments 5.5K views
16:35 Research & Benchmarks2 months ago Introducing Ornith 1.0 Sam Witteveen introduces Ornith 1.0, a new family of open-weight models from Deep Reinforce that takes a fundamentally different appr... 0 comments 9.2K views
09:41 Research & Benchmarks2 months ago OpenAI Just Gave Codex a Superpower & More AI News You Can Use The AI Advantage reviews OpenAI Codex's newly released record-and-replay feature — currently Mac-only and requiring at least a $20/mo... 0 comments 4.7K views
27:14 Research & Benchmarks2 months ago GLM 5.2 is SO GOOD (and almost free) This video from the How I AI channel takes a close look at GLM 5.2, the latest open-weight model from Beijing-based startup Z.AI. The... 0 comments 2.8K views