10:04 Research & Benchmarks2 months ago Mistral OCR 4 Is Built Different – 170 Languages, and Does It Beats Them All? Fahd Mirza, a Mistral AI ambassador, delivers a no-hype hands-on review of Mistral OCR 4, the company's new document extraction model... 0 comments 1.6K views
12:16 Research & Benchmarks2 months ago I Battle Tested Sakana Fugu’s Fable Killer Sakana AI, a Japanese company, has launched Fugu Ultra — a multi-agent orchestration system that routes tasks through multiple fronti... 0 comments 58.7K views
21:01 Research & Benchmarks2 months ago GLM-5.2 vs MiniMax-M3 vs Qwen3.7-Max — 3 Coding Tests, One Winner Fahd Mirza runs a hands-on three-way coding showdown between GLM-5.2, MiniMax M3, and Qwen 3.7 Max using the Hermes agent framework.... 0 comments 2.2K views
12:45 Research & Benchmarks2 months ago Master Boogu-Image: The New 10B Open Source King of AI Design|Pro Commercial Design Veteran AI delivers a structured evaluation of Boogu-Image, a newly released open-source image generation and editing model with appr... 0 comments 752 views
08:12 Research & Benchmarks2 months ago GLM 5.2 – Why Everyone is Loving It? And How to Run It Locally Fahd Mirza covers GLM-5.2, the 744 billion parameter mixture-of-experts model from Chinese AI lab Zhipu AI that has become one of the... 0 comments 423 views
09:49 Research & Benchmarks2 months ago GLM 5.2 Failed… But Not At Everything Creator Magic puts GLM 5.2 — a recently released open-weight Chinese model that benchmarks close to Claude Opus on several metrics —... 0 comments 2.3K views
19:57 Research & Benchmarks2 months ago GLM-5.2 vs Claude Opus 4.8: Two AI Models, the Same Brutal Tests Fahd Mirza runs a structured three-round head-to-head between GLM 5.2 — THUDM's fully open-weight 744B parameter MoE model (40B activ... 0 comments 2.3K views
14:34 Research & Benchmarks2 months ago GLM-5.2 is Basically Opus (For 1/5 the Price) Nick Saraev puts GLM 5.2 head-to-head against Anthropic's Opus 4.8 across roughly 40 real-world creative and coding tasks — including... 0 comments 7.3K views
08:30 Research & Benchmarks2 months ago New 225B Coding Model Laguna M.1 – Honest Test (Bugs + Creative Code) Fahd Mirza takes Laguna M.1 — Poolside's new 225-billion-parameter mixture-of-experts coding model with 23 billion parameters active... 0 comments 1.1K views
29:57 Research & Benchmarks2 months ago New #1 open-source AI model is here! AI Search puts GLM 5.2 — the latest release from ZAI — through a series of demanding real-world tasks, claiming it currently tops ope... 0 comments 162.2K views
31:47 Research & Benchmarks2 months ago PewDiePie Wants To Take Down The Big AI Companies Matt Wolfe reviews Project Odysseus, an open-source self-hosted AI workspace built by PewDiePie (Felix Kjellberg) that aims to recrea... 0 comments 5.1K views
08:29 Research & Benchmarks2 months ago Best AI Visual Effects Tools in 2026 (Full Ranking) Youri van Hofwegen puts six AI visual effects tools through a standardized three-part test — background replacement, relighting, and... 0 comments 6.4K views
12:24 Research & Benchmarks2 months ago Why I’ll Pay Full Price for Fable 5 When It Comes Back Chris Raroque tests Fable 5, Anthropic's newest coding model that was available for roughly 48 hours before being pulled by the US go... 0 comments 3.5K views
13:22 Research & Benchmarks2 months ago GLM 5.2 – The Top NEW Open Weights Model Sam Witteveen delivers a timely hands-on review of GLM 5.2, the latest open-weights model from Z.A. AI (ZhipuAI), released just hours... 0 comments 5.6K views
12:49 Research & Benchmarks2 months ago Best AI Video Generator on Your Phone (2026) Youri van Hofwegen runs eight leading AI video generators through a structured mobile usability test in 2026, evaluating the experien... 0 comments 7.7K views
14:00 Research & Benchmarks2 months ago Kimi K2.7 vs GLM-5.2: Real Coding Showdown in Hermes Agent Fahd Mirza runs a live, head-to-head coding comparison between two of China's most capable open-source models — Kimi K2.7 Code from M... 0 comments 2.7K views