18:35 Benchmarks24 hours ago I Gave Two AI Supercomputers a Real Job Alex Ziskind and guest Wendell put two high-end local AI machines to work: a system with eight RTX Pro 6000 GPUs and an Nvidia DGX St... 0 comments 3.5K views
20:38 Benchmarks7 days ago M5 Ultra vs 2 DGX Sparks… The Number You’re Not Looking At Alex Ziskind pits a 256 GB M5 Ultra Mac Studio against a cluster of two NVIDIA DGX Sparks, each with 128 GB, linked by a 200 Gb QSFP... 0 comments 11.6K views
09:03 Benchmarks7 days ago I Tested Codex’s $500/mo Ultrafast. What You Need to Know. Nate Herk puts Codex's new Ultrafast mode through real-world tests to see whether it justifies the $500-per-month plan it requires. U... 0 comments 526 views
12:44 Benchmarks1 week ago GPT-6-Astra on The New ULTRAFAST $500 Plan Is Scary… OpenAI has launched a new $500-a-month ChatGPT Pro plan, roughly $6,000 a year, and Nick Saraev gives an overview and tests it on a r... 0 comments 11.9K views
44:03 Benchmarks2 weeks ago Decision Models, DSV4.1 Flash & New Benchmarks Developer and YouTuber sentdex puts two new fast models head to head: GLM 5.3 Flash and DeepSeek V4.1 Flash. The video walks through... 0 comments 3.4K views
21:53 Benchmarks2 weeks ago i got one….and it’s FAST!!! NetworkChuck gets hands-on with Apple's M5 Ultra, on loan from Apple, and puts it head to head with his M3 Ultra to see how much fast... 0 comments 41.2K views
14:52 Benchmarks2 weeks ago M5 Ultra… Apple Wasn’t Messing Around Alex Ziskind puts Apple's new M5 Ultra chip through a rigorous set of benchmarks against the previous-generation M3 Ultra, focusing h... 0 comments 6.3K views
16:57 Benchmarks2 weeks ago Did Elon catch up? (Grok 4.7 is here) Matthew Berman breaks down xAI's newly released Grok 4.7, testing Elon Musk's own prediction that the model would land roughly on par... 0 comments 33.6K views
26:34 Benchmarks3 weeks ago I Gave 1000 AI Agents One Computer Alex Ziskind puts the Camino Grando, a liquid-cooled workstation packing eight Nvidia RTX Pro 6000 Blackwell GPUs and 768GB of total... 0 comments 10.4K views
39:23 Benchmarks3 weeks ago Agentic Engineering Benchmarks: How I RANK Astra, Fable 5.1, and Open-Weights IndyDevDan presents a practical framework for selecting AI benchmarks tailored to agentic engineering, arguing that broad composite i... 0 comments 4.5K views
19:21 Benchmarks4 weeks ago I Ran A 27B Model On A Hand-Sized PC… Didn’t Expect This Alex Ziskind tests the Kadas Mind Pro — a palm-sized PC powered by Intel's Panther Lake architecture with an ARC B390 integrated GPU... 0 comments 119.6K views
14:30 Benchmarks4 weeks ago MiniCPM5-2B: The Best Sub-Agent Model Yet? Sam Witteveen examines MiniCPM 2B (technically 2.5B), the latest small model from OpenBMB, which claims to beat 4B-class models on fu... 0 comments 4.2K views
22:26 Benchmarks4 weeks ago How long can your skills be before your agent forgets what you told it? — Laurie Voss, Arize AI Laurie Voss, head of developer relations at Arize AI and co-founder of npm, presents original benchmark research at AI Engineer tackl... 0 comments 350 views
12:20 Benchmarks1 month ago K2 Horizon: 0.9B, 7B, and 32B Tested Locally, Real Results The Institute of Foundation Models — the AI research arm of MBZUAI, a university in Abu Dhabi with satellite labs in Paris and Silico... 0 comments 1.2K views
32:20 Benchmarks1 month ago GPT-6 Astra blew away every one of my benchmarks The How I AI channel documents hands-on benchmark testing and practical workflow demonstrations using GPT-6 Astra, OpenAI's new front... 0 comments 30.6K views
08:11 Benchmarks1 month ago JetSpec Locally: Breaking the Speed Ceiling of LLM Inference – Up to 9x Fahd Mirza installs JetSpec — a speculative decoding framework from UCSD — and benchmarks it locally on an H100 GPU running Qwen 3 8B... 0 comments 7.4K views