Grok 4.5 Just Shocked The AI Community – SpaceXAI Grok 4.5

Grok 4.5 Just Shocked The AI Community – SpaceXAI Grok 4.5

More

Summary

TheAIGRID takes a comprehensive look at Grok 4.5, the latest model from xAI, examining where it fits in the current frontier model landscape and whether its pricing makes it a compelling alternative to significantly more expensive options. The video leads with benchmark results across several evaluations: Terminal Bench 2.1 (where Grok 4.5 scores 83.3%, matching GPT-5.5), SWE-bench software engineering tasks (outperforming Opus 4.8 Max), and the GDP-Val real-world professional work benchmark (29% vs GPT-5.5’s 22% and Opus 4.8’s 21%).

Pricing is a central theme throughout. Grok 4.5 comes in at $2 per million input tokens and $6 per million output tokens — less than half the cost of comparable frontier models, and a fraction of Claude Fable’s $10/$50 pricing. The video also highlights Cursor’s own benchmarks showing Grok build outperforming Claude Code with Opus 4.8 on coding tasks, a result the host attributes in part to xAI’s acquisition of Cursor and access to its training data.

The creator is careful to note benchmark reliability concerns — including OpenAI’s own finding that SWE-Pro has a 30% error rate — and positions Grok 4.5 as sitting between Opus 4.7 and Opus 4.8 in terms of intelligence, rather than accepting xAI’s own Opus 4.8 comparison at face value. For developers evaluating cost-performance tradeoffs in 2026, this is a useful orientation video on where Grok 4.5 actually fits.


📺 Source: TheAIGRID · Published July 09, 2026
🏷️ Format: Review

1 Item

Channels

1 Item

Companies