Grok 4.5 ranks second on the APEX-SWE leaderboard with 51.2% accuracy, trailing Anthropic's Fable 5 at 65.5% in real-world software engineering tests.

xAI releases Grok 4.5, trained on tens of thousands of Nvidia GB300 GPUs. In coding benchmarks, the model trails Fable 5 and GPT-5.5 but needs 4.2 times fewer tokens than Opus…

xAI's Grok 4.5 tops the SWE Marathon benchmark with a 29.0% resolution rate, beating Claude Opus 4.8. Here's what it means for crypto and AI markets.