Artificial Analysis has released version 4.2 of its Intelligence Index, likely in response to criticism that its benchmarks failed to capture GPT-6 Astra's actual progress. Astra now scores four points above its predecessor but still trails Anthropic's Claude Fable 5.1.

OpenAI launched Astra and its GPT-5.6 family to claim it has surpassed Anthropic, but revenue numbers and benchmark disputes complicate the

Viral claims of GPT-6 Astra scoring 98.6% on ARC-AGI-3 lack verification. The actual leaderboard leader, Claude Opus 5, sits at roughly 30.2%.

An impressive system that can (to some unknown extent) build symbolic world models