Astra was pretrained on more than 100,000 GPUs at the Stargate facility in Texas. OpenAI researcher Aidan Clark called it the company's largest training run ever. The jump from Sol to Astra represents a bigger capability gain than the jump to Sol from earlier models, Clark said, in part because previous AI models played a role in monitoring training.
In benchmarks OpenAI published, Astra scores well above its predecessor GPT-5.6 Sol and Anthropic's Fable models. Astra hits top marks across a range of disciplines: logical reasoning (99.9 percent on ARC-AGI-3, though under its own test conditions), math (97.6 percent on FrontierMath Tier 4 v2), software engineering (74.1 percent on DeepSWE v1.1), expert knowledge (96 percent on GPQA Diamond), engineering (95.9 percent on BenchCAD), and cybersecurity (100 percent on ExploitBench).
Computer Use
Benchmark
Astra










