OpenAI’s newest flagship model just claimed the top spot in a benchmark category that measures something surprisingly practical: how good an AI is at making PowerPoint decks and Excel files that actually look professional. GPT-5.6 Sol (max) earned the highest Presentation Elo score in Artificial Analysis’s AA-Briefcase benchmark, outperforming every other model tested, including Anthropic’s Claude Fable 5 (max).

For crypto natives, the name “Sol” immediately triggers a different association. The Solana blockchain’s native token trades under the same three letters. There’s no reported connection between OpenAI’s naming convention and the Layer 1 blockchain, and no measurable impact on SOL token prices has materialized from the coincidence.

What the AA-Briefcase benchmark actually measures

The AA-Briefcase benchmark, launched on June 18, 2026, by independent evaluation firm Artificial Analysis, is designed to test AI models on complex, long-horizon professional tasks. Models are evaluated on the kind of knowledge work that fills actual corporate calendars: building presentations, analyzing data in spreadsheets, and producing polished deliverables.

The benchmark breaks performance into multiple dimensions, including overall task quality, analytical depth, presentation polish, speed, and cost efficiency. GPT-5.6 Sol (max) didn’t win across the board. It ranked second overall behind Claude Fable 5 (max), trailing on analytical metrics and broader quality evaluations.