Microsoft's documented position on MAI-Image-2 is more specific than broad claims about a global text-to-image leaderboard. The company says MAI-Image-2 is the third-ranked model family on Arena.ai, while MAI-Image-2.5 holds a No. 2 ranking for image editing on Arena. Those are distinct measurements, involving different model labels and leaderboard categories.
The distinction matters for developers and enterprise buyers evaluating image-generation systems. A ranking can be useful evidence of competitive performance, but its meaning depends on the benchmark, the task being measured, the model variant, and the point in time when the leaderboard was viewed. Treating an image-editing result as a general text-to-image result can materially overstate what an official announcement establishes.
Microsoft's official MAI-Image-2 announcement describes the model family as third-ranked on Arena.ai. It also identifies MAI-Image-2.5 as having launched with a No. 2 ranking in image editing on Arena. The announcement does not document a public MAI-Image-2.6 release or assign that version a world No. 2 text-to-image position.
The ranking categories are not interchangeable
Arena leaderboards can represent different image-generation tasks. Text-to-image evaluation concerns the creation of images from written prompts. Image editing evaluates a different workflow: modifying an existing image in response to instructions. A strong placement in either category is relevant, but it is not automatically transferable to the other.








