Microsoft’s MAI-Image-2.6 Is No. 2 on Arena. The Product Test Is Still to Come
Microsoft’s latest in-house text-to-image model is second in Arena’s August 10 ranking, behind OpenAI’s GPT-Image-2. The result is an early competitive signal from a live head-to-head-vote leaderboard, while Microsoft’s promised reference, grounding and control features—and their wider rollout—remain to be detailed.
- MAI-Image-2.6 is second in Arena’s August 10 text-to-image snapshot, behind OpenAI’s GPT-Image-2.
- The preview had 3,488 comparison votes, versus 70,065 for the leader and 46,329 for Microsoft’s older MAI-Image-2.5 entry.
- Microsoft has described new reference, grounding and control capabilities, but has not yet explained how they work or when all will ship.
Microsoft’s MAI-Image-2.6, its latest in-house text-to-image model, has reached No. 2 in Arena’s public ranking. That puts the new preview ahead of a crowded field in this particular measure, but it does not yet establish that Microsoft has delivered the most capable image product—or even that the rank will be durable as more comparisons arrive.
The distinction matters because Arena is not a feature checklist or a standardised test suite. Its leaderboards aggregate human choices from head-to-head battles into model ratings. The August 10 text-to-image snapshot therefore captures a live comparative result, not evidence for every capability Microsoft is promoting.

Microsoft’s official artwork accompanying its MAI-Image-2.6 announcement. Source: Neowin.

Vote counts for the top three models in Arena’s August 10 text-to-image leaderboard. Source: Arena.
No. 2 is a result with a small early record
The leaderboard lists mai-image-2.6-preview at 1336±11 from 3,488 votes. Its displayed rank spread is 2–3; Arena says that field expresses the range of possible ranks based on confidence intervals. The leading gpt-image-2 (medium) is at 1381±5 after 70,065 votes. The third-place Grok image variant is at 1316±12 and marked preliminary after 2,676 votes.