Sys.Intel // Generators/Weekly snapshot

The best AI image generators, ranked.

78 text-to-image models scored from 1.9M blind, side-by-side human votes on Arena. Refreshed every week — so you always know which generators an AI detector has to keep up with.

78
Models ranked
1.9M
Human votes
Sep 7, 2026
Data as of
8
Categories
  1. #1OpenAIPreliminary
    gpt-image-2.5-sunburst
    Proprietary · 1,104 votes
    1419
    ±21
  2. #2OpenAI
    gpt-image-2 (medium)
    Proprietary · 30,140 votes
    1389
    ±6
  3. #3OpenAIPreliminary
    gpt-image-2.5-flare
    Proprietary · 928 votes
    1388
    ±21
Text-to-Image Arena · Product, Branding & Commercial Design

Product, Branding & Commercial Design

Product shots, packaging, adverts and other commercial imagery.

1,916,350 votes · 78 models · as of Sep 7, 2026

Product, Branding & Commercial Design rankings

RankModelScore
1
OpenAI·Proprietary
1419±21
2
OpenAI·Proprietary
1389±6
3
OpenAI·Proprietary
1388±21
4
Microsoft AI·Proprietary
1336±11
5
Reve·Proprietary
1314±12
6
SpaceXAI·Proprietary
1309±19
7
Reve·Proprietary
1279±9
8
1269±7
9
Meta·Proprietary
1265±8
10
Microsoft AI·Proprietary
1261±6
11
Alibaba·Proprietary
1257±10
12
1251±9
13
1246±5
14
OpenAI·Proprietary
1240±5
15
1237±7
16
Bytedance·Proprietary
1235±6
17
Ideogram·Ideogram open model
1224±6
18
Luma AI·Proprietary
1197±9
19
Alibaba·Proprietary
1187±9
20
Reve·Proprietary
1181±6
21
Luma AI·Proprietary
1180±6
22
Microsoft AI·Proprietary
1175±6
23
SpaceXAI·Proprietary
1174±4
24
Black Forest Labs·Proprietary
1173±5
25
SpaceXAI·Proprietary
1168±5

Showing 25 of 78 models · data as of Sep 7, 2026

Source: Arena
Statistics

How settled is the top 20?

Arena scores come with a 95% confidence interval. Where intervals overlap, the order between two models is not statistically settled — that is exactly what the “rank spread” column in the table expresses.

Arena score with 95% confidence interval

Top 20 models · hover a rank for details

11501200125013001350140014501234567891011121314151617181920RANK
#1gpt-image-2.5-sunburstOpenAI1419 ±2113981440 · 1,104 votes

Average win rate against every opponent

Ties excluded · as published by Arena · a model can out-rank another yet win less often, because it has faced stronger opponents

  1. #1
    gpt-image-2.5-sunburst
    76.8%
  2. #2
    gpt-image-2 (medium)
    74.7%
  3. #3
    gpt-image-2.5-flare
    71.7%
  4. #6
    grok-imagine-image-2.0 (low)
    62.6%
  5. #5
    reve-2.1
    60.5%
  6. #4
    mai-image-2.6
    59.9%
  7. #7
    reve-2.0
    54.8%
  8. #9
    muse-image
    50.3%
  9. #15
    gemini-3-pro-image-preview (nano-banana-pro)
    49.6%
  10. #12
    gemini-3.1-flash-lite-image (nano-banana-2-lite)
    47.6%
  11. #8
    gemini-3.1-flash-image (nano-banana-2) [web-search]
    46.5%
  12. #11
    qwen-image-3.0-pro
    46.3%
  13. #10
    mai-image-2.5
    46.1%
  14. #13
    gemini-3-pro-image-2k (nano-banana-pro)
    42.6%
  15. #14
    gpt-image-1.5-high-fidelity
    42.4%
  16. #17
    ideogram-4.0-quality
    40.9%
  17. #16
    seedream-5.0-pro
    38.7%
  18. #19
    qwen-image-2.0-pro-2026-06-22
    37.0%
  19. #18
    uni-1.1-max
    36.9%
  20. #20
    reve-v1.5
    34.4%
Why we track this

Every model on this board is a model our detector has to catch.

Sightova builds AI image detection. The generators people prefer are the generators that end up in fraudulent claims, fake profiles and manipulated news — so we watch the same leaderboard the labs do, and we make it public. Right now that means keeping pace with gpt-image-2.5-sunburst from OpenAI.

Photorealism is the threat

Switch to the Photorealistic or Portraits category: those rankings show which models produce the images most likely to pass as real photographs.

Metadata-independent

Our detection reads pixels, not watermarks or EXIF, so a new generator climbing this board is a training signal, not a blind spot.

Human-judged, not benchmarked

Arena ranks by anonymous side-by-side votes from real people. It measures what people actually prefer — and therefore what gets used.

Kept weekly

Snapshots are archived every week. Movement badges show who climbed, fell or entered since the last release.

About the data

How the scores are made

On Arena, a visitor types a prompt, two anonymous models each generate an image, and the visitor votes for the better one. Millions of those pairwise votes are fitted with a Bradley–Terry style model to produce each model's Arena score.

The ± figure is a 95% confidence interval; the rank spread is the range of ranks a model could hold once that uncertainty is taken into account. Preliminary marks models added in the last release, which still have comparatively few votes.

Sightova republishes the numbers as published — no re-weighting, no editorial adjustments. Scores are rounded to whole numbers for display.

Attribution

Rankings, scores, confidence intervals and vote counts are data from Text-to-Image Arena · Product, Branding & Commercial Design by Arena Intelligence, data as of Sep 7, 2026. © Arena Intelligence. Sightova is not affiliated with, sponsored by or endorsed by Arena; we mirror the board weekly with attribution. Model and lab names are trademarks of their respective owners.

View the live board on arena.ai
FAQ
As of Sep 7, 2026, gpt-image-2.5-sunburst from OpenAI holds the top Arena score (1419 ±21). Check the rank spread before calling it a clear winner: when confidence intervals overlap, the top few models are statistically tied.
The highest-ranked model with a non-proprietary license is ideogram-4.0-quality (Ideogram open model) at #17. Use the “Open weights” filter above to see them all — note that several licenses are non-commercial.
Every week. Arena publishes a new release with a vote cutoff date; we capture it, archive it, and show movement badges relative to the previous release. The exact cutoff is printed under the table and in the attribution box.
Because the generators people prefer are the generators whose output ends up in fraud, misinformation and fake identities. Tracking this board is part of how we prioritise what our detector is trained and tested against — and publishing it keeps us honest about what is out there.
Our detector is trained on output from the major generator families and is evaluated continuously against new releases. See the testing methodology for how we measure accuracy, or upload an image to try it.