The best AI image generators, ranked.
78 text-to-image models scored from 2.4M blind, side-by-side human votes on Arena. Refreshed every week — so you always know which generators an AI detector has to keep up with.
- 78
- Models ranked
- 2.4M
- Human votes
- Sep 7, 2026
- Data as of
- 8
- Categories
- 1440±21
- 1428±21
- 1400±6
Cartoon, Anime & Fantasy
Stylised illustration, anime and fantasy art.
2,437,796 votes · 78 models · as of Sep 7, 2026
Cartoon, Anime & Fantasy rankings
| Rank | Model | Score | ||
|---|---|---|---|---|
1 | 1 – 2 | gpt-image-2.5-sunburstPreliminary OpenAI·Proprietary | 1440±21 | 1,198 |
2 | 1 – 2 | gpt-image-2.5-flarePreliminary OpenAI·Proprietary | 1428±21 | 1,149 |
3 | 3 | OpenAI·Proprietary | 1400±6 | 31,995 |
4 | 4 – 5 | Microsoft AI·Proprietary | 1348±10 | 4,634 |
5 | 4 – 6 | grok-imagine-image-2.0 (low)Preliminary SpaceXAI·Proprietary | 1322±19 | 935 |
6 | 5 – 6 | Reve·Proprietary | 1310±11 | 2,920 |
7 | 7 – 9 | Meta·Proprietary | 1286±8 | 9,957 |
8 | 7 – 13 | Alibaba·Proprietary | 1276±10 | 4,082 |
9 | 7 – 13 | Reve·Proprietary | 1275±9 | 5,562 |
10 | 8 – 13 | Bytedance·Proprietary | 1271±7 | 23,116 |
11 | 8 – 13 | Microsoft AI·Proprietary | 1270±6 | 22,528 |
12 | 8 – 13 | Google·Proprietary | 1263±7 | 15,963 |
13 | 8 – 14 | Google·Proprietary | 1258±9 | 6,253 |
14 | 13 – 16 | OpenAI·Proprietary | 1245±5 | 59,294 |
15 | 14 – 16 | Google·Proprietary | 1242±5 | 61,392 |
16 | 14 – 16 | Google·Proprietary | 1238±7 | 33,270 |
17 | 17 – 22 | Ideogram·Ideogram open model | 1195±6 | 16,324 |
18 | 17 – 24 | Luma AI·Proprietary | 1194±9 | 5,067 |
19 | 17 – 24 | Alibaba·Proprietary | 1191±9 | 4,471 |
20 | 17 – 24 | Luma AI·Proprietary | 1190±7 | 12,472 |
21 | 17 – 26 | Nvidia·OpenMDW 1.1 | 1183±11 | 2,897 |
22 | 18 – 26 | Microsoft AI·Proprietary | 1180±7 | 20,299 |
23 | 18 – 26 | Nvidia·OpenMDW 1.1 | 1179±9 | 4,844 |
24 | 17 – 30 | Recraft·Proprietary | 1177±19 | 923 |
25 | 21 – 26 | SpaceXAI·Proprietary | 1176±4 | 96,620 |
Showing 25 of 78 models · data as of Sep 7, 2026
How settled is the top 20?
Arena scores come with a 95% confidence interval. Where intervals overlap, the order between two models is not statistically settled — that is exactly what the “rank spread” column in the table expresses.
Arena score with 95% confidence interval
Top 20 models · hover a rank for details
Average win rate against every opponent
Ties excluded · as published by Arena · a model can out-rank another yet win less often, because it has faced stronger opponents
- #186.2%gpt-image-2.5-sunburst
- #276.3%gpt-image-2.5-flare
- #375.0%gpt-image-2 (medium)
- #460.6%mai-image-2.6
- #960.5%reve-2.0
- #659.1%reve-2.1
- #559.1%grok-imagine-image-2.0 (low)
- #753.2%muse-image
- #1647.7%gemini-3-pro-image-preview (nano-banana-pro)
- #1347.5%gemini-3.1-flash-lite-image (nano-banana-2-lite)
- #1245.7%gemini-3.1-flash-image (nano-banana-2) [web-search]
- #845.4%qwen-image-3.0-pro
- #1145.2%mai-image-2.5
- #1044.7%seedream-5.0-pro
- #1442.1%gpt-image-1.5-high-fidelity
- #1541.9%gemini-3-pro-image-2k (nano-banana-pro)
- #1934.4%qwen-image-2.0-pro-2026-06-22
- #1734.1%ideogram-4.0-quality
- #1831.9%uni-1.1-max
- #2031.1%uni-1.1
Every model on this board is a model our detector has to catch.
Sightova builds AI image detection. The generators people prefer are the generators that end up in fraudulent claims, fake profiles and manipulated news — so we watch the same leaderboard the labs do, and we make it public. Right now that means keeping pace with gpt-image-2.5-sunburst from OpenAI.
Photorealism is the threat
Switch to the Photorealistic or Portraits category: those rankings show which models produce the images most likely to pass as real photographs.
Metadata-independent
Our detection reads pixels, not watermarks or EXIF, so a new generator climbing this board is a training signal, not a blind spot.
Human-judged, not benchmarked
Arena ranks by anonymous side-by-side votes from real people. It measures what people actually prefer — and therefore what gets used.
Kept weekly
Snapshots are archived every week. Movement badges show who climbed, fell or entered since the last release.
How the scores are made
On Arena, a visitor types a prompt, two anonymous models each generate an image, and the visitor votes for the better one. Millions of those pairwise votes are fitted with a Bradley–Terry style model to produce each model's Arena score.
The ± figure is a 95% confidence interval; the rank spread is the range of ranks a model could hold once that uncertainty is taken into account. Preliminary marks models added in the last release, which still have comparatively few votes.
Sightova republishes the numbers as published — no re-weighting, no editorial adjustments. Scores are rounded to whole numbers for display.
Rankings, scores, confidence intervals and vote counts are data from Text-to-Image Arena · Cartoon, Anime & Fantasy by Arena Intelligence, data as of Sep 7, 2026. © Arena Intelligence. Sightova is not affiliated with, sponsored by or endorsed by Arena; we mirror the board weekly with attribution. Model and lab names are trademarks of their respective owners.
View the live board on arena.ai