Text-to-Image leaderboard
Models ranked by blind head-to-head votes. Scores are Elo ratings and update as new matchups complete.
Generation highlights
How the top models compare
Best models, by
Best Text-to-Image Models by Price
1 model without pricing omitted.
Best Text-to-Image Models by Speed
6 models waiting for enough speed data.
Best AI Models for Text-to-Image
| # | Model | Elo |
|---|---|---|
| 1 | 1284 | |
| 2 | 1282 | |
| 3 | 1281 | |
| 4 | 1278 | |
| 5 | 1271 | |
| 6 | 1271 | |
| 7 | 1271 | |
| 8 | 1263 | |
| 9 | 1262 | |
| 10 | 1259 |
| 11 | 1255 | |
| 12 | 1253 | |
| 13 | 1249 | |
| 14 | 1249 | |
| 15 | 1247 | |
| 16 | 1246 | |
| 17 | 1245 | |
| 18 | 1245 | |
| 19 | 1245 | |
| 20 | 1240 | |
| 21 | 1239 | |
| 22 | 1237 | |
| 23 | 1237 | |
| 24 | 1237 | |
| 25 | 1236 | |
| 26 | 1234 | |
| 27 | 1234 | |
| 28 | 1233 | |
| 29 | 1232 | |
| 30 | 1228 |
As of July 2026, Google’s Nano Banana 2 leads the leaderboard with a 1284 Elo and a 76.1% win rate, narrowly edging out its sibling, Nano Banana Pro (1281 Elo), by a slim three-point margin. The top tier remains highly competitive, with a mere 7-point gap separating the first-place model from OpenAI’s GPT Image 2 (1277 Elo), which boasts the highest overall win rate at 83.0%. Notably, budget efficiency is challenging premium dominance, as the third-ranked FLUX.2 [dev] Turbo (1279 Elo) offers top-three performance at just 12% of the cost per image compared to the runner-up.
Rivalries
Aggregate head-to-head across the arena
Highlighted challenges
The Reversed Rodeo
This competition tests how well AI image models truly understand language versus how much they rely on visual habits from their training data. The prompt is deliberately simple on the surface but devilishly hard in practice. Most models default to the familiar trope of an astronaut riding a horse. By forcing the reversal, we measure three critical capabilities that separate good models from great ones: Strict instruction following (including negations) Accurate subject-object relationships and spatial hierarchy Resistance to strong dataset biases
Geometric Composition
A spatial-reasoning test. Each object has a precise relationship (inside, on top, behind, seen through the glass), so it measures whether a model follows explicit placement instructions and handles transparency and refraction rather than just approximating the scene.
Recent SOTA shifts in Text-to-Image
Full history- Gemini 3.1 Flash Image Preview overtook Gemini 3 Pro Image Preview
- Gemini 3 Pro Image Preview overtook ImagineArt 1.5 (Preview)
- ImagineArt 1.5 (Preview) overtook GPT Image 1 Mini
- GPT Image 1 Mini overtook FLUX.1 Kontext [max]
- FLUX.1 Kontext [max] overtook FLUX.1 [dev]
- FLUX.1 [dev] overtook DALL-E 3
- DALL-E 3 overtook DALL-E 2
FAQ
What is the best AI text-to-image model?
Based on blind community voting, Nano Banana 2 is currently the #1 ranked AI text-to-image model with an Elo rating of 1284. Rankings update in real time as new votes come in.
How are AI text-to-image models ranked on Lumenfall?
Lumenfall Arena ranks AI models through blind community voting. In each matchup, two models generate from the same prompt and voters pick the better result without seeing model names. Votes are processed using TrueSkill, a Bayesian rating algorithm developed by Microsoft Research, that produces a single Elo score reflecting each model's relative quality.
What is an Elo rating for AI models?
An Elo rating is a numerical score representing a model's skill relative to other models. Under the hood, Lumenfall uses TrueSkill, which tracks two values per model: mu (estimated skill) and sigma (uncertainty). The displayed Elo is calculated as 1000 + 10 x (mu - 3*sigma), a conservative lower bound. A model must prove itself consistently across many matchups to earn a high rating.
Keep the arena honest
Cast your vote
Pick winners in blind matchups. Every vote nudges the Elo and shapes these rankings.
Cast Your VoteSuggest a prompt
Got an idea worth testing? Submit a prompt and watch the models battle it out.