Product, Branding & Commercial · Text-to-Image
Elo rankings from blind votes across 6 challenges in this category.
Highlights
10 imagesBest models, by
Best Text-to-Image Models by Price
1 model without pricing omitted.
Best Text-to-Image Models by Speed
5 models waiting for enough speed data.
Best AI Models for Product, Branding & Commercial
| # | Model | Elo |
|---|---|---|
| 1 | 1308 | |
| 2 | 1283 | |
| 3 | 1279 | |
| 4 | 1265 | |
| 5 | 1258 | |
| 6 | 1253 | |
| 7 | 1249 | |
| 8 | 1248 | |
| 9 | 1245 | |
| 10 | 1243 |
| 11 | 1241 | |
| 12 | 1237 | |
| 13 | 1235 | |
| 14 | 1228 | |
| 15 | 1226 | |
| 16 | 1226 | |
| 17 | 1220 | |
| 18 | 1219 | |
| 19 | 1216 | |
| 20 | 1215 | |
| 21 | 1214 | |
| 22 | 1209 | |
| 23 | 1206 | |
| 24 | 1205 | |
| 25 | 1204 | |
| 26 | 1203 | |
| 27 | 1203 | |
| 28 | 1197 | |
| 29 | 1197 | |
| 30 | 1195 |
FLUX.2 [dev] Flash leads the category with a 1308 Elo and 61.5% win rate, outperforming premium rivals like Nano Banana 2 (1283 Elo) at a significantly lower $0.005/img price point. While specialized branding models like Recraft V4 hold high price tiers, OpenAI's GPT Image 1.5 and 2 remain more competitive in this commercial space, maintaining a 25+ Elo lead over Black Forest Labs' flagship FLUX.2 [pro].
Rivalries
Aggregate head-to-head across the arena
Highlighted challenges
The Reversed Rodeo
This competition tests how well AI image models truly understand language versus how much they rely on visual habits from their training data. The prompt is deliberately simple on the surface but devilishly hard in practice. Most models default to the familiar trope of an astronaut riding a horse. By forcing the reversal, we measure three critical capabilities that separate good models from great ones: Strict instruction following (including negations) Accurate subject-object relationships and spatial hierarchy Resistance to strong dataset biases
Geometric Composition
A spatial-reasoning test. Each object has a precise relationship (inside, on top, behind, seen through the glass), so it measures whether a model follows explicit placement instructions and handles transparency and refraction rather than just approximating the scene.
Keep the arena honest
Cast your vote
Pick winners in blind matchups. Every vote nudges the Elo and shapes these rankings.
Cast Your VoteSuggest a prompt
Got an idea worth testing? Submit a prompt and watch the models battle it out.