Head to head
Esc

Models · slot A

to navigate to pick

FLUX1.1 [pro] Black Forest Labs GPT Image 2 OpenAI

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX1.1 [pro]

18.4 arena score

#50 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 2

27.7 arena score

#4 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX1.1 [pro]

0%

win rate

Ties

0%

GPT Image 2

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent photorealistic lighting and window reflections on the glass and sphere.
  • + Highly realistic texture on the wooden table and red book cover.
  • + Sharp details and high visual clarity.
  • The glass container is a tall rectangular prism rather than a cube as requested.
  • The sphere appears to be levitating unnaturally rather than resting.

GPT Image 2

  • + Perfect adherence to the geometric 'cube' shape requested.
  • + Natural placement of the small sphere resting on the bottom of the cube.
  • + Soft, realistic lighting that accurately captures the window light direction.
  • The texture of the book and plant is slightly softer and less detailed than Image A.
  • The glass edges look a bit thick, resembling a fish tank more than a solid glass cube.

Verdict: Both models followed the prompt instructions very well, correctly including all requested elements. GPT Image 2 is the preferred overall choice because it accurately rendered a 'cube' whereas FLUX1.1 [pro] rendered a tall rectangle, and GPT Image 2 also placed the sphere in a more physically grounded position.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent cinematic atmosphere and lighting
  • + Captures a gritty urban feel with beautiful wet pavement reflections
  • + Follows the motion blur and framing instructions well
  • The anatomy of the bicycle is mangled and physically impossible
  • The man is not actually 'repairing' the bike, just standing over it
  • Heavy artifacts around the man's feet and the bike frame

GPT Image 2

  • + Highly realistic bicycle and tool set details
  • + Clearer adherence to the 'repairing' action with appropriate tools and posture
  • + Features readable and contextually accurate Japanese text on the sign
  • Does not capture the 'motion blur' requested in the prompt
  • The depth of field feels less like a 50mm lens and more like a standard smartphone photo
  • The lighting is flat compared to the cinematic request

Verdict: FLUX1.1 [pro] excels at the cinematic atmosphere, lighting, and mood requested by the prompt, but fails significantly on the structural integrity of the objects (the bicycle is a mess of lines). GPT Image 2 is much more grounded and realistic, successfully depicting the act of repairing with a tool kit and a coherent bicycle, making it the better choice despite its flatter lighting.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Extremely high-fidelity skin texture with realistic pores and sweat details.
  • + Excellent execution of bokeh sparks and dramatic torchlight contrast.
  • + Captures an intense, battle-worn expression that fits the 'paladin' theme well.
  • The braids are very subtle and located mostly in the background hair, making them difficult to see.
  • The armor is less prominent due to the extreme close-up crop.

GPT Image 2

  • + Excellent adherence to the 'hair braided with small beads' part of the prompt.
  • + Displays beautiful, intricate engraving on the plate armor and visible leather straps.
  • + Effective use of warm lighting and shallow depth of field for a cinematic look.
  • The character looks relatively clean compared to the 'battle-worn' and 'faint scars' request.
  • The sparks / bokeh elements are much less prominent than in the other image.

Verdict: FLUX1.1 [pro] excels in hyper-realistic textures and atmospheric lighting, creating a visceral sense of a 'battle-worn' warrior, though it misses the specific detail of the beads. GPT Image 2 provides a more balanced composition that showcases the engraved armor and braids beautifully, though its skin textures are softer and less detailed. FLUX1.1 [pro] is preferred for its superior realism and character intensity.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent presentation of the menu in a real-world context with silverware
  • + Maintains a clean, minimalist aesthetic with plenty of white space
  • Text is largely illegible gibberish
  • Food photos are small and lack specific detail

GPT Image 2

  • + Perfect text rendering for all menu items and descriptions
  • + High-quality, distinct food photography in a clear grid
  • + Strict adherence to all prompt sections (appetizers, pizza, mains)
  • Slightly less 'minimalist' than a high-end editorial design due to dense information

Verdict: GPT Image 2 performed significantly better by providing fully legible text and distinct, appetizing images for every category mentioned in the prompt. While FLUX1.1 [pro] captured a stylish lifestyle layout, its failure to generate readable text or clear food details makes it less useful as a menu design than the highly functional and professional output from GPT Image 2.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Crisp and clean text rendering.
  • + Good photorealistic texture on the meat and buns.
  • Failed to provide an 'exploded' burger, showing a mostly assembled one instead.
  • Included blue neon text which contradicts the 'fiery, glowing effect' request.
  • Failed to include the starburst for the price.

GPT Image 2

  • + Perfect adherence to the 'exploded' burger concept with suspended components.
  • + Excellent fiery, glowing text effects applied to all requested text elements.
  • + Strictly followed the starburst requirement for the price.
  • The composition is slightly more cluttered than Model A.
  • Minor artifact where some sauce droplets look a bit plastic-like.

Verdict: GPT Image 2 followed every specific detail of the prompt, including the exploded burger layout, the fiery text effects, and the price starburst. FLUX1.1 [pro] produced a professional-looking ad but ignored several key instructions, such as the exploded view and the specific fiery aesthetic for the typography.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent background blurring and café atmosphere
  • + Clean, legible handwriting style
  • Confused the octopus menu item and price, listing it twice with errors ('$228' and 'Herbss')
  • Text rendering looks more like a digital font than actual chalk on board
  • Spelling error in 'Chipkies' and 'giluten'

GPT Image 2

  • + Perfect adherence to the complex list of menu items and prices
  • + Very realistic chalk texture with dusty smudges and varying pressure
  • + Accurate handwriting variations as requested in the prompt
  • The frame of the chalkboard is slightly cropped at the bottom
  • Composition is a bit tighter than Model A

Verdict: GPT Image 2 is the clear winner as it followed all textual instructions perfectly, including specific prices and long item names without spelling errors. FLUX1.1 [pro] struggled with the specific text content, repeating lines with incorrect prices and including several typos, while also failing to capture the 'chalk texture' requested.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + High cinematic visual quality and lighting
  • + Detailed space environment
  • + Coherent horse anatomy and textures
  • Failed the negative constraint: the astronaut is riding the horse instead of the horse on top

GPT Image 2

  • + Perfect adherence to the complex spatial constraint of horse over astronaut
  • + Surreal and humorous interpretation of the prompt
  • + Clean texture rendering on the spacesuit and horse fur
  • The horse's front legs terminate awkwardly into the reins
  • The anatomy of the astronaut as a four-legged mount is slightly distorted

Verdict: While FLUX1.1 [pro] produced a much more visually stunning and professional-looking cinematic image, it completely failed to follow the specific 'horse on top' instruction. GPT Image 2 successfully navigated the difficult prompt instruction, creating a surreal scene of a horse riding an astronaut, making it the superior choice for prompt adherence.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent photographic lighting and cinematic bokeh.
  • + High resolution and very clean textures on the capybara fur.
  • + Well-rendered side profile of the businesswoman.
  • Compositional error: the woman appears to be sitting in the front passenger seat next to the driver, not the back seat.
  • The capybara's paws are not on the steering wheel as requested.
  • The capybara looks too small or incorrectly positioned relative to the seat.

GPT Image 2

  • + Accurately places the businesswoman in the back seat as requested.
  • + Successfully shows the capybara's paws on the steering wheel.
  • + The capybara's proportions and placement in the driver's seat are much more realistic.
  • The image quality is slightly softer with less sharp detail compared to the other model.
  • The cap looks a bit like a police hat rather than a standard yellow taxi cap.
  • Minor blurring on the woman's face and hands.

Verdict: GPT Image 2 followed the spatial instructions of the prompt much more accurately, correctly placing the woman in the back seat and the capybara's paws on the steering wheel. While FLUX1.1 [pro] has superior lighting and textures, it fundamentally failed the layout requirements by putting the passenger in the front seat.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent character rendering on the jack-o-lantern
  • + Clean thorn border artwork
  • + Atmospheric use of the moon and silhouettes
  • Significant text repetition and spelling errors in the invitation details
  • Missed several text elements from the prompt
  • Large empty black space at the bottom feels unbalanced

GPT Image 2

  • + Perfect adherence to all text requirements with 100% spelling accuracy
  • + Rich, detailed gothic composition with webs, thorns, and integrated NYC skyline
  • + Followed the square format request perfectly while maintaining a vintage parchment feel
  • The jack-o-lantern is a bit dark compared to the 'glowing' request
  • The border is very busy, though it fits the requested theme

Verdict: GPT Image 2 is the clear winner as it perfectly rendered all the complex text requirements, including the specific date, location, and multiple banner messages. While FLUX1.1 [pro] captures a clean illustrative style, its failure to generate correct text and its repetitive layout make it less useful as an actual invitation.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent soft 3D cartoon style consistent with the prompt
  • + Beautifully integrated soft lighting and refined textures
  • + Perfectly matches the 'minimal garnish' and clean background aesthetic
  • Missed the small flag icon requested in the prompt
  • Text for 'SUSHI' is relatively thin and lacks professional hierarchy

GPT Image 2

  • + Successfully included all prompt elements including the small flag icon
  • + Outstanding text rendering with bold, clean typography
  • + Very intricate textures on the diorama base and food items
  • Less of a 'cartoon' feel and more of a realistic CGI render
  • Diorama is much busier than the requested 'minimal' design

Verdict: Both models followed the core prompt instructions well. GPT Image 2 is the most technically complete, including the flag and delivering higher-impact typography, whereas FLUX1.1 [pro] captures the 'soft 3D cartoon' aesthetic with more finesse. GPT Image 2's inclusion of all specific UI/text elements makes it slightly superior for this specific design challenge.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent whimsical lighting and backlighting effects on fur.
  • + Very high artistic consistency and cute factor.
  • + Strong focus on the 'expressive eyes' part of the prompt.
  • Failed to include the red fox kit.
  • The animals are mostly sitting rather than 'tumbling' or 'chasing'.
  • The butterflies are glowing blobs rather than detailed insects.

GPT Image 2

  • + Successfully included all four requested animals: puppy, kitten, bunny, and fox.
  • + Better adherence to the dynamic action of 'chasing and tumbling'.
  • + Features clear god rays and more detailed butterfly renderings.
  • The fox kit has somewhat unnatural black legs/paws that look like socks.
  • The kittens face is slightly distorted in its anatomy.
  • Background wildflowers are a bit cluttered compared to the subject focus.

Verdict: GPT Image 2 is the winner because it successfully included all four specific animals requested in the prompt, whereas FLUX1.1 [pro] missed the red fox kit entirely. While FLUX1.1 [pro] has a more polished, magical aesthetic, GPT Image 2 better captured the requested action and the variety of species mentioned.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent vector emblem style that feels cohesive and professional.
  • + Sophisticated use of warm brown and cream tones.
  • + Strong decorative flourishes that match a vintage aesthetic.
  • Failed the primary text spelling by rendering 'CAFFÈ FRANILIAN' instead of 'Caffè Florian'.
  • The dome structure looks more like a building cupola than a food cloche.

GPT Image 2

  • + Perfect adherence to text, including proper spelling and accents.
  • + Accurately depicts a food cloche dome as requested.
  • + Exceptional subtle texture on the background and logo elements.
  • The 'EST. 1720' text is slightly less crisp than the main brand name.
  • The steam lines are a bit more illustrative and less 'minimalist' vector-style than Model A.

Verdict: While FLUX1.1 [pro] produced a beautiful vector emblem, it failed the core requirement of spelling the brand name correctly. GPT Image 2 not only followed all text instructions perfectly but also correctly interpreted the requested 'cloche dome' and 'subtle texture' prompts, resulting in a superior logo for the specific brand.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX1.1 [pro]
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro]

  • + Features a minimalist artistic style with a large, central rocket graphic.
  • + Uses the requested NASA-inspired color palette effectively.
  • Text is largely illegible gibberish and full of spelling errors (e.g., 'Earna Orbit', 'Desceing').
  • The logical flow of information is disorganized, numbering from 1 up to multiple 6s and placing steps out of order.
  • The icons do not match the specific requested items accurately.

GPT Image 2

  • + Excellent adherence to all six requested infographic steps with accurate icons.
  • + Perfect text rendering for headers, steps, and crew names.
  • + Clean, professional layout that genuinely resembles a modern educational poster.
  • The 'Descent' and 'Landing' modules are slightly over-detailed relative to a strictly 'flat-vector' style.
  • The Saturn V rocket shows a small amount of photo-realistic smoke at the base which breaks the vector aesthetic.

Verdict: GPT Image 2 is the clear winner for its superior information architecture and perfect text rendering. While FLUX1.1 [pro] attempted a more conceptual vector style, it failed on nearly every instructional level, including numbering, spelling, and logical flow. GPT Image 2 followed every step of the prompt and produced a coherent, usable infographic.

Next steps

Explore each model