Head to head
Esc

Models · slot A

to navigate to pick

FLUX1.1 [pro] Black Forest Labs GPT Image 1.5 OpenAI

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX1.1 [pro]

18.4 arena score

#50 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1.5

27.1 arena score

#6 of 62 in Text-to-Image

Top 3 in Image Editing
Vote tally

Where the votes landed

FLUX1.1 [pro]

0.0%

win rate

Ties

0.0%

GPT Image 1.5

100.0%

win rate

0.0% 0.0% ties 100.0%
Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent photorealistic texture on the book cover and wood grain.
  • + Highly detailed glass refractions and reflections with high clarity.
  • + Dramatic, professional-style lighting and bokeh.
  • The 'cube' is significantly taller than it is wide, appearing more like a rectangular prism.
  • The sphere is levitating rather than sitting on a surface, which might be an unintended interpretation.

GPT Image 1.5

  • + Perfectly adheres to the 'cube' geometry with equal dimensions.
  • + Realistic physical interaction with the sphere resting on the bottom of the glass.
  • + Accurately represents the soft window light from the left as requested.
  • The plant visibility through the glass looks slightly flat/less refractive than Model A.
  • Overall image sharpness and texture detail are slightly lower than competitors.

Verdict: While FLUX1.1 [pro] produces a more visually striking and detailed image with superior textures, it fails to generate an actual cube, providing a rectangular prism instead. GPT Image 1.5 followed the spatial and geometric instructions perfectly, rendering a true cube with a resting sphere that creates a more believable physical scene.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent atmosphere and cinematic lighting.
  • + Good adherence to 'imperfect framing' with the subjects positioned off-center.
  • + Captures reflections and a wet street aesthetic effectively.
  • Serious anatomical and structural failures: the man's legs and the bicycle frame are nonsensical and mangled.
  • The bicycle is missing its rear wheel entirely.

GPT Image 1.5

  • + High anatomical and structural realism in both the man and the bicycle.
  • + Great attention to detail with the repair tools and the 'light rain' visible on the man's jacket.
  • + Accurate representation of the requested 50mm shallow depth of field.
  • Lacks the requested 'motion blur from passing cars' as the vehicle in the background is static.
  • Framing is more conventional and less 'imperfect' than requested.

Verdict: While FLUX1.1 [pro] captures the requested cinematic mood and imperfect framing better, it falls apart upon close inspection with gross anatomical errors and a nonsensical bicycle structure. GPT Image 1.5 provides a much more coherent and realistic image with impressive textures and details, making it the superior choice despite missing the motion blur requirement.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Extremely lifelike eye textures and reflections
  • + Excellent skin pore and individual hair strand resolution
  • + Subtle and realistic dirt application
  • Missed the beaded hair requirement almost entirely
  • The armor engraving is less ornate compared to the competition

GPT Image 1.5

  • + Excellent adherence to the 'hair braided with small beads' prompt
  • + Beautiful ornate engraved plate armor with visible leather straps
  • + Better representation of warm torchlight and glowing bokeh
  • Scars look slightly painted on rather than integrated into the skin
  • The hair textures are a bit clumped and less realistic than the other model

Verdict: GPT Image 1.5 is the overall winner for its superior prompt adherence, particularly regarding the braided beads and the specific details of the leather straps and ornate armor. While FLUX1.1 [pro] achieved a higher level of photographic realism in the eyes and skin, it failed to incorporate several specific elements requested in the prompt.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent presentation for a marketing mockup
  • + Follows the 'grid' request for photos very well
  • + Dynamic layout with depth using shadows and cutlery
  • Text is largely illegible and contains gibberish
  • The section headings do not accurately reflect the prompt (e.g., 'Man & Dreafry')

GPT Image 1.5

  • + Perfect text rendering with clear, readable fonts
  • + Accurate adherence to all requested sections: Appetizers, Pizza, and Mains
  • + Clear, high-quality food photography
  • Slightly more traditional layout than 'modern minimalist' requested
  • The grid of photos is restricted to the side rather than being integrated into the main design

Verdict: While FLUX1.1 [pro] creates a more stylish and visually attractive mockup, GPT Image 1.5 is the superior tool for an actual menu design due to its perfect text legibility and strict adherence to the requested content sections. FLUX1.1 fails significantly on the textual requirements, producing gibberish despite the nice aesthetic.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent typography rendering with clean, professional neon effects
  • + High-quality photographic texture on the burger patties and bun
  • + Clear, uncluttered composition with a good sense of depth
  • Failed to create an 'exploded' view; the burger is mostly assembled
  • Includes redundant text by repeating 'Magic Burger' twice
  • Missing the starburst element for the price

GPT Image 1.5

  • + Perfectly captures the 'exploded' effect with all components suspended separately
  • + Accurately includes all requested text elements, including the fiery starburst
  • + Strongly adheres to the fiery/burning atmosphere requested in the prompt
  • The image is slightly oversaturated and busy
  • The 'Magic Burger' text at the top is partially cropped

Verdict: GPT Image 1.5 followed the complex layout instructions much better than FLUX1.1 [pro], successfully rendering the exploded burger view and the price starburst. While FLUX1.1 [pro] produced a cleaner, more professional-looking advertisement, it failed the core requirement of showing the individual components suspended in mid-air.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent background depth and atmospheric lighting in the café setting.
  • + Clean and legible text rendering.
  • Failed the prompt by splitting the 'Grilled Octopus' item into two separate lines with incorrect prices ($228 and $98).
  • Text looks more like a digital font overlay than real chalk on a board.
  • Added extra text 'Chipkies' that is misspelling 'Cookies'.

GPT Image 1.5

  • + Perfect adherence to the menu item list and prices.
  • + Realistic chalk texture with dusty smudges and varying pressure.
  • + Authentic handwriting style that matches the 'hand-drawn' request.
  • The composition is a tight crop, showing less of the 'cozy café' environment.
  • Minimalist background compared to the other model.

Verdict: GPT Image 1.5 is the clear winner as it followed every detail of the prompt, including the specific menu items and prices, while maintaining a very realistic chalk-on-blackboard texture. FLUX1.1 [pro] failed the text accuracy significantly by hallucinating prices ($228) and repeating menu lines.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent high-contrast cinematic lighting and atmospheric clouds.
  • + High level of detail in the horse's fur and the astronaut's suit.
  • Failed the negative constraint: the astronaut is on top of the horse instead of the horse on top of the astronaut.

GPT Image 1.5

  • + Highly detailed space environment including planets, nebulae, and a lunar lander.
  • + Sharp rendering of textures on the lunar surface and asteroid debris.
  • Failed the negative constraint: the astronaut is on top of the horse instead of the horse on top of the astronaut.
  • Anatomical issues with the horse's front legs and hoof structure.

Verdict: Both FLUX1.1 [pro] and GPT Image 1.5 completely ignored the specific spatial instruction to place the 'horse on top' of the astronaut, instead opting for the cliche astronaut-riding-horse composition. FLUX1.1 [pro] is slightly preferred for its superior artistic composition and cleaner anatomical rendering, as GPT Image 1.5 suffers from cluttered background elements and distorted horse limbs.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent atmospheric lighting and bokeh effect through the windshield
  • + High aesthetic quality with clean textures and a cinematic feel
  • + Accurate interpretation of the 'bored' expression for the human character
  • The capybara and the passenger appear to be sitting side-by-side in the front seat instead of front/back.
  • One paw is missing from the steering wheel, failing the 'both front paws' prompt.
  • The perspective feels cramped and less like a realistic vehicle interior.

GPT Image 1.5

  • + Perfect adherence to the spatial requirements, showing front seat vs back seat layout.
  • + Exactly follows the instruction for 'both front paws on the steering wheel'.
  • + Great detail on the capybara's professional expression and specifically requested clothing.
  • The passenger's face and hands are slightly soft/blurry compared to the foreground.
  • Overall lighting is a bit flatter and less cinematic than the competitor.

Verdict: While FLUX1.1 [pro] has a more polished and artistic photographic look, GPT Image 1.5 is the clear winner for prompt adherence. GPT Image 1.5 correctly positioned the passenger in the back seat and placed both of the capybara's paws on the wheel, whereas FLUX1.1 [pro] incorrectly sat the passenger in the front and missed the steering wheel interaction.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Includes all elements of the prompt including the thorn border and bats
  • + High contrast lighting on the jack-o-lantern
  • Significant text errors including redundant 'Party' word and repetitive event details
  • Graphic design layout is unbalanced with too much black space at the bottom
  • Failed the square format request, producing a portrait-oriented poster

GPT Image 1.5

  • + Excellent layout that follows the square format request perfectly
  • + Accurate and well-rendered text for all requested fields
  • + Superior gothic aesthetic with a convincing dark parchment texture
  • The jack-o-lantern is slightly off-center
  • The 'thorns' in the border are somewhat repetitive in texture

Verdict: GPT Image 1.5 is the clear winner as it followed all instructions, including the square format and specific text strings, with zero typos. FLUX1.1 [pro] produced a portrait-oriented image with redundant, misspelled text and a less cohesive design despite having good individual illustrative elements.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent adherence to the 'cartoon' and 'soft refined texture' style requirements.
  • + Perfectly clean and minimalist aesthetic for the diorama base.
  • + High-quality text rendering for 'JAPAN' and 'SUSHI'.
  • Missed the request for a flag icon.
  • The garnish includes items like tomatoes which are not typical for sushi.

GPT Image 1.5

  • + Includes the requested flag icon.
  • + Rich, realistic PBR materials with impressive texture on the wood and ceramics.
  • + Excellent 45-degree isometric composition with more complex accessories like the teapot and soy sauce.
  • The scene is much busier than the 'minimal' request.
  • The 'cartoon' style specified in the prompt is less apparent, leaning more toward realism.

Verdict: FLUX1.1 [pro] better captured the 'cartoon scene' and 'minimal' aesthetic requested, resulting in a cleaner, more focused graphic. However, GPT Image 1.5 was the only model to include all prompt elements (the flag icon) and showcased superior technical mastery of PBR materials and realistic textures, even if it strayed from the 'casual cartoon' vibe.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX1.1 [pro]
GPT Image 1.5
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent backlighting and bokeh effect
  • + Clean and aesthetically pleasing composition
  • Failed to include the fox kit requested in the prompt
  • Animals are sitting still rather than 'tumbling together'

GPT Image 1.5

  • + Excellent prompt adherence including all four specific animals
  • + Captures the 'tumbling together' action much better
  • + Beautiful god rays and dew sparkles as requested
  • Some minor anatomical blending where the animals overlap
  • The butterfly sizing is slightly inconsistent

Verdict: GPT Image 1.5 is the clear winner as it successfully included all four requested animals (puppy, kitten, bunny, and fox), whereas FLUX1.1 [pro] missed the fox entirely. GPT Image 1.5 also captured the dynamic 'tumbling' energy of the prompt, while FLUX1.1 [pro] produced a more static, though very pretty, portrait.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent vintage illustrative style with fine cross-hatching detail
  • + Beautiful border ornamentation that enhances the vintage aesthetic
  • + The layout is balanced and traditional.
  • Serious spelling error in the primary brand name (FRATILIAN instead of FLORIAN)
  • The steam is very faint and hard to see

GPT Image 1.5

  • + Perfect text rendering of the restaurant name
  • + Cleverly integrated 'Est. 1720' into a clear banner as requested
  • + Strong contrast and clean vector style elements
  • Failed to follow the 'light background' prompt, opting for black instead
  • The texture is a bit blotchy on the dome rather than 'subtle'

Verdict: GPT Image 1.5 is the clear winner for its superior text accuracy, correctly spelling 'Caffè Florian' while maintaining the requested vintage color scheme. Although FLUX1.1 [pro] produced a more intricate illustrative style and followed the light background instruction, the significant misspelling of the brand name makes it less useful as a logo design.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX1.1 [pro]
GPT Image 1.5

AI Judge Analysis

FLUX1.1 [pro]

  • + Excellent vector art style with professional, clean lines.
  • + Captures the NASA navy and light gray palette well.
  • + Creative use of a large central moon with connecting data points.
  • Text is largely unintelligible and includes many typos (e.g., 'Apolo 1', 'Infor the').
  • The sequence logic and iconography are messy and redundant (listing '1' twice).
  • Failed to include clear, distinct icons for all 6 requested steps.

GPT Image 1.5

  • + Perfect adherence to the 6 requested steps with clear, legible text.
  • + Consistent iconography following the requested flat-vector style.
  • + Accurate names of the astronauts and location (Tranquillity) added intelligently.
  • Composition feels more like a collection of comic panels than a single cohesive poster.
  • The silhouette of the earth in 'Earth Orbit' uses a slightly different green that deviates from the strictly requested palette.

Verdict: GPT Image 1.5 is the clear winner as it perfectly follows the instructional requirements of the prompt, providing all six sequential steps with legible text and accurate iconography. While FLUX1.1 [pro] has a more sophisticated 'high-end' design aesthetic, its failure to render readable text and its confusing, repetitive numerical labeling makes it a failure as an infographic.

Next steps

Explore each model