Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 [dev] Black Forest Labs GPT Image 1 OpenAI

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 [dev]

24.6 arena score

#16 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1

22.6 arena score

#32 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 [dev]

0%

win rate

Ties

0%

GPT Image 1

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Sophisticated glass rendering with realistic internal reflections and refraction.
  • + Exhibits high photographic quality with naturalistic wood grain and soft bokeh.
  • + Accurately captures the lighting source from the left as requested.
  • The plant is positioned more to the side than 'behind' compared to the other model.
  • The sphere appears to be floating mid-air rather than resting on the bottom of the cube.

GPT Image 1

  • + Follows the spatial arrangement perfectly, with the plant clearly visible behind the glass.
  • + Clear, logical geometry in the composition of the cube and book.
  • + Correctly interprets the sphere's placement inside the cube.
  • The glass cube is rendered with thick, teal-tinted edges that look a bit artificial.
  • Less impressive handling of complex refractions through the glass compared to FLUX.1 [dev].

Verdict: Both models followed the prompt successfully, including all requested elements and lighting directions. FLUX.1 [dev] produced a more aesthetically pleasing and photorealistic image with superior glass physics, while GPT Image 1 provided a more literal and accurate spatial arrangement of the plant behind the cube.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent handling of wet pavement reflections and light rain particles
  • + Cinematic composition with a realistic 50mm lens feel
  • + Highly detailed clothing textures and realistic skin
  • The man is holding the handlebars rather than performing a repair
  • The bicycle design is a bit modern and generic for the 'candid' prompt

GPT Image 1

  • + Strong adherence to the 'repairing' action with a crouching pose
  • + Captures the 'imperfect framing' and 'candid' feel more effectively
  • + Very natural, non-stylized skin texture on the man's face
  • The bicycle frame geometry is physically nonsensical near the rear hub
  • Less noticeable motion blur on passing cars compared to Image A

Verdict: Both models followed the prompt well, but they succeeded in different areas. FLUX.1 [dev] produced a more polished, high-quality image with superior environmental effects, while GPT Image 1 better captured the specific action of 'repairing' and the requested 'candid, imperfect framing.' FLUX.1 [dev] is the likely winner due to the significant structural errors in the bicycle and less convincing depth of field in the GPT output.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent photographic skin texture and lifelike eyes
  • + Vibrant and high-quality bokeh effect
  • + Clean and symmetrical composition
  • Missed the 'small beads' in the hair braids
  • The character looks too pristine and model-like for the 'battle-worn' prompt
  • Plate armor is basic rather than 'ornate engraved'

GPT Image 1

  • + Includes small beads in the braids and ornate engravings on the armor
  • + Character has a realistic 'battle-worn' expression with dirt and grime
  • + More dramatic and atmospheric warm lighting
  • Slightly lower skin texture resolution compared to Model A
  • Composition is a bit tighter, cutting off some detail
  • Armor texture looks slightly more digital than metallic

Verdict: GPT Image 1 followed the prompt's specific details much better, capturing the ornate engravings, the beads in the hair, and a truly battle-worn appearance. FLUX.1 [dev] produced a more aesthetically pleasing portrait with superior skin texture, but it failed on several key prompt descriptors like the engravings and beads, and the character looks far too clean for the context.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent professional graphic design layout with whitespace
  • + Includes a clean restaurant branding area
  • + High quality, integrated food photography
  • Text consists largely of jibberish/placeholder characters
  • Relies on a diagonal layout rather than a standard grid

GPT Image 1

  • + Successfully follows the requested grid layout for food photos
  • + Text is legible and uses bold sans-serif fonts as requested
  • + Includes specific sections for Appetizers and Pizza
  • Lacks the 'Mains' category section requested in the prompt
  • Slightly less 'professional' or 'modern' aesthetic compared to Model A
  • Text description is repetitive and contains minor typos like 'descrigion'

Verdict: FLUX.1 [dev] produces a more convincing professional design with a sophisticated layout, but fails to provide legible text or the specific grid requested. GPT Image 1 follows the grid structure and legibility requirements much more closely, though it missed the specific 'Mains' heading. GPT Image 1 is the winner for better prompt adherence regarding content and layout structure.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent photorealistic texture on the meat patties and vegetables
  • + Clean, professional composition with a focused lighting effect
  • Completely failed to include the primary 'MAGIC BURGER' text
  • Failed to render the price in a starburst as requested
  • The explosion effect is very vertical and static, lacking dynamic motion

GPT Image 1

  • + Successfully integrated all requested text including the fiery effect
  • + Perfectly followed the instruction for the starburst price tag
  • + Dynamic composition with embers and liquid splashes creating a sense of motion
  • The price '€.99' is slightly incorrect, missing the '6'
  • The 'MAGIC BURGER' text is cut off at the top of the frame

Verdict: GPT Image 1 followed the complex prompt instructions much more effectively than FLUX.1 [dev], capturing the specific text elements, the starburst, and the fiery background style. While FLUX.1 [dev] produced a more realistic burger texture, it failed to include the primary title and several key layout requirements.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent text rendering with no spelling errors.
  • + Beautiful high-contrast composition and realistic café background.
  • + Followed the prompt to complete the third menu item creatively with cookies.
  • The 'chalk' texture looks a bit more like a digital marker or smooth paint than dry chalk.
  • The handwriting style for the body text looks slightly too uniform, like an 'ink' font.

GPT Image 1

  • + Exceptional realistic chalk texture, including the grainy edges of the strokes.
  • + The title font feels more authentic to a handwritten chalkboard style.
  • Missing the dollar sign for the last item (Cookies 9).
  • Lower visual contrast compared to the first image, making it look a bit dull.
  • The letter 'i' in 'Risotto' looks slightly malformed.

Verdict: FLUX.1 [dev] produced a much cleaner and more professional-looking image with perfect spelling and high-contrast visuals, though the text lacks a true chalky grain. GPT Image 1 captured the realistic texture of chalk much better, but failed on minor details like the currency symbol for the final price and overall image lighting. FLUX.1 [dev] is the winner for its clarity and complete adherence to the text requirements.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Features a smooth, ethereal lighting that fits the cinematic request.
  • + The white horse and suit create a cohesive color palette.
  • Completely failed the semantic challenge of 'horse on top' of the astronaut.
  • Anatomical issues with the horse's legs, including a strange double-jointed appearance on the hind leg.

GPT Image 1

  • + Excellent texture work on the space suit and horse hair.
  • + Higher level of detail in the background with visible planets and nebulae.
  • Failed to follow the specific instruction to put the 'horse on top' of the astronaut.
  • The composition is a cliché interpretation of the 'space cowboy' trope despite the prompt's push for surrealism.

Verdict: Both models completely failed to follow the logical reversal requested in the prompt ('horse on top, not vice versa'), instead providing standard images of astronauts riding horses. GPT Image 1 is the superior image due to its higher texture detail, better anatomical accuracy of the horse, and more complex background compared to FLUX.1 [dev].

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + High resolution and crisp rendering of the capybara's fur.
  • + Very aesthetically pleasing lighting and bokeh effect.
  • + Accurately represents the 'bored' expression requested for the passenger.
  • The passenger is sitting in the front passenger seat rather than the back seat as requested.
  • The capybara's paws are not placed correctly on the steering wheel, appearing to float or hold nothing.
  • The steering wheel is on the right side, which is incorrect for a New York taxi.

GPT Image 1

  • + Accurately places the passenger in the back seat as requested.
  • + Correctly places both paws of the capybara on the steering wheel.
  • + The taxi driver cap is much more authentic to a professional uniform.
  • The passenger's face is slightly out of focus and less detailed.
  • The steering wheel is also on the right side, which contradicts the New York setting.
  • The overall image is slightly darker and less vibrant than the competitor.

Verdict: While FLUX.1 [dev] produced a more visually striking image with superior fur texture and lighting, it failed major compositional requirements of the prompt by placing the passenger in the front seat. GPT Image 1 followed the spatial instructions much better, placing the passenger in the back and correctly positioning the capybara's paws on the wheel, making it the more accurate interpretation despite slightly lower aesthetic polish.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Strong cinematic lighting with a vibrant glowing effect
  • + Intricate thorny border framing the entire composition
  • + High contrast and sharp, clean illustration style
  • Several text hallucinations including 'Falloween Rantcy' and 'You Tre'
  • Formatting errors in the event details such as '30, 10, 2026' and repeated time tags

GPT Image 1

  • + Perfect adherence to the requested text for the header and subtitle banner
  • + Captures the 'vintage gothic' aesthetic well with muted tones and parchment texture
  • + Includes both webs and thorns in the border as requested
  • The location 'The Arches' is incorrectly mapped to the 'TIME' label
  • Visuals are a bit muddy and lower contrast compared to Model A

Verdict: GPT Image 1 is the superior choice because it successfully renders the complex header and subtitle text exactly as requested, which is critical for an invitation. While FLUX.1 [dev] has more striking visual lighting and a cleaner border, its significant legibility errors and strange spelling hallucinations make it unusable as a party invitation.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent soft-focus photography aesthetic
  • + Beautifully rendered translucent textures on the fish
  • + Accurate 45-degree isometric projection
  • Text rendering is poor with spelling errors like 'SUSH CATON'
  • Background/text relationship is cluttered by the flowers
  • Contrast is a bit low

GPT Image 1

  • + Perfect text rendering for both 'JAPAN' and 'SUSHI'
  • + Strong '3D cartoon' aesthetic as requested
  • + Clean, bold composition with high clarity
  • The diorama base is a bit large relative to the plate
  • Materials look slightly more like clay than 'realistic PBR'

Verdict: While FLUX.1 [dev] produces a more sophisticated lighting environment and realistic textures, GPT Image 1 is the superior choice for this specific prompt due to its perfect text rendering and adherence to the '3D cartoon' style. GPT Image 1 followed all layout instructions, including the specific text and flag placement, whereas FLUX.1 [dev] struggled with spelling and clear typography.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Features soft, vibrant colors that enhance the 'wholesome' vibe.
  • + Clean composition with distinct, well-placed characters.
  • Fails the prompt's 'hyper-photorealistic' requirement, looking more like a 3D animation or digital illustration.
  • Missing the kitten entirely, and depicts one animal as a hybrid squirrel/rabbit creature.
  • Characters are posed in a static line rather than 'tumbling together'.

GPT Image 1

  • + Excellent adherence to 'hyper-photorealistic' with realistic fur textures and anatomy.
  • + Accurately includes all four requested animals: golden retriever, tabby kitten, bunny, and fox kit.
  • + Dynamic composition captures the 'chasing' and 'tumbling' action requested.
  • The fox kit has three front paws visible, indicating a minor anatomical artifact.
  • The butterfly in the top left corner is a bit large in scale compared to the animals.

Verdict: GPT Image 1 far exceeds FLUX.1 [dev] in this challenge by delivering a truly photorealistic result that includes all specified animals in a dynamic pose. FLUX.1 [dev] produced a stylized, cartoonish image that missed the kitten and lacked the requested realistic detail.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Elegant layout with a professional vector emblem feel
  • + Good use of negative space and classic typography flourishes
  • Significant spelling errors including 'Flarilaan' and 'Reseaurant'
  • Includes hallucinated numbers ('11011', '1941') not present in the prompt

GPT Image 1

  • + Flawless text rendering exactly matching the prompt including 'Caffè'
  • + Strong adherence to the texture requirement and minimalist vector style
  • + Perfectly placed banner and cloche icon
  • Interpreted 'light background' as a dark background (though the asset itself is brown/cream)
  • Minimalist design might feel slightly simple compared to traditional emblems

Verdict: GPT Image 1 is the clear winner because it correctly spells all text, including the specific accents in 'Caffè Florian' and the 'Est. 1720' banner. While FLUX.1 [dev] produced a more complex and visually interesting vintage layout, it suffered from severe typos and added nonsensical characters.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 [dev]
GPT Image 1

AI Judge Analysis

FLUX.1 [dev]

  • + Elegant layout with a professional infographic feel.
  • + Strong adherence to the flat-vector style with sophisticated color blending.
  • + Excellent visual balance and clean iconography.
  • Text consists of gibberish despite being legible.
  • Fails to follow the logical sequence of steps requested in the prompt.

GPT Image 1

  • + Excellent text rendering with correct names and labels.
  • + Successfully includes almost all requested technical steps and icons.
  • + Clean, communicative design that functions as an actual infographic.
  • Simple composition that feels slightly cluttered compared to model A.
  • Minor spelling error with 'EARLLUNAR' in the bottom label.

Verdict: While FLUX.1 [dev] produced a more aesthetically pleasing and high-quality artistic design, the text is nonsensical and the actual infographic content is poor. GPT Image 1 is the clear winner because it understood the functional requirements of the infographic, rendering readable and mostly accurate text that tells a coherent story of the mission.

Next steps

Explore each model