Head to head
Esc

Models · slot A

to navigate to pick

FLUX1.1 [pro] Ultra Black Forest Labs GPT Image 2 OpenAI

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX1.1 [pro] Ultra

17.9 arena score

#51 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 2

28.1 arena score

#3 of 62 in Text-to-Image

Top 3 in Text-to-Image
Vote tally

Where the votes landed

FLUX1.1 [pro] Ultra

0%

win rate

Ties

0%

GPT Image 2

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent photorealistic rendering of glass refraction and dust particles.
  • + Sophisticated lighting that accurately reflects the 'soft window light' prompt.
  • + High-quality text rendering on the book spine adding to the realism.
  • The sphere is floating unnaturally in the center of the cube without support.
  • The cube's proportions and perspective relative to the table feel slightly off.

GPT Image 2

  • + Natural grounding of objects with the sphere resting on the bottom of the cube.
  • + Realistic wood texture and grain on the tabletop.
  • + Logical composition with the plant and window placement.
  • The glass cube has thick, beveled edges that look more like an acrylic case than a standard glass cube.
  • The lighting is a bit flatter and less dynamic than Model A.

Verdict: Both models followed the prompt instructions perfectly, including all requested objects and lighting conditions. FLUX 1.1 [pro] Ultra created a more visually stunning and realistic image regarding textures and light, though it chose a surreal floating placement for the sphere; GPT Image 2 provided a more physically grounded and logical scene despite slightly less impressive material rendering.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent execution of long-exposure motion blur on passing vehicles.
  • + Strong cinematic atmosphere with vibrant red tones and wet reflections.
  • + High technical clarity in the subject's face and jacket texture.
  • The composition feels overly staged rather than 'candid' or 'imperfectly framed'.
  • The bicycle design is a bit simplified and generic.
  • Lacks literal interaction with tools or specific mechanical repair.

GPT Image 2

  • + Successfully captures a 'candid' street photography aesthetic with messy, realistic framing.
  • + Excellent attention to detail in the bike's mechanics and the inclusion of a tool kit.
  • + Highly realistic skin textures and authentic-looking Japanese street signage.
  • Fails to include the requested motion blur from passing cars.
  • The background cars look slightly static given the prompt's focus on motion.
  • The lighting is a bit flat compared to the cinematic request.

Verdict: FLUX1.1 [pro] Ultra excels at the cinematic and motion elements, creating a visually striking image with beautiful long-exposure effects. However, GPT Image 2 better captures the 'candid' and 'imperfect' nature of the prompt, showing a more realistic repair scene with specific tools and authentic street details, though it missed the motion blur requirement.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent attention to all prompt details including beads in hair and ornate engraving.
  • + Superior lighting and texture rendering on the armor and skin.
  • + Very lifelike, detailed eyes with realistic reflections.
  • The bokeh sparks are a bit large and might distract slightly from the main subject.

GPT Image 2

  • + Good interpretation of the braid and beads requirement.
  • + Effective lighting and atmosphere that captures the torchlight feel.
  • + Strong textures on the engraved armor plating.
  • Face appears a bit 'too perfect' and youthful for a 'battle-worn' description despite the dirt.
  • Lower overall resolution and sharpness compared to the other model.

Verdict: FLUX1.1 [pro] Ultra produced a significantly more detailed and realistic image with superior texture work on the skin, armor, and hair. While GPT Image 2 followed the prompt well, it suffered from a slightly soft focus and a character model that looked less like a battle-hardened paladin and more like a clean model with a bit of dirt applied.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent high-angle mockup perspective
  • + Clean use of white space that fits the minimalist prompt
  • + Professional font choices for the headers
  • Text is mostly gibberish or misspelled ('Pizzzans')
  • Visual layout of the food photos is cluttered and overlaps poorly
  • Inconsistent alignment between text and images

GPT Image 2

  • + Perfect text rendering with realistic dish descriptions and pricing
  • + Highly organized grid layout that strictly follows the sections requested
  • + Vibrant food photography that looks consistent and professional
  • Slightly less 'minimalist' than Model A due to the high density of information
  • Small icons for social media are clear but generic

Verdict: GPT Image 2 is the clear winner as it provides a fully functional, legible menu design with perfect spelling and a logical grid layout. While FLUX1.1 [pro] Ultra captures the 'minimalist' aesthetic well through its use of white space and mockup perspective, the text and image placement are too chaotic for a professional design. GPT Image 2 successfully incorporates vibrant accents and all requested categories with high-quality, realistic food images.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent typography with a professional, clean neon-glow effect.
  • + High-quality rendering of the fire and ember environment at the bottom.
  • + Clean, centered composition that looks like a polished commercial ad.
  • Failed to include the starburst element for the price.
  • The burger stack is repetitive with multiple patties, losing the 'exploded component' look.
  • Includes some small AI-generated gibberish text at the very bottom corners.

GPT Image 2

  • + Perfect adherence to all prompt elements, including the starburst and glowing fire text.
  • + Dynamic 'exploded' view that clearly shows every individual component.
  • + Exceptional texture on the meat, vegetables, and bun.
  • The composition is a bit crowded with the text assets overlapping the burger effects.
  • The lettuce appears slightly oversaturated compared to the other ingredients.

Verdict: GPT Image 2 is the clear winner as it followed every specific detail of the prompt, including the 'starburst' for the price and the fiery texture for the text, while FLUX1.1 [pro] Ultra missed the starburst and created a less dynamic burger stack. GPT Image 2 also captured the 'exploded' motion more effectively, showing sauce splashes and distinct component separation.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Features a very high contrast, legible image
  • + Successfully renders the requested date and most key phrases
  • Serious spelling errors and repetitive gibberish (e.g., 'mushroommm ris', 'fresh fresh', 'friistratnisnios')
  • The handwriting style is more like a digital marker than realistic chalk on a board
  • Failed to use elegant cursive for the title as requested

GPT Image 2

  • + Excellent adherence to the 'elegant cursive' title instruction
  • + Perfect spelling and completion of the menu items, even the truncated prompt for 'Brown But...'
  • + Highly realistic chalk texture with slight variations in letter size and slant
  • Slightly lower contrast making some of the bottom text a bit faint
  • The background environment is slightly more cluttered than the clean frame of the other model

Verdict: GPT Image 2 followed the prompt instructions near-perfectly, successfully intuiting the completion of the truncated 'Brown Butter' item and using elegant cursive for the header. FLUX1.1 [pro] Ultra suffered from significant 'hallucinated' text and repetitive spelling errors, failing to maintain the realistic texture of chalk.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent cinematic lighting and composition
  • + High resolution with realistic textures on the horse's mane and space suit
  • + Dynamic action pose with a beautiful planetary background
  • Completely failed the semantic instruction 'horse on top'
  • Interpretated 'horse riding astronaut' as the standard astronaut riding a horse

GPT Image 2

  • + Correctly followed the specific and difficult instruction of the horse being on top
  • + Captures the surreal nature requested in the prompt
  • + Good texture on the moon's surface and space suit fabric
  • Anatomical issues where the person's hands/gloves merge into the ground
  • The horse's legs are awkwardly clipped or merged with the astronaut's shoulders
  • Composition is a bit static compared to the alternative

Verdict: While FLUX1.1 [pro] Ultra produced a much more visually stunning and cinematic image, it failed the core logical constraint of placing the horse on top of the astronaut. GPT Image 2 successfully followed the complex instruction to reverse the roles, making it the superior choice for prompt adherence despite lower aesthetic quality and some anatomical artifacts.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + High resolution and crisp texture on the capybara's fur.
  • + Consistent lighting and vibrant bokeh in the background.
  • The capybara and the woman are sitting in the same seat, creating a major anatomical and spatial error.
  • The woman is holding the steering wheel instead of being in the back seat.
  • The capybara's head looks pasted onto a human body.

GPT Image 2

  • + Perfect adherence to the prompt's spatial instructions, with the capybara in the driver's seat and the woman in the back.
  • + The capybara's paws are correctly placed on the steering wheel.
  • + Realistic taxi interior and cinematic atmosphere with rain droplets on the window.
  • Lower resolution and slightly more grain compared to Image A.
  • The woman's face in the background is slightly blurry.

Verdict: While FLUX1.1 [pro] Ultra has higher image clarity, it completely fails the prompt's logic by merging the driver and passenger into the same seat. GPT Image 2 follows every instruction perfectly, placing the businesswoman in the back seat and the capybara at the wheel with correct anatomy and composition.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent typography rendering for the specific event details
  • + Crisp, clean illustration style
  • + Adheres well to the dark parchment and twisted trees requirement
  • Includes redundant and misspelled text ('noickt or a night of frights')
  • Failed to include the word 'Party' in the main title
  • The layout feels slightly clinical/modern rather than 'vintage gothic'

GPT Image 2

  • + Perfectly captures the vintage gothic aesthetic with dark, atmospheric coloring
  • + Flawless text rendering of the full title and scroll banner
  • + Highly creative composition including a city skyline and arches to match the location name
  • The date and location text at the bottom is slightly smaller and harder to read against the busy background

Verdict: GPT Image 2 is the superior output as it perfectly captures the desired vintage gothic atmosphere while fulfilling every text requirement accurately. FLUX1.1 [pro] Ultra suffered from significant text hallucinations, repeating the invite phrase with misspellings and omitting the word 'Party' from the main title.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Elegant soft-lighting and refined PBR texture rendering.
  • + Excellent integration of text into the background space.
  • + Sophisticated 3D cartoon aesthetic that feels professional and cohesive.
  • Includes a strange, non-requested ladybug artifact on the base.
  • The steam rising from the sushi is physically inaccurate for cold fish.

GPT Image 2

  • + Extremely clear and bold text rendering that perfectly matches the prompt.
  • + Higher detail in the food textures like the rice grains and salmon marbling.
  • + Strong adherence to the 'miniature diorama' request with the stone lantern and garden elements.
  • Lighting is a bit flat and lacks the 'gentle' quality requested.
  • Shadows under the text are a bit harsh, feeling less integrated than the other model.

Verdict: Both models followed the prompt exceptionally well, but GPT Image 2 edges ahead for its superior literal interpretation of 'miniature diorama' and cleaner text execution. While FLUX1.1 [pro] Ultra has a more sophisticated artistic lighting style, GPT Image 2 provided a more complete sushi variety and a better-realized base.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent soft lighting and consistent artistic style
  • + Superb fur detail and expressive, uniform eye styles
  • + Clean, dreamy composition with a high quality 'masterpiece' feel
  • Static posing that feels more like a portrait than 'tumbling and chasing'
  • Creatures look slightly stylized/Disney-like rather than hyper-photorealistic
  • The kitten is golden rather than a distinct tabby

GPT Image 2

  • + Captures the action of 'tumbling and chasing' much more effectively
  • + Correctly identifies specific species details like the tabby markings and the fox's dark paws
  • + Excellent interpretation of 'god rays' and sunburst backlighting
  • The fox kit has a slightly distorted, elongated front leg artifact
  • Lower resolution appearance in some background elements compared to Model A
  • The butterfly wings are a bit thick and less delicate looking

Verdict: While FLUX1.1 [pro] Ultra produces a cleaner, more aesthetically perfect 'portrait' style image with beautiful lighting, GPT Image 2 (DALL-E 3) followed the prompt instructions for action and species variety more accurately. GPT Image 2 successfully depicted a tabby kitten and captured the dynamic motion of the animals tumbling, whereas FLUX1.1 [pro] Ultra rendered a static group of very similar-looking golden animals.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent vector emblem style execution
  • + Clean, bold color palette
  • + High contrast and sharp lines
  • Serious spelling error in the main brand name
  • Banner placement obscures the cloche somewhat

GPT Image 2

  • + Perfect text rendering for all requested strings
  • + Beautiful vintage stippling and texture
  • + Strong composition with a balanced frame
  • Slightly less 'minimalist' than requested due to ornate detailing
  • The cloche handle is slightly off-center

Verdict: While FLUX1.1 [pro] Ultra produced a very clean vector style, it failed significantly on the primary prompt requirement by misspelling the restaurant name. GPT Image 2 followed all instructions perfectly, providing accurate text, a lovely vintage texture, and a superior layout, making it the clear winner.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX1.1 [pro] Ultra
GPT Image 2

AI Judge Analysis

FLUX1.1 [pro] Ultra

  • + Excellent flat-vector aesthetic that matches the 'modern infographic' request
  • + Creative layout using the moon's surface as a horizon for the data points
  • + Accurate NASA-inspired color palette
  • Text is largely illegible gibberish
  • Information flow is confusing and non-linear compared to a standard infographic
  • Fails to clearly delineate the requested 6 steps

GPT Image 2

  • + Follows the 1-6 step sequence exactly with logical icons for each
  • + Text and labeling are clear, legible, and accurate (e.g., Crew names)
  • + Stronger adherence to specific infographic structure and layout
  • Visual style leans more towards realistic illustration than 'flat-vector'
  • Composition is slightly crowded at the top header area
  • The Eagle patch representation is a bit messy in details

Verdict: GPT Image 2 followed the prompt's structural instructions much better, providing a clear 1-6 step process with legible text and accurate data. While FLUX1.1 [pro] Ultra captured a superior 'flat-vector' artistic style, it failed significantly on readability and the logical progression of the mission steps.

Next steps

Explore each model