Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [flex] Black Forest Labs GPT Image 1 Mini OpenAI

Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.

FLUX.2 [flex]

24.8 arena score

#14 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1 Mini

25.0 arena score

#13 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [flex]

66.7%

win rate

Ties

33.3%

GPT Image 1 Mini

0.0%

win rate

66.7% 33.3% ties 0.0%
Shared challenges 19

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [flex]
GPT Image 1 Mini
50% wins 50% ties 0% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the glass cube geometry with sharp, realistic edges.
  • + Shows clear refraction of the plant through the glass panels as requested.
  • + Effective use of bright window light creating realistic highlights and shadows.
  • The blue sphere appears slightly too large relative to the 'small sphere' prompt.
  • The plant's leaves intersect the back of the cube in a way that looks slightly merged rather than behind it.

GPT Image 1 Mini

  • + The blue sphere is levitating, which creates a striking and clean visual center.
  • + The texture on the red book cover and pages is highly detailed and realistic.
  • + Superior lighting and color palette with a warm, professional photographic feel.
  • Failed to make the plant visible through the glass; it is entirely blocked by the solid red book and the frosted-looking pot.
  • The glass cube is missing its top face where the book sits, making it more of an open frame.

Verdict: FLUX.2 [flex] adhered better to the spatial requirements of the prompt, successfully showing the plant through the glass cube. GPT Image 1 Mini produced a more aesthetically pleasing image with better textures, but it failed to show the plant through the glass and the cube structure is physically inconsistent.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of the specific white car model and its design details.
  • + Subtle but accurate hair style preservation for the subject.
  • The man is not looking at the road, staring off into the distance instead.
  • Changed the character's outfit from a plaid coat to a simple white shirt.

GPT Image 1 Mini

  • + Perfectly preserved the man's plaid coat, scarf, and facial expression.
  • + Dynamic and engaging composition with better integration of the subject and coastal background.
  • The car model was changed significantly from the source image, particularly the headlights and front end.
  • The scale of the car relative to the man feels slightly off.

Verdict: FLUX.2 [flex] was more successful at preserving the specific vehicle from the source image but failed to maintain the man's clothing. GPT Image 1 Mini captured the man's likeness and outfit perfectly but significantly altered the design of the car. GPT Image 1 Mini is the likely winner as it feels like a more natural and successful composite for a travel-style photo, despite the slight change to the car's front end.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [flex]
GPT Image 1 Mini
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to technical prompts like motion blur from passing cars and light rain visibility.
  • + Very high detail in skin texture and clothing materials.
  • + Strong bokeh effect and realistic urban lighting that adds to the cinematic feel.
  • The bicycle frame geometry is physically nonsensical, with tubes merging incorrectly.
  • The man's hands are mangled and blending into the bicycle chain area.

GPT Image 1 Mini

  • + The bicycle has a much more coherent and realistic structure compared to Model A.
  • + The lighting and color palette feel more grounded and less digitally enhanced.
  • + Natural candid composition that feels like a real street photograph.
  • Failed to include the requested motion blur from passing cars (background is static).
  • Missing the 'light rain' visual cues like raindrops or mist, though pavement is wet.

Verdict: FLUX.2 [flex] wins on technical prompt adherence, successfully incorporating the requested motion blur and rain effects with striking clarity, despite significant structural errors in the bicycle and hands. GPT Image 1 Mini produces a more physically plausible bicycle and a more authentic 'candid' feel, but it missed the specific environmental motion requirement and has lower detail in the skin textures.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the 'beads in hair' prompt
  • + High contrast and sharp lighting on the metal textures
  • + Superior rendering of leather straps and cloth underlayer
  • The facial scars look a bit like face paint or superficial decals
  • The background torches are a bit distracting compared to the subject

GPT Image 1 Mini

  • + Natural and gritty skin texture reflecting the 'battle-worn' theme
  • + Consistent warm tonal palette with realistic lighting
  • + Very fine engravings on the plate armor
  • Missing the specific request for beads in the hair
  • Slightly less clarity on the leather and cloth details compared to the other model

Verdict: FLUX.2 [flex] successfully captured more specific details from the prompt, particularly the beads in the hair and the distinct texture of the underlayer cloth. GPT Image 1 Mini provided a more atmospheric and gritty portrait but failed to include the requested hair ornaments.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent professional layout with pricing and descriptions
  • + High-quality, realistic food photography
  • + Clean typography with clear section hierarchy
  • Text under headers is mostly nonsense/placeholder Latin-style text
  • Section headers are slightly confused (pizza listed under appetizers)

GPT Image 1 Mini

  • + Perfectly clean grid layout
  • + Clear section categorization
  • + Simple and highly legible bold fonts
  • Missing all menu details like dish names, descriptions, and prices
  • The layout feels like a blank template rather than a finished design
  • Minimalist to the point of being unfinished

Verdict: FLUX.2 [flex] produced a much more realistic and professional-looking menu that includes pricing, dish names, and realistic lighting on the food, despite some of the text being gibberish. GPT Image 1 Mini created a very clean grid, but it lacks the essential content of a menu (dish names and prices), appearing more like a placeholder template.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent photorealistic texture on the meat and bun.
  • + Vibrant and dynamic composition with realistic sauce drips and crumbs.
  • + Clear, professional integration of all requested text elements, including the starburst.
  • The 'MAGIC BURGER' text is slightly clipped at the top edges.

GPT Image 1 Mini

  • + Strong fiery glow effect on all text elements.
  • + Clean separation of burger components.
  • Lighting on the burger is much flatter and less appetizing than the other model.
  • Missing the sense of motion and 'exploded' energy requested in the prompt.
  • The background lacks the requested fiery intensity, appearing mostly dark with small sparks.

Verdict: FLUX.2 [flex] produced a significantly higher quality image that looks like a professional advertisement, featuring superior lighting, textures, and a dynamic 'exploded' layout. GPT Image 1 Mini followed the text instructions well but failed to capture the photorealistic detail and energetic motion requested for the burger itself.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent chalk texture with realistic smudging and dusting on the board.
  • + Strong adherence to the elegant cursive handwriting style requested.
  • + Perfect completion of the cut-off prompt item ('Brown Butter Chocolate Chip Cookies').
  • The 's' in 'Today's' is somewhat cramped compared to the other lettering.

GPT Image 1 Mini

  • + Text is extremely legible and centered well.
  • + Successfully completed the full text of the third menu item.
  • + Good rendering of the grain in the wooden frame.
  • Failed to provide 'elegant cursive' for the title, using a print-style font instead.
  • The chalk effect looks a bit like a digital overlay rather than physical chalk.
  • The handwriting looks too uniform and lacks the 'natural variations' requested.

Verdict: FLUX.2 [flex] is the clear winner as it perfectly captured the 'elegant cursive' requirement and provided a much more realistic chalk texture including authentic smudges and varied pressure. GPT Image 1 Mini produced very clean and readable text, but failed on the specific stylistic request for cursive handwriting and looked more like a digital font.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent character preservation, maintaining the specific face, hair, and scarf pattern from Image 2.
  • + Captures the lighting and environment of Image 1 very accurately.
  • + Successfully integrates the clothing style and specific logos/text from the reference material.
  • The pose is moderately modified, lacking the extreme twist and lean of the torso from Image 1.
  • The anatomy of the feet and their positioning on the red block is slightly awkward.

GPT Image 1 Mini

  • + Successfully replicates the yellow studio background and red ottoman from Image 1.
  • + Retains the basic clothing elements like the checkered scarf and sunglasses.
  • Fails significantly on pose adherence, opting for a standard crouch rather than the dynamic, twisted pose from Image 1.
  • Loss of character detail, especially the specific facial features and the original scarf's floral pattern.
  • Anatomy of the feet is poorly rendered and lacks detail.

Verdict: FLUX.2 [flex] is the clear winner because it successfully maintained the character's unique identity, including his specific face, hair, and the complex pattern on his scarf. While neither model perfectly replicated the extreme, physics-defying pose of the first image, FLUX.2 [flex] came closer to the intended composition while GPT Image 1 Mini produced a generic character that lost most of the source material's detail.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the specific 'horse on top' spatial instruction
  • + Vibrant colors and high-quality cinematic lighting
  • + Creative interpretation of a surreal concept with the horse holding the astronaut
  • The horse's front legs/hooves morph awkwardly into hands to hold the astronaut

GPT Image 1 Mini

  • + Natural and balanced composition for a standard concept
  • + Realistic textures on the space suit and horse hide
  • Failed the primary prompt instruction to have the horse on top
  • Lacks the surreal quality requested by the prompt
  • Much darker and less cinematic than Model A

Verdict: FLUX.2 [flex] successfully followed the difficult spatial instruction to place the horse on top of the astronaut, resulting in a truly surreal and creative image. GPT Image 1 Mini ignored the specific 'not vice versa' constraint and produced a standard, uninspired image of an astronaut riding a horse. FLUX.2 [flex] is the clear winner for its superior prompt adherence and vibrant visual style.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent photorealism with sharp textures on the capybara's fur and the taxi interior.
  • + Well-lit composition that captures both the interior and the vibrant New York night life clearly.
  • + Follows all prompt instructions including the specific 'calm, professional expression' and 'paws on the steering wheel'.
  • The transition between the capybara's neck and the jacket looks slightly unnatural.
  • The hands/paws on the steering wheel have a slightly human-like anatomical structure that is a bit uncanny.

GPT Image 1 Mini

  • + Natural integration of the capybara's head into the dark jacket and hat.
  • + Captures the 'bored' expression of the businesswoman effectively.
  • + Moody, cinematic lighting that feels realistic for a night scene.
  • The passenger's face is quite blurry and lower in quality compared to the driver.
  • The 'paws on the steering wheel' are partially obscured or less defined than in Model A.
  • Only one paw is clearly visible on the steering wheel, missing the 'both front paws' instruction.

Verdict: FLUX.2 [flex] is the clear winner as it provides a much more detailed and vibrant image that follows the prompt instructions precisely, including showing both paws on the steering wheel. While GPT Image 1 Mini has good atmospheric lighting, it suffers from a lack of clarity in the background passenger and fails to fully illustrate the specified interaction with the steering wheel.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent typography with a mix of elegant gothic and serif fonts
  • + High-quality central jack-o-lantern with realistic textures and lighting
  • + Matches the invitation aesthetic perfectly with a visible parchment border
  • The parchment edges look slightly more digital/artificial than aged
  • Typography layout is a bit scattered in terms of vertical spacing

GPT Image 1 Mini

  • + Atmospheric dark color palette captures the 'moody night sky' and 'dark parchment' brief well
  • + Clean and readable center-aligned typography
  • + Intricate thorny border matches the gothic theme
  • Failed to include the specific labels 'Date:', 'Time:', and 'Location:' requested in the prompt
  • The jack-o-lantern carving is somewhat generic compared to Image A
  • The text on the scroll banner is less 'elegant' and looks like a standard sans-serif font

Verdict: FLUX.2 [flex] followed the prompt more precisely, including the specific text labels and providing a more polished, professional invitation layout. While GPT Image 1 Mini captured the atmosphere of a dark parchment very well, it omitted requested labels for the event details. FLUX.2 [flex] is the preferred choice for a functional and aesthetically pleasing invitation.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [flex]
Before After
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent source preservation of the face, glasses, and clothing
  • + Realistic texture and density for a buzz-cut style
  • + Natural integration with the existing beard and ear placement
  • The hairline is a bit too straight and blunt across the forehead

GPT Image 1 Mini

  • + High volume and length as requested for a 'full' head of hair
  • + Matches the texture and color of the beard very well
  • Significantly alters the facial structure, making the man look younger and like a different person
  • Slightly softens the overall image detail compared to the source

Verdict: FLUX.2 [flex] successfully adds hair while maintaining the identity and facial features of the original subject perfectly. In contrast, GPT Image 1 Mini provides a thicker head of hair but fails the preservation requirement, fundamentally changing the person's face to the point of being unrecognizable as the same individual.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to typographic layout with 'JAPAN' and 'SUSHI' on separate lines.
  • + Beautiful soft lighting and realistic clay-like PBR textures.
  • + Includes a wide variety of sushi types (nigiri and maki) and garnishes.
  • The spacing between the text and the flag icon is quite large, creating a slightly disconnected composition.

GPT Image 1 Mini

  • + Very clean and vibrant 3D cartoon aesthetic.
  • + The wooden diorama base provides a nice material contrast to the plate.
  • + Good centered composition and balanced color palette.
  • Failed the layout instruction by placing 'SUSHI' and the flag on the same line rather than 'SUSHI' below 'JAPAN'.
  • Added chopsticks which were not requested in the prompt.

Verdict: FLUX.2 [flex] adhered more strictly to the complex layout instructions, correctly placing the text and flag in a vertical stack at the top center. While GPT Image 1 Mini produced a very high-quality image, it deviated from the text placement prompt and added extra elements like chopsticks.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Expertly incorporates all requested elements into a cohesive news studio setting.
  • + Captures the subject's likness well in a vibrant vector caricature style.
  • + High level of detail with multiple dogs in hockey jerseys and a professional broadcast desk.
  • The style is more 'cartoon' than a traditional hand-drawn caricature.
  • The text on the news desk has minor spelling artifacts.

GPT Image 1 Mini

  • + Successfully uses a hand-drawn pencil/crayon caricature style which feels more authentic to the request.
  • + Maintains character likeness while exaggerating facial features for a humorous effect.
  • + The layout is clean and clearly includes the dog, hockey stick, and puck.
  • The hockey stick is merged awkwardly with the dog's side/back.
  • Less ambitious in terms of the environment compared to Model A.

Verdict: Both models followed the complex instructions very well, correctly identifying the subject's face, clothing, and the three required themes (TV anchor, dogs, hockey). FLUX.2 [flex] created a much more immersive and detailed scene with a full studio backdrop, while GPT Image 1 Mini captured a more traditional, humorous caricature art style with better facial exaggeration. FLUX.2 [flex] is the winner for its superior composition, creativity, and the charming addition of multiple dogs in team jerseys.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to lighting requests with clear 'god rays' and a defined sunrise background.
  • + Highly detailed environment with distinct dew sparkles on the grass and flowers.
  • + Dynamic and balanced composition that captures a sense of 'tumbling' and interaction.
  • The kitten's front paws look slightly unnatural in their mid-air position.
  • The fox's tail is exceptionally long and fluffy, bordering on stylized rather than photorealistic.

GPT Image 1 Mini

  • + Beautifully soft and realistic fur textures on all four animals.
  • + Great sense of motion with multiple animals captured in mid-leap.
  • + Consistent and warm color palette that reinforces the joyful vibe.
  • The 'god rays' are less defined compared to Model A, appearing more as a general glow.
  • The background is very blurred, losing the 'lush wildflower meadow' detail requested in the prompt.

Verdict: Both models followed the prompt exceptionally well, including all four specific animals and the golden hour atmosphere. FLUX.2 [flex] wins slightly due to its superior environmental details, particularly the dew sparkles and the more dramatic 'god rays' effect, while GPT Image 1 Mini feels a bit more like a studio portrait with a blurred background.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Perfectly captures the Studio Ghibli anime aesthetic with clean line art and watercolor-style coloring.
  • + Maintains the structural composition and specific clothing patterns of the original image.
  • + Successfully uses soft pastel tones and gentle lighting as requested.
  • The girlfriend's expression is changed from angry to smiling, losing the narrative of the original meme.

GPT Image 1 Mini

  • + Retains the original emotional expressions, including the girlfriend's facial anger.
  • + Applies a textured, hand-drawn look that feels nostalgic.
  • + Preserves the spatial arrangement of the source image well.
  • The style leans more toward generic colored pencil illustration rather than the specific 'Studio Ghibli' anime aesthetic.
  • The colors are a bit saturated and muddy rather than dreamy and pastel.

Verdict: FLUX.2 [flex] delivered a much more accurate 'Studio Ghibli' style, featuring the characteristic clean line work and ethereal background lighting associated with the studio, though it failed to preserve the girlfriend's angry expression. GPT Image 1 Mini preserved the emotional context of the scene better but failed to replicate the specific aesthetic requested, resulting in a colored pencil texture that doesn't feel like a Ghibli illustration.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [flex]
Before After
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent source preservation, maintaining the woman's face and the dog's features accurately.
  • + Successfully added wind-blown hair and green leaves that match the existing trees.
  • The dog's left leg and the woman's right hand exhibit slight warping artifacts.
  • The added leaves lack motion blur, appearing as static floating objects.

GPT Image 1 Mini

  • + The leaves have various angles and sizes, creating a better sense of depth and movement.
  • + The hair flow feels natural and dynamic across the shoulders.
  • Significantly altered the woman's facial features and the appearance of the golden retriever.
  • Modified the background scenery, including the bridge and trees, failing to preserve the source image's identity.

Verdict: FLUX.2 [flex] successfully followed the edit instructions while keeping the identity of the person and dog intact, which is critical for image editing. GPT Image 1 Mini provided a more 'energetic' feel with the leaves, but it failed significantly at source preservation, essentially generating a new image that looks like the original but with different faces and background details.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent typography rendering with the correct accent on 'Caffè'
  • + Superior minimalist vector aesthetic that perfectly aligns with modern branding standards
  • + Accurate adherence to the light background and cream/brown tone request
  • The 'Est. 1720' text is slightly off-center within the banner

GPT Image 1 Mini

  • + Strong vintage texture and gold-leaf style appearance
  • + Good use of classic serif typography and layout
  • Failed the requirement for a 'light background' by using a solid black background
  • The steam effect and cloche shading are a bit messy compared to a clean vector emblem

Verdict: FLUX.2 [flex] followed all prompt instructions including the specific color palette and background requirements, resulting in a professional and clean vector logo. GPT Image 1 Mini ignored the request for a light background and subtle texture, opting for a high-contrast black and gold style that feels more like a sign than a minimalist emblem.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [flex]
GPT Image 1 Mini

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent layout with a professional, dark theme that feels modern and high-quality.
  • + Superior text rendering with clear headers and supportive subtext.
  • + Sophisticated vector illustrations with nice use of depth and lighting.
  • Missed the final 'Landing' step requested in the prompt, stopping at 'Lunar Orbit'.
  • The flow of information is slightly fragmented by the central divider.

GPT Image 1 Mini

  • + Followed the step-by-step instructions perfectly, including all six stages from Launch to Landing.
  • + Strong adherence to the requested NASA-inspired color palette and flat-vector style.
  • + Clear, numbered sequential flow that makes the infographic easy to read.
  • The 'Translunar' graphic is a confusing scribbled line that doesn't accurately represent a trajectory.
  • Visuals are a bit more simplistic and 'cartoony' compared to the polished look of the competitor.

Verdict: While FLUX.2 [flex] produced a much more visually stunning and professional-looking poster, it failed to include critical steps requested in the prompt (Step 6: Landing). GPT Image 1 Mini adhered strictly to the requested structure and content sequence, despite having slightly less refined artwork. GPT Image 1 Mini is the winner for its superior prompt adherence regarding the specific infographic stages.

Next steps

Explore each model