Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [max] Black Forest Labs GPT Image 1 OpenAI

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

FLUX.2 [max]

25.9 arena score

#10 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1

23.2 arena score

#28 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [max]

0%

win rate

Ties

0%

GPT Image 1

0%

win rate

Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent rendering of photorealistic glass physics, including reflections and refraction.
  • + Very high detail in textures, particularly the leather-like cover of the book.
  • + Superior interpretation of 'soft light from the left' with realistic shadow patterns on the table.
  • The plant is extremely out of focus, making it less distinct as a background element.
  • The internal reflection of the blue sphere on the right face of the cube looks slightly confusing.

GPT Image 1

  • + Clean composition with a clear view of the green plant through the glass as requested.
  • + Simple, effective color contrast between the primary red, blue, and green elements.
  • + The blue sphere is perfectly centered and distinct.
  • The cube lacks a glass 'floor', making the sphere appear to hover or rest directly on a metal plate inside the frame.
  • The lighting feels a bit flat and less directional compared to the prompt's request for window light.

Verdict: Both models followed the prompt's spatial instructions perfectly. FLUX.2 [max] is the winner due to its superior photorealistic rendering of glass, light, and texture, whereas GPT Image 1 feels slightly more like a 3D render with less sophisticated light physics.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Successfully preserved the specific car model from the source image
  • + High prompt adherence with the California coastline and background detail
  • + Maintains the man's scarf and plaid coat texture
  • The man appears to be sitting in the passenger seat instead of driving
  • The man's hair texture is slightly altered and less defined than the source

GPT Image 1

  • + Correctly positions the man in the driver's seat
  • + Excellent motion blur on the wheels and road adds a sense of driving
  • + Highly realistic face preservation and lighting
  • Changed the car's wheel rims from the source image
  • Distorted reflections on the car's body panels near the door handle

Verdict: Both models handled the complex task of merging two source images into a new environment well. While FLUX.2 preserves the integrity of the car's physical details (like the rims) more accurately, GPT Image 1 is the superior edit because it correctly identifies the driver's side and creates a dynamic sense of motion that fits the prompt perfectly.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the 'motion blur from passing cars' prompt requirement
  • + Highly realistic textures on the bicycle components and wet pavement
  • + Accurate representation of an 'imperfect' street photography frame
  • The man's skin texture appears slightly over-sharpened compared to the soft rain environment

GPT Image 1

  • + Natural and convincing skin textures on the man's face
  • + Beautiful shallow depth of field and color grading
  • + Expressive subject with a very candid, realistic feel
  • Fails to include the requested 'motion blur from passing cars', as cars are static and sharp-edged
  • The bicycle geometry is slightly warped near the chain guard

Verdict: FLUX.2 [max] provides a more complete adherence to the technical prompts, successfully incorporating motion blur and a chaotic street environment. While GPT Image 1 has a very soulful subject and great skin tones, it misses key prompt elements like the car motion blur and feels less like a 50mm candid street shot.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent engraving details and realistic texture on the metal plate armor.
  • + Superior skin texture showcasing individual pores, realistic scars, and dirt.
  • + Strong adherence to the 'hair braided with small beads' requirement with clear visual definition.
  • The sparks in the bokeh are a bit sparse compared to the requested atmosphere.
  • The lighting on the face feels slightly flat despite the torchlight prompt.

GPT Image 1

  • + Excellent mood and lighting with deep shadows and warm, believable torchlight reflections.
  • + Highly effective use of bokeh sparks that enhance the battle-worn atmosphere.
  • + Good expression and character features that convey the 'battle-worn' theme well.
  • The beads in the hair are less distinct and blend into the hair texture.
  • The engraving on the armor is less sharp and detailed compared to FLUX.2.

Verdict: FLUX.2 [max] provides significantly better technical detail and texture, particularly in the armor engravings and skin fidelity, making it appear more 'lifelike'. However, GPT Image 1 captures the cinematic mood and warm torchlight atmosphere much more effectively, although it falls short on the fine details like the beads in the hair.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent grid layout following professional design standards.
  • + Clean use of color-coded headers for sections.
  • + High consistency in typography and alignment.
  • The placeholder text is mostly gibberish.
  • The food images are small and some look more like generic stock photos than high-quality menu photography.

GPT Image 1

  • + Features large, high-quality, appetising food photography.
  • + Text is mostly legible with fewer spelling errors than Model A.
  • + Stronger 'minimalist' aesthetic that feels more contemporary.
  • Missing the 'Mains' section header entirely, resulting in poor sectioning.
  • The grid layout of the text is slightly less organized than Model A.
  • Redundant text such as 'Appetizer Name' and repetitive descriptions.

Verdict: FLUX.2 [max] creates a more structurally complete and professional menu layout with clear sections for Appetizers, Pizza, and Mains as requested, though the text is garbled. GPT Image 1 produces much better food photography and legible text, but it fails to include the requested 'Mains' section and has a less sophisticated graphic design. FLUX.2 is the better design tool for this task due to its superior layout composition.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent typography with a realistic fiery texture.
  • + Crisp details in the burger ingredients and flying particles.
  • + Perfect adherence to all text requirements including the specific price.
  • The starburst for the price is a flat graphic that doesn't quite match the photographic lighting of the burger.

GPT Image 1

  • + Strong vertical composition with a great sense of depth.
  • + Fiery starburst effect matches the overall aesthetic of the image background.
  • + Rich, saturated colors contribute to a warm, appetizing feel.
  • Failed to include the '6' in the price, rendering it as '€.99'.
  • The top bun is slightly cropped out of the frame.

Verdict: FLUX.2 [max] is the winner as it correctly rendered all text elements requested, whereas GPT Image 1 missed a digit in the price. While GPT Image 1 had a more integrated 'fiery' style for the starburst, FLUX.2 [max] provided better clarity and a more professional layout for a commercial advertisement.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent chalk texture with realistic smudges and dust on the board.
  • + Perfect text rendering of all requested items including the logical completion of the cutoff prompt.
  • + Sophisticated lighting and café atmosphere within the composition.
  • The 'cursive' style in the title is more of a stylized script than traditional elegant cursive.

GPT Image 1

  • + Natural variation in letter sizing and spacing.
  • + Clear, legible chalk-style text.
  • The chalk texture looks a bit too digital/uniform compared to the more gritty, realistic look of the competitor.
  • Missing the pricing symbol highlights ($) on the third item.
  • Overall composition is tighter and less atmospheric than the café setting requested.

Verdict: FLUX.2 [max] is the clear winner as it perfectly captured the full text, including the cut-off item, whereas GPT Image 1 missed the dollar sign on the final item. FLUX.2 [max] also provided a much more realistic chalk texture and a more immersive 'cozy café' atmosphere through lighting and board framing.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the character reference, maintaining facial features, hair, and clothing details accurately.
  • + Perfect recreation of the complex pose and leg-crossing from the source image.
  • + High visual quality with realistic skin textures and lighting integration.
  • The right hand lacks the delicate 'dancer' finger extension seen in the source pose.
  • The shirt color was changed to black to match the character reference, which is accurate to the prompt but loses the vibrant contrast of the original red.

GPT Image 1

  • + Successfully captures the general pose and the yellow background from the source image.
  • + Maintains the key accessories like the sunglasses and scarf.
  • The face and hair bear little resemblance to the character reference in Image 2.
  • Significant anatomical issues, particularly with the distorted right arm and hands.
  • Poor image clarity and a lack of detail in the clothing textures.

Verdict: FLUX.2 [max] performed exceptionally well, accurately transplanting the character's facial features and specific outfit onto the complex pose from the source image while maintaining photorealism. GPT Image 1 struggled significantly with character consistency and suffered from major anatomical distortions in the limbs and hands. FLUX.2 [max] is the clear winner for its technical accuracy and superior visual quality.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent anatomical details on both the horse and the spacesuit.
  • + Stunning background with a highly detailed galaxy and atmospheric asteroid.
  • + High-quality lighting with cinematic lens flares and highlights.
  • The astronaut is sitting a bit too far back on the horse's spine.
  • The pose is somewhat static compared to the other model.

GPT Image 1

  • + Dynamic action pose with the horse galloping through space.
  • + Strong color palette with a moody, vintage cinematic feel.
  • + Good composition with the horse angled diagonally across the frame.
  • The horse's hind legs are anatomically confused and appear distorted.
  • The background is relatively empty and lacks the surreal detail of Model A.
  • Texture on the spacesuit looks slightly muddy compared to the crispness of Model A.

Verdict: FLUX.2 [max] produced a much higher quality image with superior detail and a more impressive cosmic background. While GPT Image 1 captured a more dynamic 'moving' pose, the anatomical errors in the horse's legs and the lower overall resolution make it the weaker choice compared to the polished output of FLUX.2 [max].

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the clothing details, including the white label on the scarf and the gold ring.
  • + Perfectly preserved the sand on the subject's face and the specific hair vitiligo pattern.
  • + Maintained the color grading and lighting of the background more faithfully.

GPT Image 1

  • + Captured the general structure of the outfit well.
  • + Lighting on the face is soft and aesthetically pleasing.
  • Failed to preserve face and hair details, notably removing fine sand and changing the shape of the vitiligo patch on the forehead.
  • Missing several accessories from Image 2 such as the ring and the specific tag on the scarf.
  • The background beach textures appear smoothed and lose the natural debris from the original photo.

Verdict: FLUX.2 [max] is the clear winner as it successfully transferred every minor detail of the outfit from Image 2 (including the ring, watch, and scarf tag) while keeping Image 1's subject and background virtually identical. GPT Image 1 struggled with source preservation, altering the subject's unique skin patterns and omitting several requested accessories.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent photographic quality and lighting depth.
  • + Highly detailed New York taxi interior and dashboard.
  • + Great rendering of the passenger's face and phone glow.
  • The capybara has realistic human hands instead of front paws.
  • The human hands are disproportionately large compared to the steering wheel.

GPT Image 1

  • + Correctly depicts capybara paws on the steering wheel.
  • + Clearer text on the taxi cap.
  • + Stronger adherence to the 'bored' expression of the passenger.
  • Composition feels slightly more cramped than the other image.
  • The lighting on the capybara's fur is a bit flat compared to the realism of the competitor.

Verdict: While FLUX.2 [max] creates a more visually stunning and realistic photograph with complex interior lighting, it fails a major anatomical prompt requirement by giving the capybara human hands. GPT Image 1 follows the prompt much more accurately by depicting paws on the steering wheel, despite having slightly lower overall texture detail. GPT Image 1 is the winner for its superior prompt adherence regarding the primary subject's anatomy.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent typography with perfect spelling and clear hierarchy.
  • + High visual clarity and atmospheric cinematic lighting.
  • + Accurate representation of all prompt elements, including the scroll and border details.
  • The thorns in the border are slightly repetitive.

GPT Image 1

  • + Strong 'vintage' aesthetic with a darker, moodier color palette.
  • + Good integration of the scroll banner into the upper composition.
  • Text error in the event details, merging 'Time' and 'Location' into one line incorrectly.
  • Lower resolution feel with softer, less defined edges on the jack-o-lantern and trees.

Verdict: FLUX.2 [max] followed the prompt instructions perfectly, resulting in a professional-grade invitation with flawless text and lighting. GPT Image 1 captured a moodier aesthetic but ultimately failed on the specific event details, merging the time and location info and producing a slightly blurrier image.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [max]
Before After
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent source preservation, maintaining identical facial features and background.
  • + Realistic hair texture and natural integration with the existing sideburns.
  • + Very natural-looking hairline and volume.
  • None notable.

GPT Image 1

  • + Successfully added a full head of hair as requested.
  • + Good texture on the top of the hair.
  • Subtle facial morphing makes the subject look slightly like a different person.
  • The hair volume is unnaturally high, looking almost like a wig or an overlay.
  • The integration between the new hair and the existing beard/temple area is slightly messy.

Verdict: FLUX.2 [max] performed a perfect edit by seamlessly integrating a realistic hairstyle while keeping the man's identity and the surrounding environment exactly the same as the source. GPT Image 1 successfully added hair but slightly altered the man's facial structure and proportions, resulting in a less realistic and less accurate modification.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Perfectly captures the requested isometric perspective and diorama style.
  • + Very clean, professional typography and layout that aligns with the 'miniature' aesthetic.
  • + Highly detailed 3D assets that look like high-quality clay or plastic renders.
  • The text is slightly off-center relative to the composition's vertical axis.

GPT Image 1

  • + Appealing soft, clay-like textures on the sushi and garnish.
  • + Correct text placement and inclusion of the flag icon in a central column.
  • The camera angle is a standard perspective shot rather than the requested 45° isometric view.
  • The flag icon is stylized as a square/box rather than a standard flag flag format.
  • The composition feels slightly crowded compared to Model A's diorama style.

Verdict: FLUX.2 [max] followed the prompt more accurately by providing a true 45° isometric view and a distinct diorama base, creating a more professional 'miniature' look. GPT Image 1 produced higher-quality soft textures on the food items but failed to capture the requested isometric perspective, resulting in a more standard 3D close-up.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent integration of all themes in a cohesive setting (hockey rink news desk).
  • + Maintains strong character likeness while translating to a professional vector style.
  • + Includes high-quality details like multiple dogs, scoreboards, and hockey equipment.
  • The style is more of a digital illustration/avatar than a traditional hand-drawn caricature.
  • Minor text errors on the scoreboards and desk logo.

GPT Image 1

  • + Perfectly captures the 'caricature' art style with traditional watercolor textures and exaggerated facial features.
  • + Includes all elements (news desk, dog, hockey) in a clear, humorous way.
  • + Maintains the subject's outfit (denim shirt and black top) from the source image accurately.
  • The character likeness is slightly generic compared to the source.
  • Includes fewer background details compared to its competitor.

Verdict: Both models followed the complex prompt perfectly, including the news anchor profession, hockey, and dogs. FLUX.2 [max] created a more polished, modern digital illustration with a high level of detail in the background, making the scene feel like a full production. GPT Image 1 better captured the specific 'caricature' aesthetic requested, using a traditional hand-drawn style with exaggerated expressions that feel more humorous and personal.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the 'dew sparkles' requirement with distinct droplets on the grass.
  • + Superior anatomical proportions for all four animals.
  • + Beautiful lighting with soft atmospheric fog and god rays.
  • The animals feel slightly detached from one another in their pursuit of the butterflies.

GPT Image 1

  • + Captures the 'tumbling together' aspect of the prompt much better than Model A.
  • + Exceptional facial expressions that convey a very high level of joy.
  • + Stunning god rays and warm sunset lighting.
  • Anatomy issues as the kitten's front right leg appears to be emerging from its head/neck area.
  • The rabbit's scale is a bit small compared to the other animals.

Verdict: Both models followed the complex prompt very well, but FLUX.2 [max] is the winner due to its superior anatomical correctness and the clever inclusion of dew sparkles. While GPT Image 1 captured the playful 'tumbling' interaction better, it suffered from a significant anatomical artifact where the kitten's limb is misplaced.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent preservation of the original characters' features and expressions while stylized.
  • + Beautiful watercolor texture and soft, faithful pastel color palette.
  • + Perfectly captures the Studio Ghibli '90s aesthetic with clean line work.
  • The woman in the foreground is rendered in sharp focus, losing the original's depth-of-field effect.

GPT Image 1

  • + Successfully captures a hand-drawn, crayon-like texture.
  • + Preserves the composition and character roles accurately.
  • Drastic faces change that loses the specific identities of the people in the meme.
  • Clothing details like the plaid pattern on the shirt are oversimplified.
  • Colors feel slightly too saturated/neon compared to the requested soft pastel look.

Verdict: FLUX.2 [max] is the clear winner as it masterfully balances the Studio Ghibli art style with the recognizable identity of the original meme characters. While GPT Image 1 provides a charming illustration, it loses the specific facial features and distinctive clothing patterns that make the source image iconic.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [max]
Before After
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent preservation of the subject's face and original clothing details.
  • + Realistic wind-blown hair effect that maintains natural texture.
  • + High-quality rendering of green flying leaves that blend well with the existing foliage.
  • The flying leaves in the foreground are slightly blurry compared to the background.
  • The dog's leash was altered and simplified from the original.

GPT Image 1

  • + Successfully added a large amount of falling debris/leaves for an energetic feel.
  • + Good interpretation of the hair blowing in multiple directions.
  • + Preserved the dog's bushy tail well.
  • Significant loss of detail in the subject's eyes and face compared to the source.
  • Introduced noticeable artifacts and 'muddy' textures in the background trees and water.
  • The flying leaves look more like brown flecks or wood chips than realistic leaves.

Verdict: FLUX.2 [max] is the superior model because it successfully added the requested dynamic motion while maintaining the high resolution and facial integrity of the original photo. GPT Image 1 followed the instructions well, but the overall image quality suffered, resulting in a loss of sharpness and some distortion in the subject's features.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent typography including the correct grave accent on 'Caffè'
  • + Clear adherence to the light background and subtle texture request
  • + Clean vector emblem composition with a balanced stamp-like feel
  • The steam lines are very faint and could be more stylistic

GPT Image 1

  • + Strong minimalist icon design of the cloche
  • + Good inclusion of the requested 'Est. 1720' banner
  • Failed to provide a 'light background' as requested
  • Texture is a grainy digital noise rather than a subtle paper/vintage texture
  • Missed the accent on 'Caffè'

Verdict: FLUX.2 [max] followed the prompt more accurately, particularly regarding the light background and the inclusion of the accent in the text. GPT Image 1 ignored the background color instruction and produced a much coarser texture that feels less like a professional vector emblem.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [max]
GPT Image 1

AI Judge Analysis

FLUX.2 [max]

  • + Excellent logical progression from left to right visualizing the stages.
  • + High-quality vector styling with consistent iconography and a clean layout.
  • + Accurate text rendering for all requested mission phases and astronaut names.
  • Repetitive labeling where 'Earth Orbit' appears multiple times at the start.
  • The rocket icon looks more like a generic space shuttle than the Saturn V.

GPT Image 1

  • + Features a more accurate Saturn V rocket silhouette.
  • + Captures the NASA-inspired color palette perfectly.
  • + Creative use of a location pin for 'Tranquility' base.
  • Confusing layout where icons and labels do not clearly align in a sequential order.
  • Contains spelling errors such as 'EARLLUNAR'.
  • The Descent and Landing icons are essentially the same graphic repeated.

Verdict: FLUX.2 [max] is the superior model as it creates a coherent, easy-to-read infographic that follows the chronological steps requested in the prompt. While GPT Image 1 has slightly more accurate rocket iconography, its layout is disorganized and it fails on basic spelling and sequential logic.

Next steps

Explore each model