Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [max] Black Forest Labs GPT Image 1 Mini OpenAI

Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [max]

23.7 arena score

#23 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1 Mini

25.0 arena score

#13 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [max]

0%

win rate

Ties

0%

GPT Image 1 Mini

0%

win rate

Shared challenges 17

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent handling of complex light refraction and shadows on the table surface.
  • + Highly realistic textures, especially the glittery surface of the sphere and the wood grain.
  • + Strong adherence to the spatial prompt regarding the plant behind the glass.
  • The book's placement looks slightly floating/unbalanced on the glass edge.

GPT Image 1 Mini

  • + Clean, minimalist composition with accurate item placement.
  • + Clear and simple representation of the blue sphere and red book.
  • + Good lighting consistency from the left side.
  • The glass cube lacks realistic thickness and refraction characteristics compared to Model A.
  • The plant is very blurred, making the 'visible through the glass' instruction less impactful.
  • Texture on the wood and sphere is overly smooth/plastic.

Verdict: FLUX.1 Kontext [max] provides a much more sophisticated image with realistic light play, caustics, and complex textures that make the scene feel tangible. GPT Image 1 Mini correctly follows the prompt placement but lacks the photographic depth and material realism seen in FLUX.1 Kontext [max], particularly in how the glass interacts with the light and the background.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent depiction of rain with atmospheric streaks.
  • + Vibrant reflections on the wet pavement add to the cinematic quality.
  • + Compelling composition with the bike spanning the foreground.
  • Physical interaction between the hands and the chain is messy and anatomically confusing.
  • Rain streaks sometimes appear as static white lines in front of the lens rather than 3D depth.

GPT Image 1 Mini

  • + Natural skin texture and facial expression are highly realistic.
  • + Better anatomical correctness with the hands and fingers.
  • + Accurate shallow depth of field and soft background bokeh.
  • The 'light rain' is barely visible, missing the atmospheric quality of Image A.
  • The pavement is wet, but lacks the dynamic reflections requested in the prompt.

Verdict: FLUX.1 Kontext [max] creates a more cinematic and atmospheric scene with strong rain visual effects and vibrant wet reflections, though it struggles with the fine details of the man's hands on the bike. GPT Image 1 Mini feels more grounded and realistic in its subject rendering and hand anatomy, but it is too subtle with the rain and reflections requested in the prompt. GPT Image 1 Mini is the likely winner for its superior realism and avoidance of AI artifacts, despite being less 'cinematic'.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Sublime metal engraving and reflection details.
  • + Excellent skin texture with visible pores and realistic sweat/dirt.
  • + Striking lifelike eyes that catch the warm lighting perfectly.
  • Hair braids look slightly like thick ropes rather than traditional hair braids.
  • The sparks in the background are a bit elongated and distracting.

GPT Image 1 Mini

  • + Great 'battle-worn' aesthetic with more grit and blood on the face.
  • + Sophisticated, subtle engraving on the plate armor.
  • + Better integration of the braids into the character's hairstyle.
  • Lighting is a bit flat compared to the dynamic reflections in Image A.
  • The leather strap detail is less distinct than requested.

Verdict: Both models followed the prompt exceptionally well, but FLUX.1 Kontext [max] wins on technical visual quality, particularly with the hyper-realistic skin textures and the way the armor reacts to torchlight. While GPT Image 1 Mini captured the 'battle-worn' atmosphere more convincingly with facial scarring and dirt, it lacked the depth and clarity found in the materials of the FLUX.1 output.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent professional photorealism in the food items
  • + Realistic layout that mimics a real-world printed menu
  • + Balanced use of brand accents and typography
  • Text is mostly gibberish despite looking like a real menu
  • Food variety is lacking as it is almost entirely pizza

GPT Image 1 Mini

  • + Perfect text rendering for the requested sections
  • + Distinct and varied food photos covering all categories in a clear grid
  • + Clean, high-contrast minimalist aesthetic
  • The layout is overly simplistic and lacks descriptive text placeholder areas
  • The food images look slightly more like stock graphics than integrated photography

Verdict: GPT Image 1 Mini adhered better to the specific section requests and provided perfect text rendering, though the design is very basic. FLUX.1 Kontext [max] produced a much more realistic and professional-looking menu layout with higher visual quality, but failed to include the specific requested sections and usable text.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealistic texture on the burger patty and buns
  • + Great sense of high-speed motion with flying debris and embers
  • + Clean and readable typography centered well in the layout
  • Failed to place the price in a starburst as requested
  • The burger is more of a complete sandwich with side pieces rather than a fully 'exploded' vertical stack
  • Missing the fiery/glowing effect specifically on the secondary text

GPT Image 1 Mini

  • + Perfect adherence to the 'exploded' vertical stack layout
  • + Captures the 'fiery, glowing' text effect much better than Model A
  • + Correctly included the price within a starburst graphic
  • The burger ingredients look somewhat plasticky or artificial compared to Model A
  • The texturing on the bottom bun is slightly muddy
  • Lighting on the burger is a bit flat despite the glowing background elements

Verdict: While FLUX.1 Kontext [max] produces a significantly more high-quality and photorealistic image of the food itself, it failed multiple specific formatting instructions regarding the starburst and text effects. GPT Image 1 Mini adhered perfectly to every technical detail of the prompt, including the exploded vertical composition and the specific graphic elements, making it a better advertisement design despite the lower realism in the burger's texture.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent text rendering with perfect spelling and spacing.
  • + Realistic chalk texture including smudges on the board.
  • + Strong environmental context with café elements visible.
  • The title is in print/block letters rather than the requested elegant cursive.

GPT Image 1 Mini

  • + The chalk texture on the letters is very grainy and realistic.
  • + Good layout and alignment of prices.
  • Failed to render the title in cursive as requested.
  • The handwriting looks somewhat uniform, bordering on a font style.
  • The board feels less integrated into a physical space compared to Model A.

Verdict: FLUX.1 Kontext is the superior choice because it captures the messy, authentic nature of a real chalkboard, including smudges and natural handwriting variations. While both models failed to provide the cursive title requested, FLUX.1 Kontext produced much cleaner and more legible text throughout the entire menu.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent lighting and color contrast between the brown horse and the dark space.
  • + High texture detail on the astronaut suit and horse's coat.
  • + Dynamic composition with a sense of motion.
  • Failed the negative constraint; the astronaut is riding the horse instead of the horse on top.

GPT Image 1 Mini

  • + Atmospheric cinematic lighting with subtle lens flare effects.
  • + Clean anatomical rendering of both the astronaut and the horse.
  • + Consistent starry background with good depth of field.
  • Failed the negative constraint; like Model A, it shows an astronaut riding a horse.
  • Low color saturation makes the image feel slightly flat compared to Model A.

Verdict: Both FLUX.1 Kontext [max] and GPT Image 1 Mini failed the specific prompt instruction to have the 'horse on top'. Consequently, both models generated a standard astronaut riding a horse. FLUX.1 Kontext [max] is the preferred choice as it offers significantly better lighting, more vibrant colors, and sharper textural details.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealistic texture on the capybara's fur and the leather taxi cap
  • + Captures the sense of motion and vibrant New York city lights through the window
  • + Superior rendering of the taxi's exterior and side-view perspective
  • The passenger is holding a phone to her ear like a call, rather than looking at it as requested
  • Only one paw is clearly making contact with the steering wheel

GPT Image 1 Mini

  • + Perfect adherence to the passenger's expression and action (looking at phone)
  • + Clearly shows both paws on the steering wheel as requested
  • + Nostalgic, moody lighting that fits a night-time taxi theme
  • The image is overall very dark with lower contrast in the interior
  • The capybara's paw anatomy on the steering wheel looks slightly distorted

Verdict: FLUX.1 Kontext [max] produces a much higher quality, sharper image with better lighting and textures, though it misses the specific passenger action. GPT Image 1 Mini adheres more closely to every detail of the prompt, including the passenger's bored gaze at her phone and the 'both paws' requirement, but suffers from a darker, less professional-looking render. FLUX.1 Kontext [max] is the winner for its realistic cinematic quality despite minor prompt deviations.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent gothic typography that feels period-appropriate.
  • + Superior lighting and depth, especially the glow emanating from the jack-o-lantern.
  • + Very detailed border with clear web and thorn elements.
  • Redundant text at the bottom repeats the location twice.
  • The commas in the date '30,10,2026' are a slight deviation from the requested period punctuation.

GPT Image 1 Mini

  • + Perfectly follows the specific text for location and date without repetition.
  • + The scroll banner design is very elegant and well-integrated into the composition.
  • + Clean, readable layout that balances the vintage parchment feel well.
  • The lighting is much flatter compared to the high-contrast cinematic lighting in Model A.
  • The border details are muddier and less distinct than the thorns and webs in the competing model.

Verdict: FLUX.1 Kontext [max] creates a more visually striking and atmospheric poster with impressive cinematic lighting and gothic flair, though it suffers from redundant text at the bottom. GPT Image 1 Mini is more precise with the textual instructions and layout, but the overall image quality is flatter and less detailed. FLUX.1 Kontext [max] is the winner for capturing the 'polished' and 'cinematic' mood requested in the prompt.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.1 Kontext [max]
Before After
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent preservation of the subject's unique facial structure and wrinkles
  • + Realistic hair texture and thickness
  • + High-quality skin texture and detail preservation
  • The hairline transition is slightly harsh with a faint white line artifact

GPT Image 1 Mini

  • + Natural, messy hair texture that fits the desert setting
  • + Good blending of the hair at the temples
  • Significant loss of source identity by smoothing skin and changing facial features
  • Alters the shape of the nose and glasses
  • Removed original details like the button and weathering on the jacket

Verdict: FLUX.1 Kontext [max] successfully added a full head of hair while perfectly preserving the identity, facial features, and details of the man in the source image. In contrast, GPT Image 1 Mini significantly altered the person's face (smoothing skin and changing structural features), failing the 'preserve facial features' part of the instruction.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent soft refined 3D textures on the sushi and rice.
  • + Very clean typography and centered composition.
  • + Natural-looking soft lighting and shadows.
  • Missed the small flag icon request.
  • The brown text color has slightly less contrast against the blue than Model B.

GPT Image 1 Mini

  • + Includes all prompted elements including the small flag icon.
  • + Crisp text with high contrast for readability.
  • + Good adherence to the 45-degree isometric perspective.
  • The sushi looks a bit more plasticky/artificial compared to Model A.
  • The grains of rice look slightly less realistic than those in the other image.

Verdict: Both models followed the prompt very well, but GPT Image 1 Mini is the winner because it successfully included every requested element, including the small flag icon which FLUX.1 Kontext [max] missed. While FLUX.1 Kontext [max] had slightly more refined 3D textures on the food, GPT Image 1 Mini delivered a more complete adherence to the specific instructions.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Successfully captures the likeness of the original subject's hair and denim shirt.
  • + Incorporates all three elements: TV anchor, dog, and hockey stick.
  • + Uses a clean, vibrant digital illustration style.
  • Added glasses that weren't in the original image, reducing the likeness.
  • The hockey stick is partially cut off at the top.
  • The dog looks more like a clip-art icon than a character in the scene.

GPT Image 1 Mini

  • + Stronger 'caricature' style with exaggerated facial features that still mimic the original person.
  • + Cleverly includes a hockey puck in addition to the stick for better thematic coverage.
  • + Better integration of the dog as a co-anchor, fitting the humorous request.
  • The text in the background is slightly clipped.
  • The hand holding the microphone is a bit small and awkwardly placed.

Verdict: Both models followed the instructions well, but GPT Image 1 Mini feels more like a true caricature with its exaggerated grin and hand-drawn texture. While FLUX.1 Kontext preserved the clothing better, it added glasses which weren't in the source and had a more generic cartoon style compared to the expressive personality in GPT's version.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent fur texture and individual hair detail
  • + Magical, dreamlike lighting with glowing butterflies that accentuates the 'wholesome' vibe
  • + Perfectly captures all four animals as requested in the prompt
  • The composition is a bit static with the animals sitting rather than 'tumbling' and 'chasing'
  • The puppy's tongue and mouth area look slightly AI-rendered/smooth

GPT Image 1 Mini

  • + Successfully captures the action of 'tumbling' and 'chasing' with dynamic poses
  • + Stronger adherence to the 'god rays' lighting effect
  • + Included all four requested animals with distinct characteristics
  • The bunny's anatomy is a bit awkward around the neck/ears
  • The animals appear somewhat 'cut out' from the background rather than being fully integrated in the grass

Verdict: Both models followed the prompt perfectly, including the four specific animals. FLUX.1 Kontext [max] produced a more polished, high-detail '8K masterpiece' with beautiful textures and lighting, whereas GPT Image 1 Mini better captured the requested action of the animals playing and tumbling. FLUX.1 Kontext [max] is the winner for its superior visual quality and more cohesive, high-end artistic finish.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Perfectly captures the Studio Ghibli cel-shaded animation style.
  • + Preserves the exact composition and poses of the original meme.
  • + Excellent use of soft pastel colors and hand-painted watercolor textures in the background.
  • The characters' expressions are slightly softened, losing some of the sharp intensity of the original 'distracted' and 'outraged' looks.

GPT Image 1 Mini

  • + Captures a warm, nostalgic mood with a soft-focus atmosphere.
  • + Good attention to the textures of the hair and clothing.
  • The style feels more like a colored pencil sketch than Ghibli-inspired animation.
  • The faces are somewhat distorted and lose the specific character of the original photo.
  • Relatively low contrast makes the image feel slightly washed out.

Verdict: FLUX.1 Kontext [max] is the clear winner as it successfully translates the source image into a distinct Studio Ghibli aesthetic while maintaining perfect structural preservation of the original 'distracted boyfriend' meme. GPT Image 1 Mini provides a soft, warm illustration, but fails to capture the specific Ghibli art style and loses clarity in the characters' features.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.1 Kontext [max]
Before After
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent source preservation, maintaining the woman's face and the dog's features accurately.
  • + Subtle and realistic blowing hair effect.
  • + Successfully added falling leaves as requested.
  • The hand on the right side of the image has mutated into a long, unnatural appendage reaching for the dog.
  • The addition of leaves is a bit sparse compared to the 'energetic' request.

GPT Image 1 Mini

  • + Very dynamic wind effect on the hair that looks natural and windswept.
  • + Large amount of falling leaves creates a strong sense of lively motion.
  • + Generally preserves the overall scene layout well.
  • Significantly altered the woman's facial features, losing the likeness of the source image.
  • The background elements like the bridge and trees have been altered or simplified.
  • Small artifacts around the edges of the leaves and hair.

Verdict: FLUX.1 Kontext [max] does a much better job of preserving the identity of the person and the dog from the source image, but it fails significantly on human anatomy by creating a mutated third-hand-like shape. GPT Image 1 Mini creates a more 'energetic' and 'lively' scene with more wind and leaves, but it fails at source preservation by changing the woman's face. FLUX.1 Kontext [max] is preferred for maintaining the original image's integrity despite the structural error in the arm.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with correct accent placement and spacing.
  • + Perfect adherence to the 'light background' and 'warm brown and cream tones' instructions.
  • + Authentic vintage texture that looks like a printed emblem.
  • The steam element is very small and slightly disconnected from the cloche knob.

GPT Image 1 Mini

  • + Strong iconography for the cloche and steam.
  • + Clear layout with high contrast.
  • Failed the negative constraint for a 'light background' by using a black background.
  • The typography is less refined, with an awkward accent mark over the E in 'CAFFÈ'.
  • The color palette leans towards gold and black rather than the requested brown and cream.

Verdict: FLUX.1 Kontext [max] followed the prompt instructions much more accurately, particularly regarding the light background and warm cream tones. While GPT Image 1 Mini produced a clean icon, it failed the color and background requirements and had less sophisticated typography.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [max]
GPT Image 1 Mini

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography and high-end graphic design aesthetic.
  • + Includes highly accurate silhouettes of the crew with correct names.
  • + Effective use of the requested NASA-inspired color palette and subtle gradients.
  • Failed to include exactly 6 steps as requested in the prompt.
  • The rocket icon and some lunar icons are more generic/abstract than Saturn V or Lunar Module specifically.
  • Text logic is confusing with labels pointing to the wrong visual elements.

GPT Image 1 Mini

  • + Followed the specific 6-step sequence exactly as outlined in the prompt.
  • + Icons for the Saturn V and Lunar Module are much more accurate and consistent.
  • + Clean, consistent flat-vector illustration style with a clear logical flow.
  • The 'Translunar' trajectory icon is a bit messy and over-complicated compared to other icons.
  • Text rendering is slightly less refined than Model A.
  • Composition feels slightly cramped at the top.

Verdict: GPT Image 1 Mini is the clear winner for its superior prompt adherence, delivering all six requested steps with icons that actually resemble the Saturn V and Lunar Module. While FLUX.1 Kontext [max] has a more sophisticated and professional graphic design look, it failed significantly on the instructional logic and specific content requirements of the infographic.

Next steps

Explore each model