Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [pro] Black Forest Labs GPT Image 1 OpenAI

Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [pro]

20.3 arena score

#41 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1

22.6 arena score

#32 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [pro]

100.0%

win rate

Ties

0.0%

GPT Image 1

0.0%

win rate

100.0% 0.0% ties 0.0%
Shared challenges 17

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent adherence to lighting and placement instructions.
  • + Highly realistic glass textures with clean edges.
  • + Great spatial awareness with the blue sphere resting naturally on the bottom surface.
  • The plant is significantly out of focus compared to the rest of the scene.
  • The blue sphere has a slightly fuzzy or felt-like texture rather than appearing perfectly smooth.

GPT Image 1

  • + Strong color saturation and sharp details on the book and plant.
  • + Accurate material representation for the glass cube with slight green-tinted edges.
  • + Clear visibility of the plant through the glass as requested.
  • The lighting feels slightly more omnidirectional than the requested 'soft light from the left'.
  • The bottom of the cube has an unexpected metallic reflective base rather than being pure glass.

Verdict: Both models followed the prompt instructions perfectly, including the difficult spatial positioning of multiple objects. FLUX.1 Kontext [pro] is slightly preferred for its more natural, photorealistic lighting and delicate glass render, whereas GPT Image 1 feels slightly more staged with a heavier cube construction.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Natural skin texture and realistic face rendering.
  • + Good application of rain effects and wet pavement reflections.
  • + Accurate shallow depth of field for a 50mm look.
  • The subject appears to be riding or holding the bike rather than repairing it.
  • The bicycle handlebars and cables show some structural AI artifacts.

GPT Image 1

  • + Strong adherence to the 'repairing' action in the prompt.
  • + Excellent composition with a side profile that feels candid.
  • + Realistic textures on the bicycle seat and wet metal.
  • The hands merging into the bike chain area are a bit messy.
  • Less noticeable motion blur from passing cars compared to the other model.

Verdict: GPT Image 1 captures the 'repairing' aspect of the prompt much better than FLUX.1 Kontext [pro], which depicts the man simply standing with or riding the bike. While FLUX.1 Kontext [pro] has slightly superior skin texture, GPT Image 1's composition and narrative accuracy make it the better overall response to the specific scene requested.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent photorealistic skin texture and lifelike eyes
  • + Beautifully detailed engraving on the plate armor
  • + Subtle and realistic integration of the beard and facial hair
  • The beads in the hair are very minimal and less noticeable than requested
  • Lighting is a bit flat compared to the dramatic 'torchlight' prompt

GPT Image 1

  • + Superior adherence to the 'hair braided with small beads' prompt
  • + Very gritly, battle-worn appearance with convincing dirt and scars
  • + Dramatic warm lighting that effectively highlights the metal textures
  • Skin texture is slightly more stylized/painterly than Model A
  • The pupils in the eyes look slightly inconsistent in shape

Verdict: Both models performed excellently, with FLUX.1 Kontext [pro] providing a cleaner, more photorealistic portrait with incredible armor detail. However, GPT Image 1 better captured the specific details of the prompt, such as the beads in the braids and the dramatic, battle-worn atmosphere created by the torchlight and grime. GPT Image 1 is the winner for its superior prompt adherence and storytelling through visual details.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with clean, bold sans-serif fonts.
  • + Sophisticated layout that feels like a real graphic design piece.
  • + Accurate interpretation of the sections requested including categorized food items.
  • The text content is mostly gibberish.
  • The food photos are slightly less vibrant compared to Model B.

GPT Image 1

  • + Features a very clean 2x2 grid of high-quality food photography.
  • + Text labels are closer to English words with logical price formatting.
  • + High-contrast colors and professional lighting in the images.
  • The 'Mains' section is poorly integrated into the layout.
  • Composition is a bit generic and lacks the 'graphic design' flair of Model A.
  • Some spelling errors like 'descrigion' and 'Apperoiation'.

Verdict: FLUX.1 Kontext [pro] produced a more authentic-looking menu design with superior use of negative space and typographic hierarchy, although the text is largely nonsensical. GPT Image 1 followed the grid prompt more literally and provided clearer, more appetizing food photography, though its layout is slightly more cramped. FLUX.1 Kontext [pro] is the likely winner for better capturing the 'modern minimalist design' aesthetic asked for in the prompt.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent photographic quality and lighting on the burger components.
  • + Strong typography rendering with clear, glowing effects.
  • + Creates a highly professional advertisement layout with a rich background.
  • Repeated the price three times in the image unnecessarily.
  • Missing the starburst element for the price as requested.
  • The 'exploded' effect is less pronounced than requested, as the burger is mostly stacked.

GPT Image 1

  • + Perfect adherence to the 'exploded' instruction with clear gaps between all layers.
  • + Includes the starburst element for the price as specifically requested.
  • + Consistent fiery text effect across all copy.
  • Failed the price accuracy, showing '€.99' instead of '€6.99'.
  • The bottom bun looks slightly flat and untextured compared to the rest of the image.
  • The overall composition feels a bit more cramped at the top.

Verdict: GPT Image 1 followed the structural layout and specific creative elements of the prompt better, particularly the exploded view and the starburst. However, FLUX.1 Kontext [pro] produced a much more realistic and appetizing food image, though it failed several text-related details like the starburst and the correct number of price instances. GPT Image 1 is the winner for prompt adherence despite the typo in the price.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent chalk texture with realistic dusty strokes
  • + High prompt adherence for specific menu items and date
  • + Logical chalkboard layout with a wooden frame
  • Significant spelling errors in the footer text at the bottom
  • The title is not in 'elegant cursive' as requested

GPT Image 1

  • + Perfect spelling in the footer text
  • + Consistent handwriting style throughout the entire board
  • + Clean and centered composition
  • Failed to include the '$' sign for the final price item
  • Calculated chalk texture looks slightly like a digital brush rather than physical chalk
  • The title is printed caps rather than elegant cursive

Verdict: Both models failed to produce the requested 'elegant cursive' for the title, opting for print styles instead. FLUX.1 Kontext [pro] captures the authentic texture of chalk much better, but suffers from garbled text in the fine print at the bottom. GPT Image 1 provides much cleaner legibility and perfect spelling, making it the more functional design despite the missing currency symbol.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent adherence to the complex spatial requirement of the horse being on top.
  • + Creative depiction of a smaller astronaut mounting the horse which is mounting the larger astronaut.
  • + Clean, cinematic lighting and high-quality textures on the spacesuit.
  • The anatomy of the horse's rear legs and the small astronaut's merging are a bit confusing.

GPT Image 1

  • + High visual quality with a classic cinematic aesthetic.
  • + Strong composition and clear subject rendering.
  • Failed the core prompt instruction of having the horse on top of the astronaut.
  • Standard interpretation lacking the requested surreal role-reversal.

Verdict: FLUX.1 Kontext [pro] successfully interpreted the difficult 'horse on top' instruction, creating a surreal and thought-provoking image that matches the prompt's intent. GPT Image 1 ignored the specific 'not vice versa' instruction, providing a high-quality but generic image of an astronaut riding a horse.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent fur texture rendering
  • + Deep vibrant colors in the taxi paint
  • + Highly realistic cinematic lighting
  • The passenger is on a phone call instead of looking at the phone as requested
  • Only one paw is clearly visible on the steering wheel
  • The perspective is awkward, looking in through the driver side window

GPT Image 1

  • + Perfect adherence to the passenger's action and expression
  • + Stronger composition with a front-on view of the driver and passenger
  • + Accurately shows both paws on the steering wheel as requested
  • The capybara's hands look slightly more humanoid/primate-like than paws
  • Overall image is a bit darker and less sharp than Model A

Verdict: Model B (GPT Image 1) followed the prompt much more accurately, correctly capturing the passenger's 'bored' expression while looking at a phone, and placing both of the driver's paws on the wheel. While Model A (FLUX.1 Kontext [pro]) has superior lighting and texture quality, it failed the specific instruction regarding the passenger's behavior and displayed a less effective composition.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography style that matches the 'gothic' request perfectly.
  • + Very high resolution details in the pumpkin and moon.
  • + Strict adherence to the banner text and date format.
  • Includes a line of gibberish text in the event details section.
  • The layout is a bit bottom-heavy with small text.

GPT Image 1

  • + Better 'parchment' texture and color palette for a vintage look.
  • + Organized and clean text layout that is easier to read.
  • + Accurate rendering of the requested event details without hallucinated lines.
  • Text is in a serif font rather than a truly 'elegant gothic' style.
  • Jack-o-lantern and background elements are slightly blurry or painterly compared to Model A.

Verdict: FLUX.1 Kontext [pro] captures the gothic aesthetic much better with its sharp typography and high-contrast lighting, but it suffers from a line of gibberish text. GPT Image 1 follows the text instructions more reliably and has a superior layout for an invitation, though its font choice is less stylistic. Model A is preferred for its visual flair and better interpretation of 'gothic' and 'cinematic'.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.1 Kontext [pro]
Before After
GPT Image 1
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent preservation of the original facial features and bone structure
  • + Realistic hair texture and natural integration with the sideburns
  • + Maintains identical lighting and background to the source image
  • The forehead appears slightly shortened compared to the original

GPT Image 1

  • + Successfully adds a thick head of hair
  • + Preserves the overall composition of the image
  • Significantly alters the person's face, making them look like a different individual
  • The hair texture looks somewhat like a wig and lacks a natural-looking hairline
  • The glasses and eye area were altered unnecessarily

Verdict: FLUX.1 Kontext [pro] successfully added the hair while the person remained recognizable as the same individual from the source image. In contrast, GPT Image 1 modified the internal facial features so much that it no longer looks like the original person, failing the source preservation requirement of the edit task.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography style that matches the 3D aesthetic.
  • + High-quality texture on the salmon and rice, showing great PBR material properties.
  • + Clean and minimalist composition that looks like a professional 3D render.
  • The flag icon is integrated oddly into the text line rather than being a separate element.
  • The rice grains look slightly larger than natural, resembling pearls.

GPT Image 1

  • + Perfect adherence to the layout instructions, including the flag icon placement.
  • + More variety in the sushi types which makes for a more interesting scene.
  • + Clean, soft lighting that perfectly matches the 'cartoon miniature' request.
  • Slightly less realistic PBR textures on the salmon compared to the other model.
  • The flag icon is a flat 2D graphic which slightly clashes with the 3D text styling.

Verdict: Both models followed the prompt exceptionally well, producing high-quality isometric dioramas. GPT Image 1 is preferred because it offers a more complete scene with better variety and superior layout adherence regarding the text and flag icon, whereas FLUX.1 Kontext [pro] felt a bit more simplistic in its composition.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Successfully converts the subject into a stylized vector-art caricature.
  • + Strongly preserves the clothing and selfie-pose composition from the source image.
  • Fails completely to incorporate the requested job (anchor), dogs, or hockey themes.
  • The edit is too simple, focusing only on the style change rather than the content additions.

GPT Image 1

  • + Excellent adherence to all prompt details, including the news desk, dog, and hockey elements.
  • + Captures the classic 'big-head' caricature style with high-quality watercolor texture.
  • + Maintains the subject's likeness and clothing while transforming the setting.
  • The hands holding the papers are slightly malformed.
  • Text on the hockey stick is incoherent.

Verdict: FLUX.1 Kontext [pro] failed the image editing task by ignoring all the thematic requests (job, dogs, hockey), instead only providing a basic cartoon filter of the source image. GPT Image 1 successfully interpreted the full prompt, creating a humorous and detailed caricature that creatively integrated every requested element into a cohesive scene.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent depiction of warm golden light and god rays.
  • + Includes all four requested animals with very soft, fine fur textures.
  • The characters are mostly sitting rather than 'playfully chasing and tumbling'.
  • The rabbit lacks defining bunny features, looking more like a small kitten-rabbit hybrid.

GPT Image 1

  • + Strong prompt adherence to the action, showing the animals actually running and chasing.
  • + Individual animal species are more distinctly recognizable, especially the rabbit and fox.
  • + Better composition that feels dynamic and joyful.
  • The fox kit has three front paws visible due to a developmental artifact.
  • The lighting lacks the specific 'god rays' mentioned in the prompt compared to the other model.

Verdict: GPT Image 1 captured the dynamic action of 'chasing and tumbling' much better than FLUX.1 Kontext, which opted for a static pose. While GPT Image 1 has a notable anatomical error with the fox's legs, its superior composition and clear distinction between the four species make it the more successful interpretation of the prompt.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent preservation of the original image's composition and character poses.
  • + Captures the iconic Ghibli character design style perfectly while staying recognizable.
  • + Hand-painted watercolor texture is very high quality and consistent.
  • The man's expression is slightly more neutral/melancholic than the original pucker/surprise.

GPT Image 1

  • + Strong implementation of the requested 'deeply dreamy' and 'vibrant pastel' color palette.
  • + Good use of soft lighting and hazy textures.
  • Loss of significant source detail, specifically the sharp plaid pattern on the shirt.
  • The woman in the foreground has her eyes closed, losing the gaze towards the camera.
  • Character faces are a bit more generic and less polished than model A.

Verdict: FLUX.1 Kontext [pro] is the clear winner as it masterfully balances the Ghibli aesthetic with strict adherence to the source image's iconic composition and character expressions. GPT Image 1 leans heavily into a hazy, pastel glow that obscures many of the original photo's key details like the plaid shirt and character eye contact.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.1 Kontext [pro]
Before After
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent preservation of the woman's face and original features.
  • + Highly realistic hair-blowing effect that looks naturally integrated.
  • + Subtle and tasteful addition of flying leaves that do not clutter the frame.
  • The dog's tail position changed significantly from the original source.

GPT Image 1

  • + Stronger 'energetic' feel with more numerous leaves scattered throughout the scene.
  • + Effective hair motion that suggests a gust of wind.
  • Noticeable anatomical distortion on the woman's left hand (holding the leash).
  • The leash itself has become disconnected and fragmented with a strange loop.
  • The image has a more 'photoshopped' or artificial look compared to the source.

Verdict: FLUX.1 Kontext [pro] is the clear winner for its superior preservation of the source image's quality and anatomy. While GPT Image 1 adds more leaves, it suffers from significant artifacts, including a mangled hand and a broken leash, whereas FLUX.1 Kontext [pro] creates a believable motion effect while maintaining high visual fidelity.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with a correct grave accent on 'Caffè'.
  • + Beautifully rendered vintage paper texture that matches the 'subtle texture' prompt.
  • + Highly balanced composition with a classic vector emblem feel.
  • The steam element is a bit thick, resembling a leaf or solid shape more than vapor.

GPT Image 1

  • + Strong contrast with the use of warm brown and cream tones.
  • + Clean iconography for the cloche and steam elements.
  • + Accurate text rendering for the restaurant name and date.
  • Ignored the 'light background' instruction, using a solid black background instead.
  • Incorrect accent mark on 'Caffè' (it uses a grave accent in the prompt and real life, but the model used an acute-style mark).

Verdict: FLUX.1 Kontext [pro] followed all prompt instructions, including the specific request for a light background and subtle texture, creating a highly authentic vintage logo. GPT Image 1 failed on the background color requirement and missed the nuance of the Italian accent mark, resulting in a less polished vector emblem.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [pro]
GPT Image 1

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent high-resolution aesthetics with a polished 'modern vector' feel.
  • + Creative use of a trajectory path to connect different stages of the mission.
  • + Successfully captures a NASA-inspired color palette and atmospheric depth.
  • Confused logical mapping, such as labeling a ringed planet (Saturn) as 'Earth'.
  • Fails to provide the specific 1-6 numbered steps requested in the prompt.
  • Poor text legibility and nonsensical labels like '2, f ar Collins'.

GPT Image 1

  • + Strict adherence to the iconography requested, including the specific lunar module icons.
  • + Better organized layout that resembles a standard infographic with distinct sections.
  • + Higher degree of text accuracy for the labels and crew names.
  • Significant spelling errors such as 'EARLLUNAR' and inconsistent font spacing.
  • Visual style is slightly more dated/clunky than the requested modern aesthetic.
  • Missing the 'Lunar Orbit' step specifically requested, jumping from Translunar to Descent.

Verdict: FLUX.1 Kontext [pro] creates a much more visually appealing and artistic poster, but it fails significantly on the logic of an infographic, mislabeling astronomical bodies and producing scrambled text. GPT Image 1 follows the structural instructions much more closely and includes the requested iconography, even though it suffers from typical AI spelling errors. GPT Image 1 is the likely winner for better following the functional requirements of an infographic task.

Next steps

Explore each model