Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [pro] Black Forest Labs GPT Image 1.5 OpenAI

Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [pro]

20.3 arena score

#41 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1.5

27.1 arena score

#6 of 62 in Text-to-Image

Top 3 in Image Editing
Vote tally

Where the votes landed

FLUX.1 Kontext [pro]

0%

win rate

Ties

0%

GPT Image 1.5

0%

win rate

Shared challenges 17

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent photographic realism with a shallow depth of field.
  • + Perfectly captures the 'soft window light' lighting request.
  • + Clean, modern aesthetic with realistic materials.
  • The sphere is floating unnaturally in the center and has a fuzzy, felt-like texture rather than appearing as a standard sphere.
  • The glass cube looks more like a frame than a solid glass object.

GPT Image 1.5

  • + Highly accurate adherence to all spatial instructions.
  • + The blue sphere features realistic light reflections and a ground reflection.
  • + Excellent rendering of light refraction through the glass cube.
  • The plant is slightly more generic in shape compared to Model A.
  • The tabletop texture is a bit more repetitive than Model A.

Verdict: Both models followed the prompt perfectly. GPT Image 1.5 is the winner because it handled the physics of the scene more realistically, placing the sphere on the base of the cube and including reflections, whereas FLUX.1 Kontext [pro] featured a floating, felt-textured ball.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent natural skin textures and realistic facial features.
  • + Captures a very cinematic color grade and atmosphere.
  • + Includes the requested motion blur on background vehicles.
  • The subject appears to be riding or holding the bike rather than actively 'repairing' it.
  • The bicycle handlebar and cable anatomy is somewhat mangled and unrealistic.

GPT Image 1.5

  • + Perfectly adheres to the 'repairing' action with tools and a crouching pose.
  • + Excellent rendering of wet pavement reflections and rain on the bicycle frame.
  • + Composition feels very intentional and high-quality for a street photo.
  • The background car is sharp and lacks the requested motion blur.
  • The subject's face is partially obscured by the hat and pose, though it fits the scene.

Verdict: GPT Image 1.5 is the winner because it accurately depicts the 'repairing' action requested, whereas FLUX.1 Kontext [pro] simply shows the man holding the handlebars. While FLUX.1 Kontext [pro] captured the background motion blur better, GPT Image 1.5 offered a much more coherent scene with logical tools and bike mechanics.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent high-definition skin texture and realistic eye detail
  • + Very subtle and natural integration of warm torchlight on the armor
  • + Clean, professional photographic composition with a balanced depth of field
  • Missed the 'hair braided with small beads' requirement, showing only one clasp
  • The character looks relatively clean compared to the 'battle-worn' and 'dirt on skin' prompt requirements

GPT Image 1.5

  • + Strong adherence to all prompt details including multiple beads in braids and visible scars/dirt
  • + Exceptional textures on the leather straps and the cloth underlayer
  • + Distinct 'battle-worn' aesthetic that feels more authentic to the character description
  • The lighting is a bit harsh, causing some overexposure on the metallic highlights
  • The hair strands near the top left appear slightly noisy/frizzy

Verdict: GPT Image 1.5 is the winner because it followed every specific detail of the prompt, including the beads in the hair and the gritty, battle-worn appearance. While FLUX.1 Kontext [pro] produced a very clean and realistic portrait, it lacked the specific requested accessories (beads) and the level of 'dirt and scars' necessary to fulfill the prompt's character description.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Features very bold and aesthetically pleasing modern typography
  • + Adheres well to the minimalist white background and clean professional layout
  • + Food photography is high quality and nicely integrated into the design
  • The text is nonsensical gibberish
  • Food items do not always match the section titles (e.g., pizza under 'Mains')
  • Pricing numbers are inconsistent and confusing

GPT Image 1.5

  • + Text is perfectly legible and realistic for a restaurant
  • + Food photos are accurately categorized into the correct sections
  • + Organized grid of images provides a rich visual experience
  • The layout is more traditional and less 'minimalist' than requested
  • Design borders on generic stock imagery

Verdict: While FLUX.1 Kontext [pro] captures the requested 'modern minimalist' aesthetic more effectively through its bold typography and layout, it fails completely on content by providing gibberish text. GPT Image 1.5 provides a fully functional, high-quality menu with legible items and accurate image-to-text pairing, making it the more useful and professional output despite being slightly less stylized.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent photographic rendering of the burger texture and dripping cheese.
  • + Clean, modern typography that is easy to read.
  • + Vibrant color palette and embers create a high quality 'food ad' look.
  • Failed to include the starburst for the price element.
  • The burger is not really 'exploded', just slightly separated.
  • Repeated the price text twice at the bottom which was not requested.

GPT Image 1.5

  • + Successfully captured the 'exploded' view with widely suspended components.
  • + Included the requested starburst for the price tag.
  • + Captured all prompt elements including the specific layout of text.
  • The 'LIMITED TIME ONLY' text is cut off at the bottom of the frame.
  • Visual quality has a slightly over-sharpened, digital appearance compared to Model A.
  • Some elements, like the flying tomato chunks, look a bit cluttered.

Verdict: GPT Image 1.5 adhered much better to the specific layout instructions, providing a true 'exploded' view and the requested starburst. While FLUX.1 Kontext [pro] produced a more photorealistic and professional-looking burger, it failed to follow several key prompt details like the starburst and the exploded suspension of ingredients.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent text legibility and spelling accuracy
  • + Realistic chalk texture and board frame
  • + Good adherence to the multi-line list format requested
  • The title is not in 'elegant cursive' as requested, but rather a blocky font
  • Includes a spelling error 'foiur' at the bottom

GPT Image 1.5

  • + Accurately follows the 'elegant cursive' request for the title
  • + Highly realistic chalk smudge and grain texture
  • + Completed the truncated menu item description naturally
  • Slightly less crisp text detail compared to the other model
  • Composition is a bit bottom-heavy with more empty space in the middle

Verdict: GPT Image 1.5 followed the stylistic prompts more accurately, specifically adhering to the request for elegant cursive in the title and capturing a more authentic chalk-on-board texture. While FLUX.1 Kontext [pro] had very clear text, it failed the cursive title instruction and included a typo in the footer.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Perfectly follows the specific role-reversal prompt of horse-on-top.
  • + Creative surrealism with an astronaut as the mount.
  • + Cinematic lighting and high-quality textures on the space suit.
  • Internal logic error where the astronaut mount has horse hooves for feet.
  • Small miniature astronaut rider on the horse's back is slightly confusing.

GPT Image 1.5

  • + Excellent anatomical detail on the horse and astronaut equipment.
  • + Dynamic composition with a lunar landscape and cosmic background.
  • + Clear, sharp resolution across the entire image.
  • Failed the core prompt instruction of 'horse on top, not vice versa'.
  • Standard trope of astronaut on horse without the requested surreal reversal.

Verdict: FLUX.1 Kontext [pro] successfully followed the difficult logical reversal requested in the prompt, placing a horse on top of an astronaut. While it had some anatomical artifacts like hooves on the astronaut's feet, GPT Image 1.5 failed the negative constraint entirely by generating a standard 'astronaut riding a horse' image.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent fur texture and lighting on the capybara.
  • + Sharp focus on the subject with realistic car door framing.
  • The passenger is on a phone call rather than looking down at her phone as requested.
  • The perspective hides one of the front paws, making it unclear if both are on the wheel.

GPT Image 1.5

  • + Perfect adherence to all prompt details including the phone usage and bored expression.
  • + Strong composition showing both paws on the steering wheel and a clear view of the passenger.
  • + Highly detailed taxi driver cap that reflects a classic NY aesthetic.
  • Slightly lower contrast making the background feel a bit more washed out than the foreground.

Verdict: While FLUX.1 Kontext [pro] creates a very sharp and aesthetically pleasing image, GPT Image 1.5 followed the prompt instructions much more accurately. GPT Image 1.5 correctly depicted the businesswoman looking at her phone with a bored expression and clearly showed the capybara with both paws on the wheel, making it the more successful interpretation.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent text readability and sharpness
  • + Intricate spiderweb border design
  • + Strong contrast with the glowing jack-o-lantern
  • Includes some hallucinated text lines in the event details
  • The background is very dark, losing some 'twisted tree' detail

GPT Image 1.5

  • + Perfect text accuracy with no hallucinations
  • + Rich vintage texture and cinematic lighting
  • + More visible 'thorns' in the border as requested
  • Text is slightly less crisp than Model A
  • Jack-o-lantern light has a minor digital noise pattern

Verdict: GPT Image 1.5 performed better overall by following the text instructions perfectly without adding extraneous or nonsensical information. FLUX.1 Kontext [pro] had superior font sharpness and an elegant border, but included a line of gibberish text in the middle of the invitation details.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.1 Kontext [pro]
Before After
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent full and thick hair density
  • + Maintains original lighting and jacket details almost perfectly
  • Changes the facial features significantly, making the person look like a different individual
  • Glasses frames changed color from light brown to black

GPT Image 1.5

  • + Preserves the original facial features much more accurately than the competitor
  • + Provides a very natural-looking curly texture and realistic hairline
  • + Maintains the original glasses and clothing details perfectly
  • The transition from the sideburns to the new hair is slightly less seamless than Model A

Verdict: GPT Image 1.5 is the clear winner because it successfully added the hair while preserving the identity of the person in the source image. FLUX.1 Kontext [pro] provided a high-quality hair edit but fundamentally changed the man's facial structure and eye shape, failing the source preservation requirement.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent adherence to the 'minimal' requirement
  • + Clean and playful 3D cartoon style
  • + Perfect alignment of the requested text and flag icon
  • The salmon texture looks slightly more like plastic than realistic PBR
  • The rice grains feel a bit simplified and rounded

GPT Image 1.5

  • + Impressive realistic PBR textures on the teapot and sushi
  • + Rich detail in the diorama base
  • + Excellent text rendering and flag icon accuracy
  • Ignored the 'minimal' instruction by adding many extra props like a teapot and soy sauce
  • The composition is much more crowded than requested

Verdict: While GPT Image 1.5 offers superior textures and material realism, it failed to follow the specific 'minimal' and 'small raised diorama' instructions, resulting in a cluttered scene. FLUX.1 Kontext [pro] followed the prompt more accurately, delivering a clean, isometric 3D scene that perfectly aligns with the requested layout and minimalist aesthetic.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent preservation of the source subject's identity and original clothing.
  • + Applies a clean cartoon/caricature style while maintaining the selfie composition.
  • Completely failed to include the requested profession (TV anchor), dogs, or hockey themes.
  • The edit is more of a stylization than a thematic caricature.

GPT Image 1.5

  • + Successfully incorporates all requested elements: TV anchor desk, microphone, news graphics, dogs, and hockey imagery.
  • + Effectively uses caricature proportions by exaggerating the head and smile while maintaining a high level of detail.
  • + Creates a humorous and busy scene that fits the 'exaggerated' prompt.
  • Less preservation of the source image's specific clothing and pose.
  • The face, while recognizable, leans more toward a generic caricature than a direct translation of the source person.

Verdict: FLUX.1 Kontext [pro] creates a high-quality stylized version of the source image but fails to follow the thematic instructions, resulting in a simple cartoon selfie with no profession-related elements. GPT Image 1.5, however, fully embraces the prompt by placing the subject in a newsroom with dogs and hockey references, meeting the 'exaggerated and humorous' requirement.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent soft lighting and atmosphere.
  • + High level of detail in the fur textures.
  • + Very aesthetically pleasing composition with a consistent style.
  • Failed to generate a distinct bunny; the third animal appears more like a kitten-rabbit hybrid.
  • The animals are sitting rather than tumbling or playfully chasing as requested.

GPT Image 1.5

  • + Successfully captured the 'tumbling' and 'playfully chasing' action described in the prompt.
  • + Identifiable representation of all four requested animals (dog, cat, bunny, fox).
  • + Beautiful god rays and dew sparkle effects that match the 8K masterpiece requirement.
  • The fox's paw in the bottom right corner is somewhat distorted and dark.
  • Slightly less 'photorealistic' and more illustrative/saturated than Model A.

Verdict: While FLUX.1 Kontext [pro] offers a more dreamlike and polished visual quality, it failed to differentiate the bunny from the kittens, essentially creating three cats. GPT Image 1.5 adhered much better to the prompt's action and specific animal requirements, successfully rendering the fox, kitten, puppy, and bunny in a playful, tumbling pile.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent preservation of the source image's composition and poses.
  • + Clearly captures the late 80s/early 90s Ghibli aesthetic with clean line work and watercolor textures.
  • + Maintains the distinct blue plaid pattern of the man's shirt very accurately.
  • The woman in the foreground's facial expression is slightly more generic than the source's look.
  • The paper texture is a bit uniform across the entire canvas.

GPT Image 1.5

  • + Successfully captures the dreamy, soft-glow lighting often found in modern Ghibli or Makoto Shinkai films.
  • + Good use of pastel colors and a nostalgic, warm color grade.
  • + Preserves the character's positions and essential outfits well.
  • The man's hand disappears into his back/pants unnaturally.
  • The blurring effect on the foreground woman is a bit excessive compared to the source and the style.
  • Loss of some fine detail in the background due to the heavy lighting bloom.

Verdict: FLUX.1 Kontext [pro] is the winner because it perfectly balances the Studio Ghibli art style with an incredibly high level of source preservation, maintaining the exact patterns and poses of the original meme characters. GPT Image 1.5 succeeds in creating a 'dreamy' atmosphere, but it suffers from anatomical errors in the man's hand and an overly aggressive blur that obscures the subject matter.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.1 Kontext [pro]
Before After
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent source preservation, maintaining the identity and features of the woman and dog perfectly.
  • + Subtle, realistic blowing hair that integrates seamlessly with the original style.
  • The 'flying leaves' instruction is barely met, with only a few small green dots visible.
  • Overall motion is less 'dynamic' than requested.

GPT Image 1.5

  • + Strong adherence to the 'flying leaves' prompt with numerous leaves of varied colors.
  • + High degree of energetic motion in the hair.
  • Significant loss of source preservation; the woman's face has been noticeably altered.
  • The leaf placement looks like a flat overlay rather than being integrated into the 3D space of the scene.

Verdict: FLUX.1 Kontext [pro] is much better at preserving the original image's integrity and quality, but it was too conservative with the requested edits. GPT Image 1.5 successfully added a lot of motion and leaves, but failed the editing task by changing the subject's face and creating a cluttered, less realistic composition.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with a professional serif font and accurate accented characters.
  • + Perfectly captures the 'subtle texture on light background' requested in the prompt.
  • + Clean, minimalist vector aesthetic suitable for a real restaurant logo.
  • Includes a minor spelling error in the banner ('EEST' instead of 'EST').
  • Steam interpretation is very basic compared to the rest of the illustration.

GPT Image 1.5

  • + Beautifully rendered cloche with sophisticated lighting and shading.
  • + Perfect spelling on all text elements including the banner.
  • + Dynamic and elegant steam illustration.
  • Ignored the request for a 'light background', opting for solid black.
  • The layout feels a bit cramped compared to the balanced spacing of the first image.
  • Typography for 'Caffè' is slightly less clean than the main brand name.

Verdict: FLUX.1 Kontext [pro] creates a much more authentic 'vintage minimalist' emblem that adheres strictly to the color and background requirements, despite a small typo in the banner. GPT Image 1.5 produces a more polished illustration with perfect text, but fails the background instruction and feels less like a minimalist vector logo and more like a detailed digital illustration.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [pro]
GPT Image 1.5

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Features a bold, high-quality main title
  • + Includes clean, modern vector styling for the lunar terrain
  • Confused astronomical objects, placing Saturn's rings on the rocket and the Moon
  • Text labeling is disorganized and partially nonsensical (e.g., '2, far Collins')
  • Failed to follow the requested sequential 6-step infographic structure

GPT Image 1.5

  • + Strictly adheres to all 6 requested steps in a logical infographic layout
  • + Text rendering is clear, accurate, and correctly positioned for every stage
  • + Maintains a consistent, clean flat-vector icon style across all panels
  • The 'Tranquillity' speech bubble is slightly redundant with the landing label
  • The inclusion of astronaut silhouettes was not explicitly requested, though it fits the theme

Verdict: GPT Image 1.5 is the clear winner for its superior prompt adherence, perfectly executing the 6-step sequence with accurate labels and icons. In contrast, FLUX.1 Kontext [pro] produced a disorganized image with significant factual errors, such as placing planetary rings around the rocket and celestial bodies, and failed to follow the requested step-by-step structure.

Next steps

Explore each model