Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1.5 OpenAI Wan 2.6 Alibaba

Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.

GPT Image 1.5

27.1 arena score

#5 of 62 in Text-to-Image

Top 3 in Image Editing
Skill signature · Text-to-Image

Wan 2.6

23.3 arena score

#26 of 62 in Text-to-Image

Top 2 in Image-to-Video
Vote tally

Where the votes landed

GPT Image 1.5

82.4%

win rate

Ties

5.9%

Wan 2.6

11.8%

win rate

82.4% 5.9% ties 11.8%
Shared challenges 19

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'partially visible through the glass' instruction.
  • + Clean, modern glass rendering with realistic reflections.
  • + Strong composition with a clear, sharp focus on all required elements.
  • The plant appears to be floating or lacks a visible pot.
  • The blue sphere is quite large relative to the 'small' description.

Wan 2.6

  • + Beautiful natural lighting with realistic shadows on the wooden table.
  • + The plant is realistically placed in a pot.
  • + Good weathered texture on the red book adds character.
  • The plant is positioned to the side rather than 'behind' the cube.
  • The perspective of the cube’s top edge is slightly distorted under the book.
  • The sphere lacks the 'small' scale requested in the prompt.

Verdict: GPT Image 1.5 correctly placed the plant behind the glass cube as requested, creating a more accurate spatial arrangement, whereas Wan 2.6 placed the plant to the right. Both models produced high-quality images, but GPT Image 1.5 is preferred for its superior adherence to the complex spatial layering described in the prompt.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
GPT Image 1.5
Wan 2.6
0% wins 0% ties 100% wins

AI Judge Analysis

GPT Image 1.5

  • + Excellent preservation of the man's facial features and unique hairstyle.
  • + Accurately places the car on a winding coastal road that fits the California theme.
  • + Good integration of the interior car details like the white leather seats.
  • The car is positioned on the wrong side of the road for US driving.
  • The steering wheel is oddly low and small relative to the man's hands.

Wan 2.6

  • + Great sense of motion with high-quality background blur on the wheels and road.
  • + The composition is more dynamic and cinematic for a car photo.
  • + Preserves the man's likeness and clothing patterns very well.
  • The steering wheel and hands are slightly distorted and messy.
  • The man's hair is slightly simplified compared to the source image.

Verdict: Both models successfully combined the man and the car into the requested California coastline setting while preserving their key characteristics. GPT Image 1.5 has better facial fidelity but places the car on the left side of the road, whereas Wan 2.6 provides a much more convincing action shot with superior lighting and a realistic driving perspective.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent depiction of rain, with realistic droplets on surfaces and wet fabric textures.
  • + Highly realistic mechanical details on the bicycle and tools.
  • + Great sense of depth and atmospheric light.
  • The car in the background lacks the requested motion blur, appearing almost sharp.
  • The framing is a bit more centered and clean than 'imperfect' suggests.

Wan 2.6

  • + Strong adherence to the 'motion blur from passing cars' prompt requirement.
  • + Excellent 'candid' feel with messy tool placement and street positioning.
  • + Deeply realistic skin texture on the man's hands and face.
  • Uncanny water droplet artifacts on the man's jacket that look more like physical bumps or glass beads than liquid.
  • Minor anatomical confusion where the hands meet the tool and bicycle chain.

Verdict: GPT Image 1.5 produces a cleaner and more technically sound image with superior textures on the bike and clothes, while Wan 2.6 captures the 'motion blur' and 'candid' street aesthetic more accurately. GPT Image 1.5 is preferred overall because Wan 2.6 suffers from distracting visual artifacts on the character's clothing that break the realism.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Exceptional texture detail on the skin, metal, and leather straps.
  • + Captures the 'battle-worn' aesthetic perfectly with realistic dirt and scratches.
  • + Superior lighting effects with warm torchlight glints on the ornate engravings.
  • The hair beads are somewhat generic and repetitive in design.

Wan 2.6

  • + Excellent interpretation of the braided hair with varied, colorful beads.
  • + Strong composition with a clear background light source creating nice bokeh sparks.
  • + Good inclusion of the frayed cloth underlayer as requested.
  • Skin texture appears slightly smooth/plastic under the dirt compared to the realism of Model A.
  • Light reflecting off the forehead appears more like sweat or a wet surface than torchlight.

Verdict: GPT Image 1.5 is the winner due to its superior rendering of materials and lifelike textures, particularly in the weathered metal and scarred skin. While Wan 2.6 did an excellent job with the specific hair bead request and composition, its skin and lighting effects felt slightly less natural.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1.5
Wan 2.6
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1.5

  • + Perfectly legible English text with accurate menu item names and descriptions.
  • + Excellent layout balance where the food photos correspond directly to the adjacent text sections.
  • + High-quality, appetizing food photography with clear focus and vibrant colors.
  • The grid for photos is vertical on one side rather than a centralized grid mentioned in the prompt.

Wan 2.6

  • + Stronger adherence to the 'grid' request for the layout of food images.
  • + Vibrant colorful accents on the borders create a modern aesthetic.
  • Text is largely gibberish with numerous spelling errors and artifacts.
  • The scaling of the text is inconsistent and messy in the 'Mains' column.
  • Visual quality of the food items is lower, with some blurring and strange shapes.

Verdict: GPT Image 1.5 is the clear winner as it produces a professional, usable menu with perfect text rendering and high-quality food photography. Wan 2.6 captures the 'grid' layout and colorful accents well, but fails significantly on text legibility and overall professional finish.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent photorealistic texture on the meat patty and bun.
  • + Vibrant, high-energy composition with great use of glowing embers.
  • + Perfect adherence to all text requirements with a cohesive fiery aesthetic.
  • The 'exploded' effect is slightly more vertical and static compared to Model B.

Wan 2.6

  • + Stronger sense of dynamic motion with sauce droplets and angled components.
  • + Very impressive fiery text effect for the main title.
  • + Clean background separation between the burger and the smoke.
  • The meat patty looks slightly less realistic and more like a processed disc.
  • The lighting on the lettuce and tomatoes is a bit flat compared to the surrounding environment.

Verdict: GPT Image 1.5 is the winner because it achieves a higher level of photorealistic detail, particularly in the food textures which are crucial for an advertisement. While Wan 2.6 has a more dynamic 'exploded' layout, GPT Image 1.5 features superior integration of the fiery lighting across all elements and more consistent text rendering.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent text accuracy with no spelling errors.
  • + Consistent and elegant handwriting style that looks highly realistic.
  • + Perfect spacing and layout on the board.
  • The 'chalkboard dust' effect is very uniform, looking slightly digital in its distribution.

Wan 2.6

  • + Excellent chalk texture and realistic smudging on the board surface.
  • + More dynamic and authentic-looking café environment context with visible lighting.
  • + Strong 'handwritten' aesthetic with charming character variation.
  • Repeating price tags for each item creates a cluttered and repetitive layout.
  • Slight punctuation issues such as the trailing comma after 'Specials'.

Verdict: GPT Image 1.5 provides near-perfect text rendering and adheres strictly to the layout instructions with clean, elegant handwriting. Wan 2.6 has superior surface textures and a more authentic 'chalk' feel but suffers from repetitive text elements (double pricing) and slightly messy composition.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent level of fine detail in the astronaut's gear and the lunar lander.
  • + Coherent atmospheric lighting that fits a cinematic space theme.
  • + High resolution with sharp textures on the rocks and the horse's fur.
  • Completely failed the negative constraint/spatial instruction of 'horse on top'.
  • The inclusion of dust on the ground feels slightly physically inconsistent for 'in space' without a clearer planetary surface context.

Wan 2.6

  • + Striking color palette with vibrant nebulae and lighting.
  • + Clean composition with a mystical, ethereal quality.
  • Completely failed the negative constraint/spatial instruction of 'horse on top'.
  • Anatomy issues where the horse's rear leg appears to blend into the dust/background in an unnatural way.

Verdict: Both models completely failed the intentional logic puzzle in the prompt, which requested the horse to be on top of the astronaut. Because both followed the standard 'astronaut on horse' trope instead, GPT Image 1.5 is the winner due to its superior technical execution, higher detail density, and more realistic textures compared to Wan 2.6.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent replication of the complex plaid pattern on the scarf.
  • + Accurately includes accessories like the watch seen in the second image.
  • + Successfully adapts the pose to fit the hands-in-pockets stance seen in Image 2.
  • Failed to preserve the head and face of the person from Image 1, cropping them out almost entirely.
  • Changed the overall framing and composition of the shot.

Wan 2.6

  • + Perfectly preserved the subject's face, skin patterns (vitiligo), and hair from Image 1.
  • + Maintained the exact original background, lighting, and composition of the source photo.
  • + Successfully integrated clothing layers and added sunglasses similar to the style in Image 2.
  • The scarf pattern is slightly simplified compared to the source image.
  • Did not include the watch or hands in the frame as seen in the source material.

Verdict: Wan 2.6 is the clear winner because it followed the instruction to keep the person's face and background completely unchanged while applying the edit. GPT Image 1.5 failed the primary preservation task by cropping out the subject's head and face, essentially generating a new body image rather than editing the provided photo.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent close-up detail on the capybara's fur and facial expression.
  • + Precise adherence to the 'taxi' text on the cap.
  • + The lighting is warm and cinematic, consistent with a night taxi ride.
  • The capybara's anatomy on the steering wheel looks slightly cramped due to the tight framing.
  • The background passenger is more blurred than in the competing image.

Wan 2.6

  • + Dynamic composition that shows more of the car exterior and the rainy Manhattan environment.
  • + Excellent realism in the textures of the coat and the rainy window glass.
  • + Captures the 'bored' expression of the passenger very effectively.
  • The capybara's paws on the steering wheel look somewhat distorted and unnatural.
  • The hat has a generic police-style badge rather than saying 'TAXI' as requested by the prompt's context.

Verdict: GPT Image 1.5 provides a much more intimate and detailed character study with perfect text rendering on the hat, while Wan 2.6 offers a wider, more atmospheric scene that captures the rainy New York vibe. GPT Image 1.5 is preferred for its superior rendering of the capybara and better adherence to specific costume details.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent typography integration with distinct gothic styles
  • + Perfect adherence to all text requirements with zero spelling errors
  • + Superior cinematic lighting and a cohesive vintage parchment texture
  • The thorn border is a bit repetitive in its pattern

Wan 2.6

  • + Beautiful composition with twisted trees framing the central Jack-o-lantern
  • + Vibrant color contrast between the blue night sky and orange glow
  • + High-quality rendering of the thorn and web border details
  • Slight typography misalignment in the main title
  • The text lacks the 'vintage' integrated feel of Model A
  • Text layout at the bottom is less decorative

Verdict: GPT Image 1.5 is the clear winner for this invitation challenge due to its exceptional handling of typography and layout; the text feels naturally integrated into the vintage poster aesthetic. While Wan 2.6 offers a beautiful illustration with impressive lighting and depth, GPT Image 1.5's professional-looking graphic design and accurate adherence to the specific text prompts make it much more effective as a functional invitation.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
GPT Image 1.5
Before After
Wan 2.6
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1.5

  • + Excellent preservation of facial features and identity
  • + Realistic curly hair texture that matches the beard and eyebrows
  • + Maintains original lighting and image sharpness perfectly
  • The hairline transition on the forehead is a bit sharp/uniform

Wan 2.6

  • + Impressive volume and thickness of hair
  • + Matches the overall color palette of the image well
  • Substantially alters facial features, making the man look younger and like a different person
  • The hair texture appears slightly too soft/painterly compared to the gritty detail of the source beard
  • The hairline is unnaturally low, encroaching on the glasses

Verdict: GPT Image 1.5 is the clear winner as it successfully adds the hair while perfectly preserving the identity, facial features, and detail levels of the original subject. In contrast, Wan 2.6 essentially generates a new face that resembles the original but loses the specific characteristics of the source person, which fails the preservation requirement of an image editing task.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent PBR material rendering with realistic wood, ceramic, and liquid textures.
  • + Accurate and aesthetically pleasing text placement with a drop shadow for depth.
  • + Rich, detailed scene that maintains high clarity.
  • Includes several extra elements like the teapot and soy sauce bottle not explicitly asked for.

Wan 2.6

  • + Follows the 'minimal' aspect of the prompt more closely.
  • + Clean isometric composition with a clear raised diorama base.
  • + Accurate text rendering and placement.
  • The 'JAPAN' text is slightly off-center compared to the rest of the logo block.
  • The textures appear a bit more plastic and less 'refined' compared to Model A.

Verdict: GPT Image 1.5 produced a much higher quality image with superior textures and lighting, though it added extra decorative elements. Wan 2.6 followed the 'minimal' constraint better, but its lighting and materials look flatter and less professional. GPT Image 1.5 is the winner due to its beautiful 3D execution and polished typography.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
GPT Image 1.5
Wan 2.6
50% wins 50% ties 0% wins

AI Judge Analysis

GPT Image 1.5

  • + Excellently captures the subject's facial features in a caricature style.
  • + Cleverly integrates all three prompts (news, dogs, hockey) into a cohesive scene.
  • + High-quality rendering with professional digital painting aesthetics.
  • The fingers on the hand holding the microphone are anatomicaly messy.

Wan 2.6

  • + Functional caricature style that includes all requested elements.
  • + Good use of space to show the full character and setting.
  • The character's face bears very little resemblance to the source image provided.
  • The hockey stick is being held in a physically impossible way by both the human and the dog.

Verdict: GPT Image 1.5 is the clear winner as it successfully maintains the subject's likeness while translating her into a caricature, whereas Wan 2.6 creates a generic cartoon face. GPT Image 1.5 also provides a more creative composition, integrating the hockey element into the background news ticker and a dog's accessory rather than just having characters hold a stick.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Excellent fur texture and detail on all animals
  • + Strong facial expressions with emotive 'big eyes'
  • + Perfect adherence to the requested sunset lighting and god rays
  • Anatomical issues with the kitten, which appears to have too many paws/legs in the tumbling position
  • The background is quite cluttered with bokeh particles

Wan 2.6

  • + Better dynamic composition showing the animals actually 'chasing' and 'tumbling'
  • + More natural integration of the butterfly elements
  • + Higher clarity in the mid-ground and background lighting
  • The fox's eyes look slightly less natural than the others
  • Some minor artifacting on the floating dandelion seeds

Verdict: Both models followed the prompt exceptionally well, capturing the lighting and specific animal types. GPT Image 1.5 has superior texture rendering and more adorable faces, but suffers from significant anatomical confusion in the kitten's limb placement; Wan 2.6 provides a more coherent scene with better movement and distinct, well-placed subjects.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
GPT Image 1.5
Wan 2.6
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'soft pastel' and 'warm, nostalgic' mood
  • + Captures a dreamlike aesthetic that fits more modern Ghibli styles
  • + Maintains the characteristic expressions of the source image perfectly within an illustrative style
  • The image is a bit too blurry/hazy, losing some structural detail
  • The texture feels more like a digital filter than a hand-painted medium

Wan 2.6

  • + Superior watercolor and hand-painted texture that evokes classic Ghibli backgrounds
  • + Very high fidelity to the source image's composition and character details
  • + Clearer linework and cleaner facial rendering
  • The color palette is a bit cooler and less 'nostalgic' than requested
  • Characters look slightly more Western-realistic than the Ghibli anime style usually dictates

Verdict: Both models did an excellent job of translating the 'Distracted Boyfriend' meme into an illustrative style. GPT Image 1.5 captured the warm, dreamy atmosphere of a Ghibli film better, while Wan 2.6 provided much more convincing hand-painted textures and watercolor effects that feel like authentic production art.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
GPT Image 1.5
Before After
Wan 2.6
67% wins 0% ties 33% wins

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'flying leaves' part of the prompt with high density.
  • + Strong dynamic hair movement that looks natural for a windy day.
  • + Perfectly preserves the source image subject and background.
  • The leaves appear somewhat flat and lack motion blur relative to their quantity.
  • Some leaves overlap the subjects in a slightly busy way.

Wan 2.6

  • + Successfully adds wind-blown hair movement.
  • + Maintains high fidelity to the original source image.
  • + Subtle and clean integration of a few leaves.
  • The leaf count is very low, making the 'energetic and lively' feel less apparent.
  • The background remains static, lacking the full atmosphere requested.

Verdict: Both models did an excellent job of preserving the source image while adding the requested hair movement. GPT Image 1.5 is the clear winner for its commitment to the 'lively' atmosphere, adding a significant number of flying leaves that transform the mood of the photo, whereas Wan 2.6 was too conservative with the leaves.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1.5
Wan 2.6

AI Judge Analysis

GPT Image 1.5

  • + Perfect text rendering for both the name and the date banner.
  • + Excellent vector emblem style with a professional, balanced layout.
  • + Includes a stippled texture that perfectly matches the 'vintage' prompt.
  • The 'steam' element is a bit thick, looking more like a solid shape than vapor.

Wan 2.6

  • + Good color palette following the brown and cream request.
  • + Clean vector-style illustration of the cloche dome.
  • + Background texture aligns well with the 'subtle texture' prompt.
  • The 'Est. 1720' banner is awkwardly placed and small.
  • The typography is much more generic compared to Model A.
  • The steam lines are slightly disconnected and thin.

Verdict: GPT Image 1.5 followed the prompt much more effectively, producing a cohesive logo with professional typography and a well-integrated 'Est. 1720' banner. Wan 2.6 produced a decent image, but the banner placement was awkward and the overall composition lacked the sophisticated 'vintage minimalist' feel requested.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1.5
Wan 2.6
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1.5

  • + Successfully included all six requested infographic steps in order.
  • + Excellent text rendering for labels like 'LAUNCH', 'TRANSLUNAR', and astronaut names.
  • + High-quality vector style and consistent iconography that matches the requested NASA palette.
  • The 'Descent' and 'Landing' modules are nearly identical in appearance.
  • Some minor overlapping of graphic elements in the Earth Orbit section.

Wan 2.6

  • + Clean aesthetic with a clear NASA-inspired color palette.
  • + Readable text for the title and astronaut names.
  • Completely failed to include the requested six-step infographic sequence.
  • Missing all requested icons (Saturn V, orbit rings, trajectory arc, lunar module).
  • Composition is mostly empty space with very little informational value.

Verdict: GPT Image 1.5 is the clear winner as it followed every detail of the complex prompt, creating a full six-step infographic with accurate vector icons and clear text. In contrast, Wan 2.6 failed to provide the infographic steps, offering only a minimalist poster with astronaut names that ignored the bulk of the instructional prompt.

Next steps

Explore each model