Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1.5 OpenAI Wan 2.5 (Preview) Alibaba

Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.

GPT Image 1.5

27.1 arena score

#7 of 62 in Text-to-Image

Top 3 in Image Editing
Skill signature · Text-to-Image

Wan 2.5 (Preview)

23.4 arena score

#27 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1.5

0%

win rate

Ties

0%

Wan 2.5 (Preview)

0%

win rate

Shared challenges 19

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to lighting instructions with a clear source from the left.
  • + The glass cube rendering is structurally sound with realistic reflections on the base.
  • + The plant is very clearly visible through multiple facets of the glass.
  • The sphere is slightly large for the described 'small' blue sphere.
  • The book alignment looks a bit too perfectly centered, bordering on sterile.

Wan 2.5 (Preview)

  • + Beautiful naturalistic lighting with visible dust motes and soft shadows.
  • + The sphere has a very tactile, matte finish that contrasts well with the glass.
  • + Includes the plant in a pot, adding more context to the scene.
  • The book is floating slightly above the glass cube rather than sitting on it.
  • The glass cube has distorted geometry near the top right corner where it meets the book.

Verdict: GPT Image 1.5 is the preferred image because it maintains better structural integrity and spatial logic, ensuring the book and cube are physically touching. While Wan 2.5 (Preview) has more atmospheric rendering and beautiful lighting, it suffers from a major physics error with a floating book and slightly warped glass edges.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent preservation of the man's facial features and specific hairstyle
  • + The man is clearly interacting with the vehicle
  • + High quality rendering of the car's interior and materials
  • The car is positioned awkwardly facing the 'wrong' way compared to the road's curve
  • Changes the orientation of the steering wheel to the right-hand side

Wan 2.5 (Preview)

  • + Perfectly preserves the exact look of the white car from the source image
  • + Excellent sense of motion and dynamic composition
  • + Accurately represents a California coastline road
  • The man is very small and difficult to see clearly
  • The man's unique hairstyle is not perfectly preserved at this distance

Verdict: GPT Image 1.5 does a better job of preserving the identity of the man from the second source image, placing him clearly in the frame. However, Wan 2.5 (Preview) produces a much more realistic and professionally composed image of the car itself, maintaining its exact details while creating a better sense of motion. Wan 2.5 (Preview) is the preferred choice for a cohesive final photograph.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'imperfect framing' and 'candid' aspects of the prompt.
  • + Highly realistic lighting and reflections on the wet pavement.
  • + Naturalistic skin textures and convincing interaction with the bicycle tools.
  • The motion blur on the passing car is fairly subtle.

Wan 2.5 (Preview)

  • + Strong bokeh and shallow depth of field effects.
  • + Clear rain streaks visible in the background.
  • + Correct inclusion of various bicycle parts on the ground.
  • The bike anatomy is slightly illogical with a floating kickstand and disconnected gears.
  • The lighting on the man's shirt feels a bit too clean and studio-lit for a 'candid' rain shot.
  • The composition is very centered, ignoring the 'imperfect framing' instruction.

Verdict: GPT Image 1.5 followed the prompt's stylistic cues much better, specifically the 'imperfect framing' and 'no stylization' requirements, resulting in a photo that feels truly candid. Wan 2.5 (Preview) produced a higher-contrast, more 'perfect' image that feels like a staged photoshoot, and it suffered from several anatomical errors in the bicycle structure.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Exquisite texture on the weathered leather and hammered metal
  • + Highly realistic skin texture with believable dirt and scars
  • + Subtle and professional color grading using warm torchlight
  • The braids are somewhat messy and loose compared to the prompt's bead-focused request

Wan 2.5 (Preview)

  • + Very distinct and clear rendering of hair beads
  • + Strong adherence to the 'ornate engraved' armor description
  • + Good use of bokeh and shallow depth of field
  • Skin texture appears slightly smooth and plastic compared to Model A
  • Lighting on the face is a bit flat despite the presence of a bright torch

Verdict: GPT Image 1.5 wins due to its superior textural details, particularly in the lifelike rendering of the skin and the weathered appearance of the leather and metal. While Wan 2.5 (Preview) executed the hair beads more literally, it lacked the cinematic depth and micro-detail present in GPT Image 1.5's composition.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent typography with perfect legibility and no spelling errors.
  • + Realistic food photography that accurately represents the menu items.
  • + Professional layout that mimics a real-world restaurant menu.
  • The color accents at the top are very subtle compared to the request for 'vibrant accents'.
  • The grid layout for photos is less symmetrical than Model B.

Wan 2.5 (Preview)

  • + Creative use of a grid layout for food photography.
  • + Strong adherence to the 'vibrant accents' and minimalist aesthetic.
  • + Good spatial balance between images and text blocks.
  • Nonsensical text with numerous spelling errors and gibberish characters.
  • Visual glitches in some food images, making the dishes look artificial.
  • Poor font choice for small body text that is blurred and unreadable.

Verdict: GPT Image 1.5 is the clear winner as it produces a functional, professional menu with perfectly rendered text and realistic food photography. While Wan 2.5 (Preview) attempts an interesting grid-based minimalist layout, it fails significantly on text legibility and image coherence, resulting in a design that is unusable in a real-world context.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent photorealistic texture on the meat patty and toasted bun
  • + Dynamic, chaotic energy that perfectly matches the 'exploded' request
  • + Strong integration of fiery glowing effects throughout the entire composition
  • The text 'LIMITED TIME ONLY' is slightly plain compared to the rest of the graphics
  • The starburst shape for the price is a bit busy with overlapping sparks

Wan 2.5 (Preview)

  • + Very clean and readable typography with a creative 'dripping' fiery effect
  • + Elegant, clear distribution of burger components that allows each ingredient to be seen
  • + High-quality rendering of the price starburst as a distinct graphic element
  • The background is a bit too clean and lacks the intense 'fiery' atmosphere requested compared to the other model
  • The image feels a bit more like a composite graphic than a photorealistic physical explosion

Verdict: GPT Image 1.5 wins on raw photorealism and capturing the intense, chaotic energy of a fiery explosion, making the food look highly appetizing and dynamic. Wan 2.5 (Preview) provides superior graphic design and clearer text rendering, but it feels more like a static advertisement than the high-motion, detailed scene requested.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent chalk texture with realistic grit and varying opacity.
  • + Highly accurate spelling and adherence to the prompt's specific date and items.
  • + Convincing natural handwriting style that feels authentic to a café environment.
  • The composition is a tight crop, showing less of the 'cozy café' atmosphere.
  • The cursive for the title is somewhat simplified rather than 'elegant'.

Wan 2.5 (Preview)

  • + Strong composition showing the chalkboard within a blurred, cozy café background.
  • + Consistent, bold handwriting style with good contrast.
  • + Effective use of layout with underlines for each menu item.
  • Failed to include 'Herbs' in the Grilled Octopus item.
  • The price for the cookies is repeated awkwardly on two lines.
  • Text appears significantly cleaner and less grainy than real chalk, bordering on a digital look.

Verdict: GPT Image 1.5 is the winner because it followed the text requirements perfectly, including the full text of all menu items and the specific date, while maintaining a very realistic chalk texture. Wan 2.5 (Preview) produced a better overall composition with a nice background, but it failed on text accuracy by omitting words and repeating price tags inconsistently.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + High level of textural detail in the spacesuit and lunar ground.
  • + Cinematic lighting with a gritty, realistic atmosphere.
  • + Excellent composition featuring a variety of space elements like asteroids and a moon lander.
  • Failed to follow the specific spatial instruction (horse on top, not vice versa).
  • Anatomical issues with the horse's front legs and hoof structure.

Wan 2.5 (Preview)

  • + Vibrant, surreal color palette with a beautiful nebula effect.
  • + Clean rendering of the astronaut and horse with fewer anatomical glitches.
  • + Dynamic sense of motion with the flowing mane and dust trails.
  • Failed to follow the specific spatial instruction (horse on top, not vice versa).
  • Less surface detail on the astronaut's gear compared to the competitor.

Verdict: Both GPT Image 1.5 and Wan 2.5 failed the negative constraint/logic trap at the heart of the prompt, which requested the horse to be on top of the astronaut. GPT Image 1.5 produced a more detailed, textured environment that feels cinematic, while Wan 2.5 opted for a cleaner, more colorful surreal aesthetic. GPT Image 1.5 is slightly preferred for its complexity and atmospheric lighting, despite both failing the primary logic test.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Successfully preserved the lower half of the original person's face and skin condition.
  • + Maintained the background and the wooden structure accurately.
  • + Correctly applied all components of the outfit including the scarf and watch.
  • Cropped the head off the person, failing to manage the full composition.
  • Could not verify if the eyes/hair were preserved due to the crop.

Wan 2.5 (Preview)

  • + Applied the outfit, accessories, and jewelry with high fidelity to the source.
  • + Great integration of lighting and shadows on the clothing layers.
  • Completely failed to preserve the person from Image 1, replacing him with the face from Image 2.
  • Ignored the explicit instruction to keep the person's exact face and hair unchanged.

Verdict: GPT Image 1.5 followed the core instruction to preserve the original person (preserving his unique skin condition), though it failed significantly by cropping the head off the image. Wan 2.5 produced a much cleaner visual result but completely failed the primary identity preservation task, treating the prompt as a face-swap from Image 2 onto the background of Image 1. GPT Image 1.5 is the winner because it actually attempted the difficult task of dressing one person in another's clothes.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent photorealism with gritty, cinematic lighting
  • + Fur texture and lighting on the capybara are highly detailed
  • + Captures the bored expression of the passenger perfectly
  • The capybara's paws look slightly more like human hands/fingers in a capybara skin than natural paws

Wan 2.5 (Preview)

  • + More colorful and vibrant background representing Manhattan night lights
  • + Better rendering of the capybara's natural paws on the steering wheel
  • + Good inclusion of the taxi roof sign and interior details
  • The capybara's face is a bit too smooth and symmetrical, losing some realism
  • The passenger's hands and phone interaction are slightly less natural than Model A

Verdict: Both models followed the prompt exceptionally well, capturing the surreal humor of a capybara cab driver. GPT Image 1.5 is the winner due to its superior photorealistic quality and cinematic atmosphere, whereas Wan 2.5 (Preview) has a slightly more digital, processed look, particularly in the lighting and fur textures.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent text integration and vintage typography that feels cohesive with the art style.
  • + Superior adherence to the 'vintage dark parchment' aesthetic with a unified sepia-tone palette.
  • + Atmospheric and spooky border that blends naturally into the composition.
  • The parchment texture is very heavy, slightly obscuring the clarity of the background graveyard.

Wan 2.5 (Preview)

  • + High-quality rendering of the fire inside the jack-o-lantern.
  • + Clean and legible event details at the bottom.
  • + Good contrast between the blue night sky and the parchment border.
  • The text 'Halloween Party Invitation' has some slight layering artifacts on the 'H' and 'y'.
  • The border feels like a separate digital overlay rather than an integrated part of a 'vintage poster'.

Verdict: GPT Image 1.5 is the winner as it perfectly captures the 'vintage gothic' request, making the text and art feel like a single cohesive physical object. While Wan 2.5 (Preview) has vibrant colors and sharp details, its composition feels more like a modern digital graphic, and the typography is less polished than GPT Image 1.5.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
GPT Image 1.5
Before After
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent texture and realistic hair density
  • + Natural integration of the hairline with the existing forehead
  • + Perfect preservation of the facial features, clothes, and lighting
  • The hair color is slightly cooler in tone than the beard

Wan 2.5 (Preview)

  • + Successfully applied the requested edit
  • + Preserved the composition and lighting well
  • The hairline transition on the forehead looks slightly artificial
  • The hair texture is a bit smooth and looks styled rather than natural hair growth
  • Small artifact in the hair strands above the forehead

Verdict: GPT Image 1.5 is the clear winner as it adds the hair with a much more convincing and natural hairline, whereas Wan 2.5 (Preview) has a visible seam where the hairline meets the skin. GPT Image 1.5 also provides a superior texture that perfectly matches the realism and lighting of the source photo.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'isometric miniature diorama' request with a detailed scene.
  • + High-quality PBR material rendering on wood, ceramic, and food textures.
  • + Perfect text layout including the flag icon centered as requested.
  • Slightly more complex than the 'minimal' request, though still very clean.
  • The background is not a perfectly solid blue, featuring a slight gradient.

Wan 2.5 (Preview)

  • + Successfully captures the '3D cartoon' style with soft, rounded textures.
  • + Very clean and minimal composition on a small raised base.
  • + Accurate text rendering and placement.
  • Does not follow the 45-degree top-down isometric angle as well as Model A.
  • The background contains significant vignette/gradient rather than being solid.
  • Less detail in the sushi model compared to Model A.

Verdict: GPT Image 1.5 followed the prompt more precisely, particularly regarding the isometric perspective and the layout of the diorama base. While Wan 2.5 captured the 'cartoon' aesthetic well, GPT Image 1.5 provided a more professional-looking 3D render with superior material details and better alignment with the specific top-down angle requested.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent caricature style that maintains a strong facial resemblance to the source image.
  • + Rich, detailed background incorporating the news studio and hockey game seamlessly.
  • + Creative integration of the prompt by putting one of the dogs in a hockey helmet.
  • Slightly crowded composition with many elements competing for attention.

Wan 2.5 (Preview)

  • + Successfully incorporates all elements: news anchor, dogs, and hockey.
  • + Preserves the subject's original denim shirt from the source image.
  • + Clean, vector-like caricature style with high contrast.
  • The facial features lose the specific likeness of the subject in favor of a generic cartoon face.
  • The composition is a bit flat with a simple white background border.

Verdict: GPT Image 1.5 is the clear winner as it creates a much better caricature that actually looks like the woman in the source image. It also does a superior job of blending the requested themes (hockey, news, dogs) into a cohesive and visually interesting scene, whereas Wan 2.5 (Preview) provides a more generic, clip-art style illustration.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent fur texture details on all animals
  • + Realistic interaction and tumbling poses
  • + Natural integration of lighting and dew sparkles
  • The cat's anatomy is slightly messy with too many visible paw pads
  • God rays are a bit soft/blurry compared to the crispness of the animals

Wan 2.5 (Preview)

  • + Dynamic sense of movement and 'chasing' as requested
  • + Vibrant, distinct colors in the wildflower meadow
  • + Consistent lighting across all subjects
  • The fox's eyes appear unnaturally blue and metallic
  • Dew drops are rendered as floating glass orbs instead of surface moisture
  • Fur texture looks somewhat flat or plastic compared to Model A

Verdict: GPT Image 1.5 produced a much more realistic and 'masterpiece' quality image with superior fur textures and a more natural group interaction. While Wan 2.5 captured the 'chasing' action better, it suffered from unnatural eye rendering on the fox and unrealistic floating dew drops that detracted from the photorealism.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent soft pastel color palette and warm, nostalgic mood.
  • + Captures the 'painterly' texture and dreamy lighting requested.
  • + Preserves the depth of field and composition of the source image well.
  • Faces are a bit generic and lose some of the specific character expressions of the source.
  • Overuse of lens flare/sparkle effects can be distracting.

Wan 2.5 (Preview)

  • + Strong adherence to the Ghibli character design style with clean line art.
  • + Perfectly captures the facial expressions of the original meme characters in an anime style.
  • + High clarity and faithful preservation of the background elements and clothing patterns.
  • Lighting is a bit flat compared to the 'dreamy' request.
  • Color palette is slightly more desaturated/cool than 'warm pastel'.

Verdict: Both models did an exceptional job translating the 'Distracted Boyfriend' meme into an illustration. GPT Image 1.5 captured the requested soft, warm, and painterly atmosphere better, but Wan 2.5 (Preview) was much more successful in translating the specific comedic expressions of the original photo into a recognizable Studio Ghibli character aesthetic. Wan 2.5 is the winner for its superior character work and line art clarity.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
GPT Image 1.5
Before After
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'flying leaves' instruction with various sizes and depths
  • + Modified the hair to show dramatic wind-blown movement
  • + Maintained the overall composition and lighting of the source image
  • The leash handle has significant structural artifacts and looks broken
  • Some leaves look like they are floating static rather than moving due to lack of motion blur

Wan 2.5 (Preview)

  • + Natural-looking hair movement that feels consistent with a breeze
  • + Preserved the details of the woman's face and the dog very accurately
  • + Clean preservation of the leash and park architecture
  • The flying leaves are sparse and look like flat lime-green stickers
  • Less 'energetic and lively' feel compared to the other model

Verdict: GPT Image 1.5 produced a much more dynamic and energetic image that fully embraced the prompt, though it suffered from some structural artifacts on the leash. Wan 2.5 (Preview) was much more conservative with the edit, resulting in a cleaner image but one that failed to capture the 'flying leaves' or 'lively feel' as effectively. GPT Image 1.5 is the winner for its superior creativity and adherence to the spirit of the edit.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent typography with a hand-drawn feel
  • + Strong vector-style shading and texture on the dome
  • + Accurate rendering of the accent in 'Caffè'
  • Ignored the 'light background' instruction, delivering it on solid black
  • Text alignment on the bottom banner is slightly off-center

Wan 2.5 (Preview)

  • + Perfectly followed the 'light background' and 'subtle texture' instructions
  • + Clean, minimalist vector aesthetic that feels professional
  • + Well-composed with a decorative border that adds to the vintage feel
  • The 'a' in 'Florian' is slightly malformed
  • The steam icon is a bit generic compared to the first image

Verdict: Wan 2.5 (Preview) more accurately followed the prompt by placing the logo on a textured light background, whereas GPT Image 1.5 failed that negative/positive space instruction entirely. While GPT Image 1.5 has slightly more stylish typography, Wan 2.5 (Preview) provides a more complete and usable design for the requested minimalist vintage theme.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1.5
Wan 2.5 (Preview)

AI Judge Analysis

GPT Image 1.5

  • + Excellent text rendering for all panel titles.
  • + Strict adherence to the requested grid-based infographic layout.
  • + Consistent flat-vector style with a clean, cohesive NASA-inspired palette.
  • One minor spelling error in 'Tranquillity' (though both English spellings are technically valid).

Wan 2.5 (Preview)

  • + Dynamic trajectory illustration linking the Earth and Moon.
  • + Clean iconography for the astronaut profiles.
  • + Follows the NASA-inspired color scheme well.
  • Confusing visual hierarchy where 'Descent' and 'Landing' text are floating without clear icons.
  • Inconsistent rocket designs; the ascent stage looks somewhat like a shuttle.
  • The trajectory lines are cluttered and cross over each other awkwardly.

Verdict: GPT Image 1.5 followed the prompt more effectively by creating a clear, step-by-step panel layout where each numbered stage corresponds to a distinct icon and title. While Wan 2.5 (Preview) had a more integrated composition, it failed to clearly represent the 'Descent' and 'Landing' steps as separate phases and had less consistent vector styling.

Next steps

Explore each model