Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1.5 OpenAI Wan 2.7 Alibaba

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

GPT Image 1.5

27.1 arena score

#7 of 62 in Text-to-Image

Top 3 in Image Editing
Skill signature · Text-to-Image

Wan 2.7

20.5 arena score

#39 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1.5

0%

win rate

Ties

0%

Wan 2.7

0%

win rate

Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent photographic clarity and lighting
  • + Clean glass reflections and refraction of the plant
  • + Accurate adherence to all spatial instructions
  • The sphere appears slightly large relative to the 'small' descriptor
  • Table surface is a bit generic compared to model b

Wan 2.7

  • + Highly realistic wood texture and weathered table details
  • + Beautiful soft lighting and color grading
  • + Detailed book binding and pages
  • Physical logic errors with an extra vertical glass pane inside the cube
  • Ghosting artifact of the blue sphere on the left side
  • The 'green plant behind' instruction is less clear due to overlapping composition

Verdict: GPT Image 1.5 is the winner because it provides a logically consistent scene that follows all prompt instructions perfectly, whereas Wan 2.7 introduces strange internal glass geometry and a phantom sphere artifact. While Wan 2.7 has superior textures and a more authentic vintage aesthetic, GPT Image 1.5's better understanding of 3D space and refraction makes it more successful.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Successfully preserved the specific man's facial features and hair from the source image.
  • + Accurately rendered California coastline features like palm trees and Highway 1 scenery.
  • + Maintained the interior details and color of the car well.
  • The car is RHD (right-hand drive) despite being in California, and he is sitting in the passenger seat relative to the road side.
  • The crop cuts off most of the car, focusing heavily on the driver.
  • The steering wheel hand looks slightly unnatural.

Wan 2.7

  • + Shows the full car in a dynamic driving shot that matches the prompt.
  • + The lighting on the car and landscape is vibrant and aesthetically pleasing.
  • + Accurately captured the cliffside road composition.
  • Completely failed to use the man from the source image, replacing him with a distorted, generic face.
  • The driver is poorly rendered and appears to have anatomical artifacts.
  • Lost the specific identity provided in the source material.

Verdict: GPT Image 1.5 is the clear winner for its success in identity preservation, accurately placing the man from the source image into the car. While Wan 2.7 created a better full-car composition, it completely failed the image editing task by ignoring the source photo of the man and replacing him with a low-quality, distorted figure.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent shallow depth of field and bokeh
  • + Highly detailed textures of the bicycle, clothing, and wet pavement
  • + Dynamic lighting with vibrant reflections of brake lights
  • The car in the background is sharp rather than having the requested motion blur
  • Composition is a bit tight for a 'candid street photo'

Wan 2.7

  • + Successfully captures a wider street scene with a documentary feel
  • + Realistic urban environment in Japan
  • + Good depiction of wet pavement reflections
  • The subject is holding the bicycle rather than repairing it
  • Fails to provide the requested shallow depth of field and motion blur
  • Image clarity is lower compared to Model A

Verdict: GPT Image 1.5 follows the stylistic prompts much better, providing a cinematic shallow depth of field and rich textures, even though it missed the specific 'motion blur' on the car. Wan 2.7 feels like a generic street photo and ignores the 'repairing' action and the specific lens-style requirements like shallow depth of field.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent depiction of warm torchlight reflections on weather-beaten gold armor.
  • + Highly intricate texture on the underlying cloth and leather straps.
  • The composition is slightly tight, cutting off the top of the head.

Wan 2.7

  • + Strong adherence to the braid and bead requirement with clear separation.
  • + Effective use of shallow depth of field with visible sparks in the background.
  • The armor texture appears a bit smoother and less 'battle-worn' than requested.
  • Lighting lacks the dramatic orange glow expected from a nearby torch.

Verdict: GPT Image 1.5 succeeds with superior lighting and texture, capturing the 'battle-worn' feel through grimy skin and distressed metal. Wan 2.7 provides a cleaner interpretation of the braids and beads but fails to match the atmospheric warmth and material detail of its competitor.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Exceptional text rendering with perfect spelling and logical item descriptions.
  • + Clear categorization with distinct sections for Appetizers, Pizza, and Mains as requested.
  • + High-quality, appetizing food photography that complements the text layout.
  • The layout is a bit more conventional than a 'modern minimalist' grid.
  • The grid of photos is aligned only to the right rather than integrated throughout.

Wan 2.7

  • + Strong aesthetic appeal with a professional-looking grid layout.
  • + Includes atmospheric elements like rosemary and oil for a lifestyle feel.
  • + Excellent use of vibrant accents and branding elements.
  • Poor text quality with many spelling errors and illegible fine print.
  • The food photos contain logic errors and anatomical artifacts upon close inspection.
  • The grid includes items like cake and desserts which were not explicitly in the requested sections.

Verdict: GPT Image 1.5 is the clear winner because it functions as an actual menu with perfect legibility and logical alignment between the text and images. While Wan 2.7 has a more stylish modern layout, its text is riddled with typos ('Calanrfri', 'Sannon') and the small print is entirely illegible, making the design unusable for its intended purpose.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent photorealistic texture on the meat and bun
  • + Very dynamic and chaotic 'exploded' composition
  • + Seamless integration of fiery text into a believable background
  • The 'exploded' effect feels more like a stack than individual pieces flying apart

Wan 2.7

  • + Clean layout with very legible and creative fiery typography
  • + Highly distinct separation of burger components creating a strong sense of motion
  • Illustration style lacks the requested photorealistic detail
  • Odd floating items like seeds/nuts and a cucumber slice that don't fit the 'Magic Burger' theme well

Verdict: GPT Image 1.5 is the winner because it adheres much better to the request for photorealism and a gritty, fiery atmosphere. While Wan 2.7 has cleaner text rendering and a more creative 'exploded' layout, its overall look is too much like a digital illustration or vector art, failing the photorealism requirement.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Text has a very realistic dry chalk texture with smudges.
  • + The handwriting looks authentic and non-digital.
  • + Perfectly follows all text prompts including completing the truncated 'Brown But...' item.
  • The composition is a tight crop showing only the board.
  • Light source is a bit uneven at the top.

Wan 2.7

  • + Excellent environmental context with warm café lighting and background.
  • + High contrast text that is very easy to read.
  • + Accurately rendered all requested text including price details.
  • The text looks like a clean digital font rather than natural chalk.
  • The 'handwriting' is too uniform and lacks the grain/texture requested in the prompt.

Verdict: GPT Image 1.5 followed the stylistic requirements much better, producing incredibly realistic chalk textures and natural handwriting variations. Wan 2.1 created a more aesthetically pleasing scene with a better background, but the text feels like a digital overlay (font) rather than actual chalk on a board.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Successfully integrated the specific accessories (scarf, sunglasses) and clothing from Image 2.
  • + Followed the complex leg pose and arm extension instructions from Image 1.
  • + Adapted the background and lighting style from Image 1 to the new character.
  • The head angle is slightly less dynamic than the original reference.
  • Anatomical issues in the feet, which appear flattened and somewhat distorted.

Wan 2.7

  • + Almost perfectly preserved the composition and visual quality of Image 1.
  • Completely failed the main instruction to use Image 2 as a character reference.
  • No visible changes were made to the person's face, clothing, or gender from the original pose reference.

Verdict: GPT Image 1.5 successfully performed the complex task of recontextualizing the character from Image 2 into the pose of Image 1, effectively carrying over the clothing, face, and accessories. Wan 2.7 failed the prompt entirely, returning what appears to be a slightly modified version of the source Image 1 without incorporating any character details from Image 2.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent dynamic lighting and highly detailed textures on both the horse and spacesuit.
  • + Strong cinematic composition with realistic-looking dust and moon surface effects.
  • + Great attention to detail in the horse's tack and anatomy.
  • Failed the specific spatial instruction; the astronaut is riding the horse instead of the horse being on top.

Wan 2.7

  • + Clean, clear resolution with an interesting surrealist vibe in the background.
  • + Good use of color and balance between the earth and the deep space elements.
  • + Accurately represents an astronaut and horse set in space.
  • Failed the specific spatial instruction; the horse is being ridden by the astronaut.
  • Anatomical issues with the horse's rear right leg looking disconnected and rubbery.

Verdict: Both models failed the negative constraint/spatial instruction to have the 'horse on top', instead defaulting to the traditional 'astronaut riding horse' trope. GPT Image 1.5 is the superior image due to its 훨씬 richer textures, cinematic lighting, and more believeable anatomy compared to the flatter, more artificial look of Wan 2.7.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent transfer of the specific jacket and scarf pattern
  • + Maintains high resolution and realistic fabric textures
  • Crop fails to show the full person, violating the instruction to keep background unchanged
  • Completely cut off the subject's face
  • Changed the lighting and color grade of the overall scene

Wan 2.7

  • + Keeps the full subject, pose, and background intact
  • + Preserves the subject's face and unique features perfectly
  • + Understands the prompt as a full-body outfit replacement task
  • Failed to use the specific outfit from Image 2, generating a generic gold-embroidered coat instead
  • Lower visual fidelity compared to the source image

Verdict: GPT Image 1.5 successfully captured the specific clothing items from Image 2 but failed the core editing task by cropping out the subject's head and changing the image composition. Wan 2.7 followed all structural instructions regarding the subject's identity and pose but failed the visual reference task by replacing the requested outfit with a completely different style. Wan 2.7 is the likely winner because it actually performed the edit on the original person and scene, whereas GPT Image 1.5 basically generated a new image that threw away the source face and framing.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the prompt's specified viewing angle from inside the taxi.
  • + Highly realistic texture on the capybara's fur and the taxi driver's cap.
  • + Perfectly captures the 'bored' expression of the businesswoman in the background.
  • The capybara's paws look more like human-animal hybrid hands with dark fingers.
  • Slightly messy rendering of the taxi's dashboard in the foreground.

Wan 2.7

  • + Strong cinematic lighting and sharp details on the exterior of the vehicle.
  • + Captures the bored businesswoman effectively in the passenger seat.
  • Violates the prompt by placing the passenger in the front seat instead of the back.
  • The viewing angle is from outside the car, whereas the prompt requested a scene from 'inside'.
  • The capybara's fur texture appears stylistically illustrated/rendered rather than photorealistic.

Verdict: GPT Image 1.5 is the clear winner because it correctly follows the complex spatial instructions of the prompt, placing the passenger in the back and viewing the scene from inside the cabin. Wan 2.7 fails on composition by placing the passenger in the front seat and using an exterior camera angle, and its capybara looks significantly less realistic than the one generated by GPT Image 1.5.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent atmospheric lighting and texture that matches the 'vintage gothic' request
  • + Perfect text rendering of all requested details without spelling errors
  • + Strong cinematic composition with a central glowing focal point
  • The thorns and webs are a bit dense, making the border look slightly cluttered

Wan 2.7

  • + Clean layout with clear separation between different elements
  • + Added creative elements like the cauldron and crows that fit the theme
  • + Followed all text instructions accurately including the scroll banner
  • The 'Est. 1847' text was not requested and feels out of place
  • The lighting is flat and more illustrative than 'cinematic'
  • The bright parchment creates a cartoonish rather than 'dark gothic' mood

Verdict: GPT Image 1.5 is the clear winner for its superior atmospheric rendering and adherence to the 'dark parchment' and 'moody' descriptors, creating a cohesive gothic aesthetic. While Wan 2.7 has accurate text, it feels more like a modern illustration than a vintage gothic poster, lacking the depth and cinematic lighting found in GPT Image 1.5.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
GPT Image 1.5
Before After
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent texture that matches the existing beard
  • + Preserves the facial features and glasses near-perfectly
  • The transition from the sideburns to the top of the head feels a bit disconnected
  • Slight change to the shape of the forehead

Wan 2.7

  • + Highly realistic hair texture and styling
  • + Excellent integration of the new hair with the existing facial structure and hairline sides
  • Slightly alters the nose and eye area from the original source image

Verdict: Both models performed very well on this editing task, maintaining the lighting and background of the original. Wan 2.7 provides a more modern and natural-looking hairstyle that integrates seamlessly into the portrait, whereas GPT Image 1.5 offers a texture that better matches the person's existing beard but has a slightly more artificial hairline transition.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to text instructions with perfect placement of 'JAPAN' and 'SUSHI'.
  • + Rich, realistic textures and PBR materials on the wood, ceramic, and food items.
  • + Superior isometric diorama composition with varied heights and interesting secondary objects like the teapot.
  • The composition feels slightly crowded compared to the 'minimal' request.

Wan 2.7

  • + Captures the '3D cartoon' and 'minimal' aesthetic very well with clean, smooth shapes.
  • + Very clean layout on a simple raised diorama base as requested.
  • + Accurate text and flag placement.
  • The sushi textures look a bit like plastic or clay rather than 'realistic PBR materials'.
  • The text is slightly off-center relative to the diorama base.

Verdict: GPT Image 1.5 followed the prompt more effectively by balancing the 'cartoon miniature' look with 'realistic PBR materials', resulting in a much more visually compelling diorama. While Wan 2.7 achieved a cleaner 'cartoon' look, it sacrificed the material realism requested and had less depth in its composition compared to the rich textures found in GPT Image 1.5.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent caricature style that maintains a strong likeness to the source person.
  • + Clearly incorporates all three prompt elements: news desk, multiple dogs, and hockey in the background.
  • + High-quality rendering with professional lighting and legible text.
  • The fingers on the hand holding the papers are slightly merged and anatomicaly awkward.

Wan 2.7

  • + Creative use of speech bubbles and a tiny hockey rink prop.
  • + Successfully preserves the 'selfie' pose from the original source image.
  • + Includes multiple dogs with fun hockey-themed accessories.
  • The caricature style is a bit more generic and loses some of the subject's unique facial features.
  • The 'TV anchor' part of the profession is less clear, looking more like a living room with a TV.
  • Text in the speech bubbles contains spelling errors like 'Rolee'.

Verdict: GPT Image 1.5 performed better by creating a more recognizable caricature that perfectly captured the professional setting of a TV anchor while seamlessly blending the hobby elements. Wan 2.7 maintained the original selfie pose better, but the facial likeness was weaker and the image contained more artifacts and spelling errors.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent depiction of texture, particularly the soft fur on all animals.
  • + Dynamic, playful composition that captures the 'tumbling together' aspect of the prompt.
  • + Stunning lighting effects with realistic god rays and integrated dew sparkles.
  • The fox's paw in the bottom right corner is slightly malformed.
  • Some of the pink flowers in the foreground exhibit minor digital smudging.

Wan 2.7

  • + Clear individual subjects with good spatial separation.
  • + Accurately includes all four requested species with distinct features.
  • + Good rendering of the wildflower meadow and golden hour lighting.
  • The animals feel somewhat static and posed rather than 'tumbling' or 'playfully chasing'.
  • The kitten has a strange, small tongue artifact.
  • The lighting feels more like a generic filter compared to the organic warmth in Model A.

Verdict: GPT Image 1.5 is the clear winner for its superior ability to capture the energy and texture requested in the prompt. While Wan 2.7 provides a clear layout, it lacks the 'masterpiece' quality and dynamic interaction of the animals seen in GPT Image 1.5, which excels at the atmospheric lighting and soft fur details.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'dreamy' and 'warm nostalgic' mood through soft lighting and bloom effects.
  • + Successfully adopts a modern anime/Ghibli-inspired character design style.
  • + Perfectly captures the pastel color palette requested in the prompt.
  • The faces, while stylized, lose much of the original actors' likenesses.
  • The heavy bloom effect makes the foreground figure quite blurry compared to the source.

Wan 2.7

  • + Outstanding source preservation, maintaining the recognizable facial structures of the original meme actors.
  • + Accurate hand-painted watercolor texture that feels very authentic to classic animation backgrounds.
  • + Maintains the composition and clarity of the source image perfectly.
  • The style leans more toward a western watercolor sketch than the specific Ghibli anime aesthetic.
  • The lighting is flat compared to the requested 'gentle, dreamy' atmosphere.

Verdict: GPT Image 1.5 better captures the emotional and stylistic vibe of a Ghibli film with its warm, glowing lighting and rounded character designs, though it loses more of the original image's detail. Wan 2.7 is technically superior at preserving the source image and applying a beautiful watercolor texture, but it feels less like an anime and more like a comic illustration. GPT Image 1.5 is preferred for its superior adherence to the 'dreamy' mood and 'Ghibli' aesthetic instructions.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
GPT Image 1.5
Before After
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the 'hair blowing in wind' instruction with a very dynamic shape.
  • + Adds a significant amount of leaves to create the requested energetic and lively feel.
  • + High source preservation with minimal structural changes to the original composition.
  • The hair flow feels slightly symmetrical and artificial.
  • Some leaves are blurry or lack texture compared to the rest of the image.

Wan 2.7

  • + Natural-looking hair motion that integrates well with the original hairline.
  • + Leaves feel naturally integrated into the environment with realistic orientation.
  • + Excellent source preservation of the woman and dog's features.
  • The 'energetic and lively' feel is more subtle compared to the other model.
  • Fewer flying leaves added than might be expected from the prompt.

Verdict: Both models performed excellent image-to-image editing, maintaining nearly all details of the source image. GPT Image 1.5 is the winner for better following the 'energetic' and 'lively' part of the prompt by adding more motion to the hair and a higher volume of flying leaves, whereas Wan 2.7 opted for a more subtle, realistic approach.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent typography with a hand-drawn vintage feel
  • + High-quality vector lighting and texture effects on the cloche
  • + Perfect text accuracy including the accent mark in 'Caffè'
  • Failed to provide a light background as requested
  • Minimalist style is sacrificed for a more illustrative look

Wan 2.7

  • + Successfully used a light background with subtle texture
  • + Followed the vector emblem layout very strictly
  • + Good use of the 'Est. 1720' banner within a circular frame
  • Spelling error in the main name ('Florion' instead of 'Florian')
  • Cloche dome illustration is very basic and lacks the requested 'steam' detail (shows abstract line work)
  • Typography is less elegant and feels more generic

Verdict: GPT Image 1.5 produced a much more professional-looking logo with superior typography and text accuracy, though it ignored the light background instruction. Wan 2.7 followed the layout and background prompts better but failed on the critical task of spelling the brand name correctly and provided a much simpler illustration. GPT Image 1.5 is the preferred choice for its artistic quality and correct spelling.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1.5
Wan 2.7

AI Judge Analysis

GPT Image 1.5

  • + Excellent adherence to the NASA-inspired color palette and flat-vector style.
  • + Text rendering is clear, legible, and large enough to be functional.
  • + Strong layout with a logical 1-6 panel flow that fills the frame.
  • The 'Lunar Orbit' moon icon changes color slightly compared to the Earth-background moon.
  • Missing some of the technical detail requested in the 'supporting information' part of the prompt.

Wan 2.7

  • + More sophisticated vertical infographic composition with a higher information density.
  • + Includes specific technical data like UTC times and altitudes for each step.
  • + Features a cleaner, more professional header and footer typical of poster designs.
  • Poor spelling in several places, including 'Descript' for Descent and 'Tranquiliry'.
  • Icons are significantly smaller, making the visual narrative harder to read at a glance.
  • Small artifacts in the star field and some text alignment issues.

Verdict: GPT Image 1.5 followed the stylistic and sequential instructions perfectly, creating a highly legible and visually balanced set of panels with accurate text. While Wan 2.7 attempted a more complex 'poster' layout with impressive supporting data, it suffered from several spelling errors and reduced icon clarity that lowered its overall quality as an infographic.

Next steps

Explore each model