Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1 OpenAI Wan 2.7 Pro Alibaba

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

GPT Image 1

23.2 arena score

#28 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.7 Pro

20.7 arena score

#38 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1

100.0%

win rate

Ties

0.0%

Wan 2.7 Pro

0.0%

win rate

100.0% 0.0% ties 0.0%
Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent photographic clarity and clean lines.
  • + Perfect adherence to spatial instructions like the plant being behind the cube.
  • + Accurate lighting and shadow effects.
  • The glass cube looks slightly like an hollow frame rather than a solid-walled object in some areas.

Wan 2.7 Pro

  • + Naturalistic texture on the wooden table and red book.
  • + Realistic reflections on the base of the glass cube.
  • + Nice soft window lighting as requested.
  • The plant appears to be growing through or from inside the glass cube rather than being behind it.
  • The cube's geometry is slightly inconsistent in the back corners.

Verdict: GPT Image 1 followed the instructions more accurately, particularly regarding the spatial relationship of the plant being behind the glass cube. While Wan 2.7 Pro provided a more textured and 'authentic' looking table and book, it failed by merging the plant into the cube's interior space.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent preservation of the specific man's face and unique hairstyle
  • + Realistic motion blur on the wheels and road indicating driving
  • + High fidelity maintenance of the car's design and details

Wan 2.7 Pro

  • + Beautiful composition and lighting for a California coastline setting
  • + Good preservation of the car's body and wheels
  • Completely failed to use the man from the source image, replacing him with a generic individual
  • The car appears to be parked or floating slightly rather than driving due to the lack of motion blur

Verdict: GPT Image 1 successfully followed the editing instructions by combining both source images, maintaining the likeness of the man and his hairstyle while placing him convincingly behind the wheel. Wan 2.7 Pro completely ignored the second source image, replacing the subject with a random man, and failed to capture the sense of 'driving' seen in the motion of the first model's output.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent shallow depth of field and bokeh realism.
  • + Natural, highly detailed skin texture on the subject.
  • + Strong cinematic atmosphere with realistic light and rain effects.
  • The bike anatomy is slightly jumbled near the rear hub.
  • Lacks the requested motion blur on the background vehicles.

Wan 2.7 Pro

  • + Clearer street context and reflections on the pavement.
  • + Captures the entire subject and bicycle for better situational context.
  • + Finer rain rendering across the frame.
  • Static vehicles lack the requested motion blur.
  • Skin textures are smoother and less realistic than the competitor.
  • Mechanical bicycle details like the chain and spokes are messy upon close inspection.

Verdict: GPT Image 1 is the preferred output because it better captures the requested photographic style, featuring a much more convincing 50mm shallow depth of field and realistic skin textures. While Wan 2.1 Pro provides a more complete view of the bicycle, its lighting and subject texture feel more like a digital render compared to the cinematic quality of GPT Image 1.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Exceptional realism in the skin texture, showing pores, fine wrinkles, and authentic battle grime.
  • + The engraving on the armor is incredibly intricate and looks historically inspired.
  • + Moody and effective lighting that enhances the somber paladin theme.
  • The hair beads requested in the prompt are very subtle and almost easy to miss.
  • Leather straps requested in the prompt are not clearly visible in this framing.

Wan 2.7 Pro

  • + Successfully incorporates all prompt elements including multiple braids with distinct beads and visible leather straps.
  • + Captures the 'bokeh sparks' and torchlight effect very literally and clearly.
  • + Good metallic luster and clean engraving details on the plate armor.
  • The skin rendering is a bit smoother and less lifelike compared to Model A.
  • The facial expression is somewhat stiff and lacks the emotional depth of a 'battle-worn' character.

Verdict: GPT Image 1 produces a more cinematic and emotionally resonant portrait with superior facial textures, perfectly capturing the gritty feel of a battle-worn warrior. Wan 2.7 Pro is more thorough in including every specific item from the prompt, such as the colorful beads and leather straps, but falls slightly behind in sheer visual realism and atmosphere.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent font legibility for header text
  • + High-quality, vibrant food photography
  • + Clean minimalist aesthetic that focuses on the dishes
  • Nonsense filler text for descriptions
  • Incomplete layout that lacks a header or branding
  • Poorly aligned grid where text overlaps or is missing for the bottom row

Wan 2.7 Pro

  • + Complete, professional brand identity with logo, address, and social icons
  • + Perfect adherence to the 'grid' requirement with 12 distinct items
  • + Functional and realistic hierarchy including category tabs and a QR code
  • Smaller text is blurry and difficult to read
  • Slightly more cluttered than a strict 'minimalist' prompt might suggest
  • Minor spelling errors in menu titles like 'Roasted Tomato Bissue'

Verdict: Wan 2.7 Pro produces a complete, usable graphic design layout that includes all requested sections, branding, and a comprehensive grid, though the smaller text loses clarity. GPT Image 1 creates beautiful food photography and clear headings but fails to provide a finished menu design, leaving the bottom half of the image as just photos without a corresponding text layout. Wan 2.7 Pro is the superior choice for a professional marketing asset.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent photorealistic texture on the meat patty and toasted bun
  • + Text is perfectly integrated with a fiery, glowing effect as requested
  • + Composition feels balanced and professional for an advertisement
  • Failed the price accuracy, displaying '€.99' instead of '€6.99'
  • The 'exploded' effect is more of a stack than a dynamic dispersal

Wan 2.7 Pro

  • + Successfully included all requested text accurately, including the price
  • + Highly dynamic composition with ingredients flying outward
  • + Good use of extra elements like steam and falling vegetables to enhance the theme
  • The meat patty and bottom bun look slightly less realistic than Model A
  • Text effects are more of a standard gradient rather than the requested 'fiery glowing' style seen in the competitor

Verdict: Both models followed the prompt well, but Wan 2.1 Pro is the winner because it accurately rendered the price and captured the 'dynamic, exploded' movement of the ingredients much better than GPT. While GPT Image 1 had superior food textures and better glowing text effects, the missing digit in the price is a significant error for an advertisement task.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent chalk texture with visible grain and powdery edges.
  • + High accuracy in rendering all requested text including the specific date and phrases.
  • + Natural variations in letter size and spacing that mimic real handwriting.
  • Failed to provide the 'elegant cursive' style for the title as requested.
  • The tight framing cuts off context of the 'cozy café' environment.

Wan 2.7 Pro

  • + Successfully incorporated the 'cozy café' atmosphere with lighting and background elements.
  • + Followed the layout instructions well, including horizontal dividers for a clean composition.
  • + Text is highly legible and maintains a consistent style throughout.
  • The lettering looks more like a digital font or vector art than high-texture chalk.
  • The title is not in 'elegant cursive' but rather a stylized print.

Verdict: GPT Image 1 excels in the technical execution of the chalk texture and realistic handwriting imperfections, capturing the 'feel' of chalk more accurately. However, Wan 2.7 Pro provides a much better overall composition by showing the café environment, although its text looks slightly too clean and digital compared to the requested chalk style. GPT Image 1 is the winner for its superior prompt adherence regarding the physical properties of the handwriting.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Successfully integrated clothing elements like the scarf and black sweatshirt from Image 2
  • + Accurately replicated the difficult pose from Image 1
  • + Maintained the character identity including sunglasses and facial features
  • Anatomy in the lower legs is slightly distorted compared to the source
  • Finger rendering on the left hand is noticeably low quality

Wan 2.7 Pro

  • + Perfectly preserved the original background and lighting of Image 1
  • Completely failed to perform the edit, returning nearly the identical image as Image 1
  • No character traits from Image 2 were incorporated

Verdict: GPT Image 1 followed the complex multi-image instruction by successfully mapping the character from Image 2 (including accessories like the scarf and sunglasses) onto the pose from Image 1. Wan 2.7 Pro failed the task entirely, simply returning the first source image without any modifications.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent cinematic lighting and atmosphere
  • + High level of texture detail on the astronaut suit and horse's coat
  • + Realistic proportions and dynamic posing
  • The horse's front legs and hooves have anatomical distortions
  • Fails the specific prompt instruction to have the horse on top of the astronaut

Wan 2.7 Pro

  • + Clean, sharp image with bright colors
  • + Good composition with a clear view of the planet below
  • + The astronaut and saddle are well-integrated
  • Fails the specific prompt instruction to have the horse on top of the astronaut
  • The background contains many tiny, floating, repetitive planets that look cluttered
  • The horse's anatomy is slightly elongated and awkward

Verdict: Both GPT Image 1 and Wan 2.7 Pro failed the negative constraint to have the 'horse on top', instead providing standard images of an astronaut riding a horse. GPT Image 1 is the preferred image because its cinematic lighting, atmosphere, and detailed textures create a much more convincing surreal scene compared to the cluttered and flatter appearance of Wan 2.7 Pro.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Perfectly replicates the specific clothing, scarf, and watch from Image 2
  • + Maintains the person's facial features and distinctive vitiligo patterns with high accuracy
  • + Seamlessly integrates the clothing into the lighting and pose of the scene

Wan 2.7 Pro

  • + Preserves the full-body composition and background of the original image
  • + Maintains the vitiligo markings on both the face and hands consistently
  • Completely failed to use the outfit from Image 2, generating a generic gold-patterned jacket instead
  • The scale of the person relative to the beach structure is slightly altered

Verdict: GPT Image 1 followed the instructions nearly perfectly, accurately transplanting the exact coat, plaid scarf, and watch from Image 2 onto the subject while keeping the subject's face and hair unchanged. Wan 2.7 Pro failed the core task of the prompt by generating an entirely different outfit that was not present in any of the source images.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent texture on the capybara's fur and the leather hat
  • + Strong adherence to the 'professional expression' requested
  • + Realistic bokeh and color grading that matches a New York night
  • The capybara's 'hands' are somewhat distorted and lack anatomical realism compared to its head
  • The framing of the taxi's exterior roof light is slightly awkward

Wan 2.7 Pro

  • + Dynamic composition that shows more of the taxi interior and the city environment
  • + Very high level of detail on the passenger's face and phone
  • + Clearer representation of the front paws on the steering wheel
  • The passenger is sitting in the front seat instead of the back seat as requested
  • The scale of the capybara's neck is slightly unnatural compared to its body

Verdict: GPT Image 1 followed the spatial instructions more accurately by placing the passenger in the back seat, whereas Wan 2.7 Pro placed her in the front. However, Wan 2.7 Pro offered a more vibrant and detailed scene with better clarity on the streetlights and steering wheel interaction. Image A is preferred for strict prompt adherence, while Image B is superior in technical composition and lighting.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent text rendering with no spelling errors
  • + Accurate moody and gothic atmosphere with cinematic lighting
  • + Perfect adherence to the dark parchment and thorn border request
  • The parchment looks more like a dark texture than an aged page
  • The jack-o-lantern is a bit small within the overall composition

Wan 2.7 Pro

  • + Very creative and detailed illustration with extra elements like lanterns and a cauldron
  • + Highly polished, clean art style with a beautiful moody night sky
  • + Clear hierarchy of information
  • Missed the 'dark parchment' request, opting for a lighter, cleaner paper look
  • Added 'Est. 1847' and 'Dress code' which were not in the prompt

Verdict: GPT Image 1 followed the atmospheric requirements more closely, delivering a brooding, vintage aesthetic that perfectly matched the 'dark parchment' and gothic tone. While Wan 2.7 Pro produced a more vibrantly detailed and polished illustration, it deviated from the request by using a lighter color palette and adding extraneous text information.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
GPT Image 1
Before After
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent preservation of the original image background and clothing
  • + Dramatic change in hair volume as requested
  • The hair looks like an obvious overlay or a wig
  • The hairline is very sharp and unnatural
  • Changes the facial structure and forehead shape significantly

Wan 2.7 Pro

  • + Natural-looking hair texture and realistic hairline integration
  • + Maintains the original facial features and head shape almost perfectly
  • + Excellent source preservation including lighting and background
  • The hair provides slightly less 'fullness' than typical for a 'thick head of hair' request, appearing more medium-density

Verdict: Wan 2.7 Pro produces a significantly more believable result by seamlessly integrating the new hair with the existing skin textures and lighting, whereas GPT Image 1 creates a 'wig' effect that looks disconnected from the person's face. Wan 2.7 Pro also does a better job of preserving the original proportions of the man's face despite the addition of hair.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Perfectly follows the requested text placement and hierarchy.
  • + Excellent miniature 3D aesthetic with soft, clean textures true to the 'cartoon scene' prompt.
  • + Centralizes the diorama base effectively for a cleaner composition.
  • The sushi textures are a bit overly simplified, looking more like plastic toy pieces than PBR materials.
  • The flag icon is a bit large compared to the refined balance of the text.

Wan 2.7 Pro

  • + Sophisticated rendering with superior PBR materials that show realistic sheen and translucency.
  • + High level of detail in the sushi variety and garnishes.
  • + Clean, professional typography and graphic design elements.
  • Includes additional text 'Authentic Japanese Cuisine' which was not in the prompt.
  • Placement of the flag icon is off-center to the right rather than 'top-center'.
  • Composition feels slightly more cluttered with scattered elements outside the diorama base.

Verdict: GPT Image 1 followed the layout instructions more accurately, particularly regarding the text placement and the specific 'cartoon scene' style requested. However, Wan 2.7 Pro produced a significantly more visually impressive result with high-quality PBR materials and professional lighting, despite adding extra text and deviating slightly from the centered layout.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent caricature style with exaggerated facial features that still resemble the source person.
  • + Clever integration of all prompt elements including the news anchor desk, the dog co-anchor, and the hockey mascot on screen.
  • + Consistent clothing style preservation from the source image.

Wan 2.7 Pro

  • + Successfully incorporates multiple dogs in hockey jerseys.
  • + Clear, high-contrast illustration style with legible text elements.
  • The facial likeness to the source person is significantly weaker than Model A.
  • The 'hockey stick' held by the character is an incoherent hybrid of a microphone and a stick.
  • Relatively generic composition with speech bubbles that feel a bit literal.

Verdict: GPT Image 1 (Model A) is the clear winner as it delivers a high-quality caricature that maintains a recognizable likeness while effectively using exaggeration. It masterfully blends the requested elements - news desk, dogs, and hockey - into a cohesive scene. Wan 2.7 Pro (Model B) feels more like a generic cartoon and struggles with the object logic of the hockey-microphone hybrid.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Captures a dynamic sense of motion with the 'tumbling' and 'chasing' aspect of the prompt.
  • + Very strong lighting effects with soft god rays and consistent golden hour warmth.
  • + Excellent fur texture and expressive, cute facial features on all four animals.
  • The kitten's front paw positioning is slightly anatomically awkward as it interacts with the bunny.
  • Some butterflies in the background are very blurry and lose their shape.

Wan 2.7 Pro

  • + Great variety in flower colors and types, enhancing the 'lush wildflower meadow' look.
  • + Good inclusion of all four animals with distinct, recognizable identities.
  • + The lighting highlights on the grass blades create a nice dew-sparkle effect.
  • The kitten is significantly smaller relative to the other animals than is realistic.
  • The animals appear more static and posed rather than 'tumbling together' as requested.
  • The fox's face lacks the 'baby' kit features, appearing more like a small adult fox.

Verdict: GPT Image 1 is the superior image because it better captures the joyful energy and interaction requested in the prompt; the animals actually appear to be playing and tumbling together. While Wan 2.7 Pro has nice floral variety, its composition feels more like four separate animals placed in a scene rather than a cohesive, playful moment.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Successfully captures the specific Studio Ghibli character design aesthetic with simple, soft facial features.
  • + Applies a consistent, warm, nostalgic color palette across the entire frame.
  • + Excellent preservation of the source image's composition and poses.
  • The hand-painted texture results in a slightly grainy, lower-resolution appearance.
  • Some fine details from the clothing, like the specific plaid pattern, are simplified too much.

Wan 2.7 Pro

  • + Maintains high resolution and clear textures while incorporating watercolor elements.
  • + Preserves the identity and likeness of the original people more closely than the competitor.
  • + Accurately replicates the complex plaid pattern of the man's shirt in the new style.
  • The facial style leans more towards generic Western watercolor illustration rather than the specific Studio Ghibli style requested.
  • The lighting feels a bit harsh and lacks the 'dreamy' atmosphere requested in the prompt.

Verdict: GPT Image 1 followed the stylistic instructions more faithfully, successfully translating the subjects into the specific, iconic character designs associated with Studio Ghibli. While Wan 2.7 Pro produced a higher quality watercolor illustration that preserved more detail from the source image, it failed to capture the specific 'Ghibli' look, opting instead for a more literal painterly filter of the original faces.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
GPT Image 1
Before After
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Successfully added wind-blown hair and flying leaves.
  • + Maintained the subject's facial identity and body structure well.
  • + Added dynamic movement to the dog's tail.
  • The leash has been altered into an unnatural, stiff loop shape.
  • The falling leaves are somewhat small and uniform in color.

Wan 2.7 Pro

  • + Excellent hair dynamics with realistic strands catching the light.
  • + The leaves have varied sizes and textures, enhancing the sense of motion.
  • + Preserved the original shape of the leash much better than the competitor.
  • The dog's tail remains static compared to the movement in the rest of the image.
  • Slight alteration to the lighting/glow on the left side of the image.

Verdict: Both models did an excellent job of following the instructions while preserving the source image. Wan 2.1 Pro is the winner because it handled the hair and leaves with more realistic texture and lighting, and crucially, it did not distort the leash into an unnatural shape like GPT Image 1 did.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1

  • + Excellent adherence to high-contrast minimalist vector style
  • + Accurate spelling of the brand name including the accent over the 'e'
  • + Strong grain texture that fits the vintage request
  • Ignored the 'light background' instruction by using a black background
  • Steam element is very simple and looks a bit like a stray hair

Wan 2.7 Pro

  • + Successfully used warm brown and cream tones on a light background
  • + Elegant vintage ornamental composition with wreath and banner
  • + Clear and stylish rendering of 'Est. 1720' text
  • Spelling error in the main brand name: 'Florion' instead of 'Florian'
  • Composition is less minimalist than requested with many extra peripheral elements

Verdict: GPT Image 1 followed the minimalist vector emblem style and spelling instructions perfectly, though it failed the background color requirement. Wan 2.7 Pro created a beautiful vintage composition that matched the requested color palette but suffered from a critical spelling error in the name 'Florian'. GPT Image 1 is the winner for its technical accuracy in text rendering and stylistic adherence.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1
Wan 2.7 Pro
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1

  • + Strong minimalist aesthetic with bold icons.
  • + Correctly identifies the three astronauts by name.
  • + Excellent utilization of the requested NASA-inspired color palette.
  • Several spelling errors including 'EARLLUNAR' and inconsistent label placement.
  • The layout is cluttered and the flow between steps is non-linear and confusing.

Wan 2.7 Pro

  • + Perfect logical flow that follows the requested 6-step mission sequence chronologically.
  • + Sophisticated layout that truly feels like a modern infographic poster.
  • + Impressive attention to detail with supporting text and mission data.
  • Minor text rendering artifacts such as 'DESCRIPT' instead of Descent.
  • Text is quite small and may be difficult to read without zooming.

Verdict: Wan 2.7 Pro is the clear winner as it successfully interprets the request for a 'vector infographic' with a logical flow and comprehensive iconography for all six steps. While GPT Image 1 has clean vector graphics, its layout is jumbled and it contains significant spelling errors like 'EARLLUNAR' which detract from its utility as an infographic.

Next steps

Explore each model