Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1 Mini OpenAI Wan 2.7 Pro Alibaba

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

GPT Image 1 Mini

24.9 arena score

#13 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.7 Pro

20.7 arena score

#38 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1 Mini

50.0%

win rate

Ties

0.0%

Wan 2.7 Pro

50.0%

win rate

50.0% 0.0% ties 50.0%
Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent adherence to the 'soft window light from the left' instruction.
  • + Very clean and minimalist composition with accurate depth of field.
  • + Photorealistic textures on the wooden table and red book.
  • The glass cube is slightly open at the top rather than being a solid closed cube.
  • The green plant is quite blurred in the background, making it less distinct through the glass.

Wan 2.7 Pro

  • + Highly detailed glass rendering including reflections on the base of the cube.
  • + The plant is very clearly visible and positioned directly behind the cube as requested.
  • + Beautifully aged texture on the wooden table.
  • The glass cube has illogical vertical panes within its structure.
  • The red book looks slightly distorted on the right side.

Verdict: Both models followed the prompt instructions perfectly in terms of object placement. Wan 2.7 Pro offered a more rich and detailed scene with a more clearly defined plant and better surface reflections, while GPT Image 1 Mini provided a cleaner, more focused image with superior lighting consistency. Wan 2.7 Pro is the likely winner for its more intricate rendering of the refraction and the background elements.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent preservation of the subject's face, hair, and unique clothing from the source image.
  • + Accurately places the specific car from the source image into a new environment.
  • + Cinematic lighting and motion blur create a realistic sense of driving.
  • The car model is slightly modernized/altered from the source (different headlights).

Wan 2.7 Pro

  • + Successfully places the car on a coastal road with high visual quality.
  • + Retains the exact car model and features from the source image.
  • Completely fails to include the man from the source image, replacing him with a generic figure.
  • The scale of the driver relative to the car seems a bit small.

Verdict: GPT Image 1 Mini is the clear winner as it successfully combines both source images by placing the specific man, wearing his unique outfit, into the driver's seat of the car. Wan 2.7 Pro fails the core edit instruction by replacing the source subject with a random person, despite doing a good job of placing the car in the new environment.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent skin texture and realistic depth of field
  • + Accurate depiction of light rain and micro-reflections on surfaces
  • + Great emotional depth and candid feel in the close-up composition
  • The bike anatomy is slightly illogical with the chain area
  • Lack of visible Japanese street context in the background

Wan 2.7 Pro

  • + Stronger adherence to the Japanese street environment setting
  • + Full-body composition shows the entire red bicycle clearly
  • + Excellent wet pavement reflections including the vehicle headlights
  • The man's scale relative to the bicycle is slightly too small
  • The face and hands have a slightly smooth, less detailed texture compared to the other model

Verdict: GPT Image 1 Mini provides a much more intimate and high-fidelity portrait with superior skin textures and bokeh, capturing the 'candid' and 'natural' aspects of the prompt better. However, Wan 2.7 Pro better captures the environmental requirements, including the specific Japanese street setting and the full red bicycle, though it suffers from minor scaling issues with the figure.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent warm lighting that authentically reflects off the armor
  • + Highly cinematic shallow depth of field with great lens bokeh
  • + Intricate and realistic engraving on the plate mail
  • Missed the prompt requirement for beads in the hair
  • The leather strap detail is less prominent and detailed than in Model B

Wan 2.7 Pro

  • + Perfectly adhered to the requirement for small beads in the braided hair
  • + Exceptional texture and detail on the leather straps and buckles
  • + Clearer depiction of 'battle-worn' scars and dirt on the skin
  • The lighting feels slightly more artificial and less moody than Model A
  • The armor's metal texture looks a bit more like polished steel rather than heavy plate

Verdict: GPT Image 1 Mini creates a more atmospheric and cinematic image with superior lighting, but fails to include the requested beads in the hair. Wan 2.7 Pro adheres more closely to every specific detail of the prompt, particularly the beads and leather textures, making it the better literal interpretation.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Perfect text legibility for main headers
  • + High-quality, clean food photography in a structured grid
  • + Excellent adherence to the minimalist prompt
  • Lack of actual menu items/descriptions makes it look like a template rather than a finished menu
  • The layout feels a bit empty on the left side

Wan 2.7 Pro

  • + Complete, functional menu layout with descriptions and prices
  • + Sophisticated professional aesthetic suitable for a real restaurant
  • + Includes realistic branding elements like a QR code and address
  • Small body text contains significant 'lorem ipsum' style typos and gibberish
  • The grid layout is slightly less clear than Model A's simple alignment

Verdict: GPT Image 1 Mini creates a very clean, bold template that perfectly follows the minimalist instruction, though it lacks depth. Wan 2.7 Pro creates a much more comprehensive and realistic menu design with specific items and pricing, although it suffers from typical AI text hallucinations in the fine print. Wan 2.7 Pro is the preferred choice for a professional context because it demonstrates a full graphic design composition rather than just a header and a grid.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent photorealistic texture on the bun and patty.
  • + The fiery glowing text effect is perfectly integrated into the dark atmosphere.
  • + Strong sense of vertical composition and natural food physics.
  • The 'exploded' effect is a bit static compared to the dynamic scattering of elements.

Wan 2.7 Pro

  • + Highly dynamic 'exploded' composition with many flying ingredients.
  • + Excellent text rendering and layout for a commercial ad.
  • + Captures the 'fiery' background prompt with more literal flaming elements and smoke.
  • The cheese looks somewhat artificial and flat compared to the other ingredients.
  • Some ingredients like the pickles and parsley were not explicitly requested but added.

Verdict: GPT Image 1 Mini excels in photorealism and atmospheric lighting, creating a moodier and more premium feel. Wan 2.7 Pro better captures the 'dynamic' and 'exploded' aspect of the prompt with a more energetic, commercial layout. While GPT Image 1 Mini has superior textures, Wan 2.7 Pro feels more like a complete advertisement design.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1 Mini
Wan 2.7 Pro
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent chalk texture throughout the lettering
  • + Very realistic and natural handwriting variations
  • + Perfect text accuracy and spelling
  • The title is in print-style block letters rather than the requested elegant cursive
  • The background context of a 'cozy café' is barely visible beyond the board frame

Wan 2.7 Pro

  • + Beautiful 'cozy café' background with nice lighting and depth
  • + Highly consistent font/handwriting style
  • + Good composition with decorative flourishes
  • Repeated the phrase '& Herbs - $28' twice on the second item
  • The text looks more like a digital font overlay than real chalk on a board
  • Failed to render the title in cursive as requested

Verdict: GPT Image 1 Mini wins on realism and technical accuracy, providing a convincing chalk texture and perfect spelling. While Wan 2.7 Pro creates a much more appealing 'café' environment, its text contains a repetitive error and lacks the authentic grainy texture of real chalk specified in the prompt.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Successfully transferred the person, sunglasses, scarf, and clothing from Image 2.
  • + Matches the yellow background and red stool environment accurately.
  • + Good facial recognition and resemblance to the character reference.
  • Failed to match the 'exact' complex pose, simplifying the leg cross into a basic step.
  • The left foot (on the stool) is anatomically distorted with too many toes.
  • Changed the character's expression slightly to a smirk instead of the neutral look requested.

Wan 2.7 Pro

  • + Perfectly preserved the original background, pose, and composition from Image 1.
  • Completely failed to perform the edit, ignoring the character reference (Image 2) entirely.
  • Just returned a cropped/modified version of the source Image 1.

Verdict: GPT Image 1 Mini followed the complex instructions by combining the character's appearance from Image 2 with the environment of Image 1, though it struggled with the precise anatomy and the exact complexity of the dancer's pose. Wan 2.7 Pro failed the task entirely, simply returning the source image without incorporating the character reference.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent cinematic lighting and atmosphere
  • + The horse's mane and tail exhibit realistic movement for a vacuum/low-gravity environment
  • + High textural detail on the space suit and horse hide
  • The astronaut's hand placement on the reins is slightly awkward

Wan 2.7 Pro

  • + Bright, clear colors and sharp detail
  • + Detailed background featuring multiple planets and the Earth's atmosphere
  • Composition feels less cinematic and more like a collage
  • The horse's legs and hooves appear somewhat stiff and unnaturally positioned

Verdict: Both models failed the negative constraint to have the horse on top of the astronaut (not vice versa), instead choosing the typical interpretation. GPT Image 1 Mini is preferred because its cinematic lighting, moody atmosphere, and cohesive composition better capture the 'surreal' and 'cinematic' aspects of the prompt compared to the flatter, more literal rendering of Wan 2.7 Pro.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
GPT Image 1 Mini
Wan 2.7 Pro
0% wins 0% ties 100% wins

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent adherence to the specific outfit in Image 2, including the exact scarf pattern and coat style.
  • + Higher image quality and resolution with realistic fabric textures.
  • + Very accurate lighting and shadow integration that matches the beach environment.
  • Significantly changed the person's facial structure and hair, making them look older and different from the source image.
  • Lost the 'Inersy' branding details on the waistband and the specific vitiligo patterns were altered.

Wan 2.7 Pro

  • + Perfectly preserved the person's exact face, hair, and original vitiligo patterns on the skin.
  • + Maintained the full composition of the source image including the wooden structure and background 1:1.
  • Completely failed to use the outfit from Image 2, instead generating a random patterned robe/suit.
  • The hands have minor anatomical distortions and the clothing fit looks slightly unnatural at the waist.

Verdict: This is a trade-off between outfit accuracy and identity preservation. GPT Image 1 Mini correctly identified and adapted the complex outfit from Image 2 but failed to keep the base person's face and hair unchanged. Wan 2.7 Pro perfectly preserved the person and the environment but completely ignored the instruction to use the specific outfit from Image 2, making GPT Image 1 Mini a better choice for an edit task.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent atmospheric lighting and photorealism
  • + Very accurate rendering of the capybara's fur and expression
  • + Stronger focus on the requested background element of the woman looking at her phone
  • One of the capybara's paws is missing/not visible on the wheel as requested
  • The woman's hand holding the phone looks slightly distorted

Wan 2.7 Pro

  • + Better adherence to the 'both front paws on the steering wheel' instruction
  • + Clearer depiction of the New York City street through the window
  • + Shows more of the taxi's exterior and interior for context
  • The woman in the back is not looking at a phone as requested
  • The scale of the capybara and the interior layout feels slightly unnatural

Verdict: GPT Image 1 Mini creates a more moody and cinematically realistic image that captures the specific 'bored' interaction with the phone, though it fails to place both paws on the wheel. Wan 2.7 Pro succeeds in the physical positioning of the capybara but misses the key prompt detail of the passenger looking at her phone. GPT Image 1 Mini is the narrow winner due to its superior lighting, texture, and emotional adherence to the prompt's tone.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent moody, cinematic lighting that fits the 'vintage gothic' request
  • + Perfect text accuracy for all required fields
  • + Highly atmospheric with subtle, spooky thorn and web textures
  • The parchment texture is very dark, making some of the background details hard to see
  • The scroll banner is a bit flat compared to the rest of the image

Wan 2.7 Pro

  • + Beautiful, complex composition with many thematic elements like lanterns and tombstones
  • + Clear, legible gothic typography and a well-rendered scroll banner
  • + Rich detail in the thorns, webs, and twisted trees
  • The lighting is bright and illustrative rather than the requested 'moody' or 'cinematic' style
  • Includes hallucinated extra text like 'Est. 1847' and 'Dress code'
  • Small visual artifacts in the sky and web sections

Verdict: GPT Image 1 Mini captures the 'moody' and 'cinematic' atmosphere much better than the competitor, resulting in an invitation that feels more authentic to the vintage gothic theme. While Wan 2.1 Pro provides a more detailed and vibrant illustration, it fails to follow the lighting constraints and adds several unwanted lines of text.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
GPT Image 1 Mini
Before After
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Successfully added a very thick and voluminous head of hair.
  • + Maintained the subject's glasses and general jacket appearance.
  • Significantly altered the facial structure, making the man look younger and like a different person.
  • The hair looks somewhat illustrative and lacks realistic fine strands at the edges.

Wan 2.7 Pro

  • + Excellent source preservation, keeping the man's face and identity virtually identical to the original.
  • + High realism in hair texture, including natural flyaways and a convincing hairline.
  • + Perfectly matches the lighting and environment of the original photo.
  • The hair could be slightly thicker as per the 'full, thick head' prompt, though it is realistic.

Verdict: Wan 2.7 Pro is the clear winner as it successfully adds a realistic head of hair while perfectly preserving the identity, facial features, and environmental lighting of the source image. In contrast, GPT Image 1 Mini significantly altered the subject's face, resulting in a different-looking person, which fails the core requirement of image editing.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography with clean, bold text for 'JAPAN' and 'SUSHI'.
  • + Accurate 45-degree isometric perspective creates a strong 3D miniature feel.
  • + The lighting is soft and consistent across the diorama base.
  • The sushi toppings have a slightly overly plastic look compared to Model B.
  • The flag icon is integrated into the text line rather than separate, though still present.

Wan 2.7 Pro

  • + High level of surface detail on the sushi rice and toppings.
  • + Intricate modeling of the garnishes and plate with a premium feel.
  • + Text is correctly spelled and includes the requested flag icon.
  • Includes unrequested extra text 'Authentic Japanese Cuisine' and multiple floating artifacts in the corners.
  • The 'JAPAN' text uses a thin font that doesn't match the 'large bold' requirement as well as Model A.
  • Object placement is slightly off-center and the wooden chopsticks are cut off at the bottom.

Verdict: GPT Image 1 Mini adhered much better to the specific layout and composition requirements, providing the clean, bold typography and perfect centering requested. While Wan 2.7 Pro had superior texture detailing on the food, it failed to keep the image 'ultra-clean' by adding extra text and distracting small debris in the corners.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent caricature style with hand-drawn pencil/colored pencil texture
  • + Successfully captures the subject's likeness and clothing from the source image
  • + All requested elements (anchor job, dog, hockey) are integrated naturally into the scene
  • The right hand gripping the dog is small and anatomically awkward
  • The hockey stick is a bit thin and loses its shape at the top

Wan 2.7 Pro

  • + Creative use of a 'hockey stick microphone' and paw-print lapel pin
  • + High-quality vector illustration style with clear, readable text
  • + Multiple dogs included, all wearing hockey jerseys
  • Subject loses the likeness of the source image, looking more generic
  • The speech bubbles feel a bit literal and less humorous than a visual gag
  • Subject is in a suit, ignoring the specific denim shirt character of the source image

Verdict: GPT Image 1 Mini creates a superior caricature because it maintains the subject's likeness while applying the requested 'exaggerated/humorous' style. It feels like a genuine hand-drawn caricature where the hockey and dog elements are part of the composition. Wan 2.7 Pro makes some very clever creative choices (like the microphone stick), but the facial likeness is lost and the execution feels more like a generic clip-art style.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent depiction of motion with all animals leaping and tumbling
  • + Great fur rendering and soft lighting
  • + Captures the 'big expressive eyes' from the prompt very well
  • The fox's front right leg looks a bit anatomically stiff
  • Butterflies are relatively simple in design

Wan 2.7 Pro

  • + Strong 'god ray' effects and lighting atmosphere
  • + Detailed variety of wildflowers with realistic dew drops
  • + Complex interactions between the animals
  • The fox appears to have two left front paws combined in a cluster
  • The kitten's facial expression is slightly distorted
  • The fox and puppy's body proportions are a bit inconsistent where they overlap

Verdict: Both models followed the prompt well, including all requested animals and atmospheric elements. GPT Image 1 Mini captured a more cohesive sense of play and motion with cleaner animal anatomy, while Wan 2.7 Pro excelled at the environment and lighting but suffered from significant anatomical errors in the fox's legs and the kitten's face.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent Ghibli-inspired art style with soft, rounded facial features.
  • + Captures the requested pastel color palette and warm, nostalgic mood effectively.
  • + Great textures that mimic a colored pencil or soft crayon hand-painted look.
  • The characters look slightly younger and less like the original people compared to the source.
  • Slightly less clarity in the fine patterns of the man's shirt.

Wan 2.7 Pro

  • + Highly accurate preservation of the original subjects' faces and expressions.
  • + Clean watercolor aesthetic with high-quality line work.
  • + Matches the original composition and background elements very closely.
  • The style feels more like a standard digital watercolor filter than specific Ghibli-inspired art.
  • The lighting is less 'dreamy' compared to the other model.

Verdict: Both models performed well, but for different reasons. GPT Image 1 Mini leaned more into the artistic instruction, completely transforming the faces into the iconic Studio Ghibli 'look' with soft, warm lighting. Wan 2.1 Pro was much more successful at preserving the actual identity and likeness of the people in the meme, though its 'Ghibli' influence is limited to a watercolor texture. Wan 2.1 Pro is the winner for creating a faithful stylized version of the source image while maintaining recognizability.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
GPT Image 1 Mini
Before After
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent adherence to the 'hair blowing in the wind' instruction.
  • + Incorporates many flying leaves as requested.
  • Failed to preserve the source image's identity, completely changing the woman's facial features.
  • The dog's face and fur texture were modified unnecessarily.

Wan 2.7 Pro

  • + Perfectly preserves the woman's face and the dog's likeness from the source image.
  • + Effectively adds dynamic hair motion and flying leaves without altering the core subjects.
  • The transition of some hair strands on the right side looks slightly unnatural/feathery.

Verdict: GPT Image 1 Mini failed as an image editor by completely changing the woman's face and identity, essentially generating a new image based on the prompt rather than editing the original. Wan 2.7 Pro successfully applied all requested dynamic effects while perfectly preserving the identity of both the person and the dog from the source image.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography with perfect spelling and correct accent marks
  • + High contrast and clean vector style
  • + Accurate inclusion of all requested elements including the cloche and banner
  • Failed the instruction for a light background, using a black background instead
  • The texture is very subtle, almost appearing as low-resolution noise

Wan 2.7 Pro

  • + Beautifully followed the warm brown, cream, and light background color palette
  • + Sophisticated emblem style with nice illustrative details
  • + Excellent subtle texture that matches the vintage aesthetic
  • Misspelled the name as 'Florion' instead of 'Florian'
  • The steam is rendered as abstract swirls rather than natural steam
  • Cloche dome looks more like a cake stand

Verdict: GPT Image 1 Mini followed the text instructions perfectly, including spelling and accents, but completely ignored the light background requirement. Wan 2.7 Pro captured the aesthetic, colors, and 'vibe' of the prompt much better, but failed on the critical task of spelling the brand name correctly. GPT Image 1 Mini is the winner for functional logo use due to accuracy.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1 Mini
Wan 2.7 Pro

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography rendering for major labels
  • + Clean and playful flat-vector style with consistent line weights
  • + Effective use of the requested navy and muted red palette
  • Non-sensical trajectory lines for the Translunar step
  • The composition is cropped at the bottom, cutting off the crew and footer elements
  • The Earth and Moon icons are stylistically inconsistent compared to the technical lunar module icon

Wan 2.7 Pro

  • + Sophisticated and professional infographic layout with a clear hierarchy
  • + Impressive attention to detail with actual mission data and dates included
  • + Excellent adherence to the full request, including a complete poster layout with crew names
  • Typos in text labels, specifically 'DESCRIPT' instead of 'Descent'
  • The iconography for the lunar module is slightly fragile and lacks the 'clean' solid lines requested
  • Small text elements are slightly blurry or contain minor artifacts

Verdict: Wan 2.4 Pro creates a much more complete and professional-looking infographic that follows the narrative structure and data density requested in the prompt. While GPT Image 1 Mini has cleaner individual icons and perfect spelling, it fails significantly on the logic of the trajectory lines and provides a poorly cropped composition.

Next steps

Explore each model