OpenAI's previous image generation model that accepts both text and image inputs and produces image outputs
Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.
GPT Image 1
#32 of 62 in Text-to-Image
Wan 2.5 (Preview)
#27 of 62 in Text-to-Image
Where the votes landed
GPT Image 1
0%
win rate
Ties
0%
Wan 2.5 (Preview)
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1
- + Perfect adherence to the object arrangement
- + Clean and realistic glass refractions
- + High resolution with smooth textures
- − Lighting is a bit flat compared to the dramatic side light requested
Wan 2.5 (Preview)
- + Excellent handling of window light and shadows
- + Highly realistic textures on the book cover and wooden table
- + Dynamic atmosphere with visible dust motes
- − The sphere appears to be floating or poorly anchored on the glass base
- − Distracting artifacts/dots in the air that might be interpreted as noise
Verdict: GPT Image 1 followed all spatial instructions perfectly, creating a clean and logically sound composition. While Wan 2.5 achieved a more realistic photographic lighting and texture quality, it struggled with the physics of the sphere's placement and introduced heavy atmospheric noise. GPT Image 1 is the preferred output for its superior structural coherence and clarity.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the man's facial features and distinctive hairstyle
- + Maintains the specific car model and color with high fidelity
- + Realistic motion blur and high cinematic visual quality
- − The man's clothing is changed from the plaid coat to a dark scarf and sweater
Wan 2.5 (Preview)
- + Successfully places the specific car into a California coastline environment
- + Good composition with palm trees and coastal scenery
- − Significant loss of detail and accuracy in the man's face
- − The car is positioned as a right-hand drive vehicle, making the man a passenger in the context of driving in California
- − The man's iconic hairstyle is less preserved compared to Model A
Verdict: GPT Image 1 is the superior edit because it maintains high detail and recognizable likeness of both the man and the car from the source images. It also correctly places the driver on the left side of the vehicle for a California setting, whereas Wan 2.5 (Preview) makes the car right-hand drive and loses facial clarity.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1
- + Excellent shallow depth of field and bokeh realism.
- + Superior skin texture and natural lighting on the subject.
- + Effective cinematic color grading that feels authentic to a 50mm street shot.
- − The 'imperfect framing' request led to a very tight crop that cuts off much of the bicycle.
- − The bicycle's mechanical structure is slightly abstract and nonsensical in the rear wheel area.
Wan 2.5 (Preview)
- + Better inclusion of the full bicycle and tools, showing more of the 'repairing' action.
- + Strong reflections on the wet pavement that enhance the rainy atmosphere.
- + Clearer street context and background depth.
- − The skin texture appears slightly smoothed and plastic compared to the other model.
- − The rain droplets are rendered as very distinct white lines, which looks a bit artificial.
- − Lacks the motion blur from cars requested in the prompt.
Verdict: GPT Image 1 captures the cinematic and photographic qualities much more effectively, with highly realistic skin textures and a convincing 50mm lens feel. While Wan 2.5 (Preview) provides a better full-body composition and more detail of the bike repair, it fails to include the requested motion blur and lacks the authentic photographic 'grit' found in GPT Image 1.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1
- + Exceptional engraved texture on the plate armor
- + Highly realistic facial skin texture and expression
- + Effective warm, moody lighting that integrates well with the character
- − The beads in the hair are very subtle and blend into the color palette
- − Lacks visible leather straps requested in the prompt
Wan 2.5 (Preview)
- + Excellent adherence to specific details like the beads in hair and leather straps
- + Clearer bokeh effect with distinct sparks and a visible torch source
- + Impressive detail on the frayed cloth and chainmail underlayers
- − The texture on the face looks slightly smoother/less aged than Model A
- − The engraving on the armor looks a bit more repetitive/pattern-like
Verdict: Both models followed the prompt well, but they excelled in different areas. GPT Image 1 produced a more atmospheric and gritty portrait with superior skin and metal textures, while Wan 2.5 captured more of the specific prompt elements like the beads, leather straps, and the torch itself. Wan 2.5 is likely the winner for its comprehensive adherence to all requested details and a more dynamic composition.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1
- + Excellent typography rendering with readable fonts.
- + High-quality, appetizing, and realistic food photography.
- + Clean and professional layout that looks like a real-world design.
- − The grid is slightly cut off at the bottom of the frame.
- − The placeholder text 'Apperoiation descrigion' contains spelling errors.
Wan 2.5 (Preview)
- + Successfully includes all three requested sections (Appetizers, Pizza, Mains).
- + Effective use of grid layout and color-coded section dividers.
- + Captures the modern minimalist aesthetic well.
- − Typography is mostly unintelligible gibberish.
- − The food photos are repetitive and look slightly more artificial.
- − The text 'Mains' appears under a slice of pizza and several salads, which is illogical.
Verdict: GPT Image 1 produces a far more professional and usable design with high-quality food photography and readable (though slightly misspelled) text. While Wan 2.5 (Preview) follows the structural grid requirements for all three categories more accurately, its failure to generate legible text and the repetitive nature of its food images make it less effective overall.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1
- + Excellent typography rendering with consistent fiery glow
- + Clean and photorealistic textures on the burger components
- + Perfect centered composition for a traditional advertisement
- − The price text is incorrect showing '.99' instead of '6.99'
- − The 'exploded' effect is a bit stiff and lacks dynamic energy
Wan 2.5 (Preview)
- + Highly dynamic 'exploded' composition with a true sense of motion
- + All text elements including the price are rendered correctly
- + Creative integration of melting cheese effects into the typography
- − The starburst element is slightly busier and less clean than the other text
- − Some minor blurring on the edges of the floating lettuce
Verdict: Wan 2.5 (Preview) is the clear winner as it successfully rendered all requested text, including the specific price, whereas GPT Image 1 failed the price text content. Furthermore, Wan 2.5 (Preview) captured a much more dynamic 'exploded' feel with slanted angles and flying droplets that better align with the motion-oriented prompt.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1
- + Excellent chalk texture with grainy, realistic strokes.
- + Perfect adherence to the requested text and pricing.
- + Consistent letter size and spacing that feels authentic to a real menu board.
- − The handwriting leans more toward neat printing than 'elegant cursive' for the title.
- − Minimal environmental context, focusing almost entirely on the board surface.
Wan 2.5 (Preview)
- + Dynamic composition with a realistic cafe background and depth of field.
- + Strong 'handwritten' feel with varying strokes and chalk smudges on the board.
- + Applies a nice slant to the text as requested in the prompt.
- − Failed to include 'Herbs' in the second menu item title.
- − Repetitive text error on the final item, showing '$9' twice on separate lines.
- − The chalk texture looks a bit too smooth and digital compared to the grain in Model A.
Verdict: GPT Image 1 is the superior choice because it followed the text instructions perfectly, including all specific menu items and prices without errors. While Wan 2.5 (Preview) produced a more visually interesting composition with a better cafe atmosphere, it suffered from significant text hallucinations and omissions in the menu items.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1
- + Excellent cinematic lighting and texture on the horse's coat.
- + Strong anatomical rendering of the horse and suit textures.
- − Failed the core prompt instruction of placing the horse on top of the astronaut.
- − Visual composition is standard and lacks the requested surrealism.
Wan 2.5 (Preview)
- + High vibrant contrast and clear, detailed background elements.
- + Good sense of motion with the flowing mane and dust particles.
- − Completely ignored the counter-intuitive instruction for the horse to be on top.
- − Standard 'astronaut on horse' interpretation which is the opposite of the specific prompt.
Verdict: Both models failed the specific prompt instruction to place the 'horse on top' of the astronaut, instead providing the conventional 'astronaut on horse' image. GPT Image 1 is preferred for its superior cinematic lighting and grounded, realistic textures, whereas Wan 2.5 (Preview) feels more like a generic digital composite.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the subject's face, hair, and unique skin features
- + Accurately recreates the layers, pattern, and style of the clothing from Image 2
- + Natural lighting and seamless integration with the existing background
- − The watch and hands are slightly blurred compared to the rest of the image
Wan 2.5 (Preview)
- + High quality clothing reconstruction including the watch and ring accessories
- − Completely failed to preserve the person from Image 1, replacing him with the man from Image 2
- − The edit ignores the explicit instruction to keep the base person's face and hair unchanged
Verdict: GPT Image 1 followed all instructions perfectly, successfully keeping the unique identity of the person from Image 1 while dressing him in the complex outfit from Image 2. In contrast, Wan 2.5 (Preview) failed the fundamental task of image editing by replacing the target person entirely with the man from the reference image.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1
- + Excellent photorealistic texture on the capybara's fur and the jacket.
- + Cinematic lighting that accurately reflects a nighttime taxi interior.
- + The passenger is correctly positioned in the back seat as requested.
- − The passenger is seated in what looks like the front passenger seat rather than the back seat.
- − The composition is a bit tight, losing some of the 'New York' environmental context.
- − The capybara's paws have a slightly muddy, less defined texture.
Wan 2.5 (Preview)
- + Includes a highly detailed New York backdrop with vibrant city lights and rain effects.
- + Sharp focus on the capybara and the steering wheel details.
- + The passenger is clearly visible and correctly following the prompt instructions.
- − The passenger appears to be floating or sitting in an anatomically impossible position relative to the seat back.
- − The capybara's hands/paws look slightly more like human fingers with long nails than natural paws.
- − The taxi sign on top is strangely placed relative to the interior perspective.
Verdict: Both models followed the complex prompt well, but they struggled with spatial logic. GPT Image 1 feels more grounded and photorealistic in its lighting and texture, though it failed the 'back seat' instruction by placing the woman next to the driver. Wan 2.5 (Preview) captured a much better New York atmosphere and atmosphere, but the passenger's seating position looks physically disconnected from the car seat.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent typography and perfect text rendering for all requested details
- + Cohesive vintage aesthetic with a dark, atmospheric color palette
- + Sophisticated border design that integrates well with the overall composition
- − The jack-o-lantern is a bit subdued compared to the large text
- − Less dynamic lighting than model_b
Wan 2.5 (Preview)
- + Vibrant, cinematic lighting with impressive fire effects inside the pumpkin
- + High level of detail in the twisted trees and thorny border elements
- + Strong 3D pop and depth in the central illustration
- − The text is somewhat generic and lacks the 'vintage gothic' elegance requested
- − A bit cluttered, with different artistic styles clashing between the border, background, and center
- − The scroll banner is small and lacks the same presence as model_a's
Verdict: GPT Image 1 is the superior choice for a functional invitation because it perfectly captures the 'vintage gothic' mood and renders all text with 100% accuracy and professional layout. While Wan 2.5 (Preview) offers more impressive 3D lighting and illustration detail, its typography feels less integrated into the requested theme.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the original face and features
- + Matches the exact lighting of the original scene
- − The hair looks like a poorly blended toupee rather than a natural growth
- − The hairline and temples have visible masking artifacts and blurry edges
Wan 2.5 (Preview)
- + Natural-looking hair texture and professional blending with the forehead
- + Excellent integration of hair behind the ears and along the temples
- + Preserves the original facial identity perfectly
- − Slightly altered the top of the ear to accommodate the sideburns
Verdict: Both models successfully preserved the original face, clothing, and background. However, Wan 2.5 (Preview) produced a far more realistic result with a believable hairline and volume and integrated the hair into the existing sideburns seamlessly. GPT Image 1's attempt resulted in a hairpiece that looks superimposed, with blurry transition lines and an unnatural shape.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to the requested isometric 45-degree perspective.
- + Clean, legible text rendering with a centered flag icon.
- + Higher complexity of sushi types including nigiri and a maki roll.
- − The raised diorama base is a bit large for the plate size.
Wan 2.5 (Preview)
- + Beautiful soft lighting and subsurface scattering on the rice and fish.
- + High-quality text rendering and graphic layout.
- + Clean, professional aesthetic with refined materials.
- − Missed the 45-degree top-down isometric angle, opting for a lower perspective.
- − Failed to center the flag icon as requested in the prompt.
- − The 'diorama base' is less distinct, looking more like a simple tiered plate.
Verdict: GPT Image 1 followed the technical layout requirements much more closely, specifically the isometric 45-degree angle and the positioning of the flag icon. While Wan 2.5 (Preview) produced a more visually pleasing rendering of materials and lighting, it sacrificed the specific perspective and spatial instructions requested in the prompt.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1
- + Captures the likeness and specific denim shirt from the source image accurately.
- + Excellent caricature style with traditional watercolor textures and expressive facial exaggeration.
- + Cleverly integrates all elements into a unified scene (dog on desk, hockey on screen, news desk).
- − The hockey stick on the desk appears slightly disconnected from the perspective of the rest of the image.
Wan 2.5 (Preview)
- + Clean vector-style graphic that is highly readable.
- + Incorporates multiple dogs and a clear hockey theme on the background monitor.
- + Preserves the denim jacket and black shirt combination from the original photo.
- − Facial likeness to the source person is weaker compared to the other model.
- − The composition feels a bit more like a sticker or logo than a cohesive caricature scene.
Verdict: GPT Image 1 produces a superior caricature that maintains a strong facial likeness to the source image while employing a charming watercolor artistic style. While Wan 2.5 (Preview) follows all instructions, its output feels more generic and lacks the expressive personality and cohesive storytelling found in GPT Image 1's composition.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1
- + Excellent anatomical rendering with natural lighting and fur textures.
- + Realistic integration of animals into the lighting environment with believable god rays.
- + Dynamic composition that truly captures the tumbling and chasing action requested.
- − The rabbit is slightly smaller and more obscured compared to other subjects.
Wan 2.5 (Preview)
- + Vibrant color palette with clear dew sparkles and flower details.
- + All four animals are clearly visible and well-spaced in the foreground.
- − The fox kit has unnatural, glowing blue eyes that don't match the prompt's photorealistic requirement.
- − Floating water droplets and seeds feel like static overlays rather than part of a 3D scene.
- − The animal anatomy and fur textures look somewhat cartoonish compared to the competition.
Verdict: GPT Image 1 is the superior choice because it achieves a much higher level of photorealism with sophisticated lighting and natural-looking animal features. Wan 2.5 (Preview) struggles with realism, producing uncanny glowing eyes on the fox and artificial-looking environmental effects that detract from the overall quality.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1
- + Excellent hand-painted texture that mimics a watercolor or colored pencil style.
- + Perfect adherence to the 'soft pastel colors' and 'nostalgic mood' requested.
- + Successfully translates the facial expressions into a stylized Ghibli-esque simplicity.
- − The composition is slightly tighter/cropped compared to the original image.
- − The background architectural details are almost entirely lost to the painterly wash.
Wan 2.5 (Preview)
- + Very high fidelity to the original source image's composition and poses.
- + Captures the modern Ghibli digital animation aesthetic with clean linework.
- + Preserves background details and clothing patterns with high accuracy.
- − The lighting is a bit flat and lacks the 'dreamy' quality requested.
- − The 'hand-painted textures' prompt was largely ignored in favor of a clean digital look.
Verdict: Both models did an excellent job translating the 'Distracted Boyfriend' meme into an anime style. GPT Image 1 leaned heavily into the 'illustration' and 'hand-painted' aspects of the prompt, creating a beautiful watercolor-esque piece that feels like concept art. Wan 2.5 (Preview) maintained much better source preservation, perfectly recreating the scene's layout and clothing patterns in a digital animation style, though it missed the requested painterly texture and dreamy lighting.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to the 'windy' hair request with realistic strands and lift.
- + Added a high density of leaves to create a true sense of motion and energy.
- + Modified the dog's tail and fur to match the wind direction.
- − The leash has become slightly blurred and warped compared to the original.
- − Introduced some minor artifacts around the flower area in the background.
Wan 2.5 (Preview)
- + Successfully preserved the original quality and sharp details of the woman and dog.
- + Applied the wind effect to the hair while keeping the face perfectly intact.
- + Added leaves that contrast well with the background.
- − The added leaves look like flat, green shapes rather than organic, flying foliage.
- − The wind effect feels less 'dynamic' because the dog's fur is completely static.
Verdict: GPT Image 1 followed the instructions more comprehensively by applying the wind effect to both the woman and the dog, creating a more cohesive and energetic scene. While Wan 2.5 (Preview) maintained higher image fidelity and source preservation, its added leaves look artificial and it failed to add motion to the dog's fur.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1
- + Excellent typography with the correct accent on 'Caffè'
- + Precise rendering of 'Est. 1720' text within the banner
- + Clean, minimalist vector style that feels modern yet vintage
- − Failed to provide a 'light background' as requested, using black instead
- − Minimalist steam trail is a bit overly simplistic
Wan 2.5 (Preview)
- + Followed the 'light background' with 'subtle texture' constraint perfectly
- + Superior composition with decorative framing and multi-layered banners
- + Includes aesthetic highlights and shadows on the cloche dome
- − The accent on 'Caffè' is slightly misplaced or stylized as a dot
- − Banner perspective makes 'Florian' look a bit cramped
Verdict: Wan 2.5 (Preview) provided a much better overall package by adhering to the color and background requirements, resulting in a more complete 'vintage' brand feel. While GPT Image 1 had slightly cleaner typography, its failure to use a light background makes it less aligned with the specific prompt instructions than Wan 2.5 (Preview).
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to the requested flat-vector aesthetic with clean, chunky icons.
- + Legible text for most labels and accurate silhouette imagery.
- + Well-balanced composition that uses the full frame effectively.
- − Includes a misspelled word 'EARLLUNAR'.
- − The icons for Earth Orbit and Translunar are a bit abstract, with Translunar being just an arc.
Wan 2.5 (Preview)
- + Features a more complex diagrammatic flow that clearly maps out the trajectory.
- + Superior text rendering with zero spelling errors.
- + Greater detail in the icons, particularly for the Lunar Module and the Moon surface.
- − The 'Descent' and 'Landing' labels are not aligned with specific icons, creating a gap in the infographic flow.
- − The character portraits for the astronauts are slightly distorted and inconsistent in style.
Verdict: Wan 2.5 (Preview) produces a more structured and technically accurate infographic with excellent text legibility, though it struggles slightly with aligning all labels to their icons. GPT Image 1 captures the 'flat-vector' aesthetic much better, feeling like a cohesive piece of graphic design, but is let down by a significant spelling error. Wan 2.5 is preferred for its clarity and accuracy in representing the mission stages.
Explore each model
Alibaba's text-to-image and image-to-image generation model from the Wan AI suite, offering high-quality visual generation capabilities