OpenAI's cost-effective image generation model for when image quality isn't the top priority
Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Wan 2.6
#28 of 62 in Text-to-Image
Where the votes landed
GPT Image 1 Mini
0%
win rate
Ties
0%
Wan 2.6
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfectly matches all spatial requirements including sphere inside and book on top.
- + Clean and minimalist composition with very realistic lighting.
- + Glass edges are sharp and accurately rendered.
- − The sphere is quite small compared to the scale of the cube.
- − The background plant is very blurred, making it harder to see 'through' the glass as requested.
Wan 2.6
- + Excellent texture on the red book and wooden table.
- + Lighting through the window is very dynamic and natural.
- + The plant is clearly positioned behind and visible through the glass.
- − The glass cube has strange internal geometry with a visible divider in the middle.
- − The blue sphere is quite large, stretching the 'small' descriptor in the prompt.
Verdict: GPT Image 1 Mini followed the spatial instructions perfectly, creating a clean and logical arrangement of objects. While Wan 2.6 has superior surface textures and more interesting lighting, the glass cube contains a strange internal pane that breaks the physics of a simple cube, making GPT Image 1 Mini the more accurate interpretation.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent preservation of the man's facial features and specific hairstyle.
- + High fidelity to the original plaid coat and black snood.
- + Natural lighting and motion blur that creates a cohesive scene.
- − The scale of the man inside the car is slightly too large, making the car look small.
Wan 2.6
- + Very accurate representation of the California coastline with palm trees and cliffs.
- + Good composition with the car and environment.
- + Maintains the car's branding and design elements well.
- − The man's facial likeness is significantly altered, losing the specific features of the source image.
- − The man's hair texture and volume are less accurate than the first model.
Verdict: GPT Image 1 Mini is the winner because it successfully preserves the identity and clothing of the subject from the source image while placing him in the requested environment. While Wan 2.6 creates a more iconic California backdrop, it fails to maintain the man's specific facial features and likeness, essentially generating a new person in similar clothing.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent skin texture and hyper-realistic facial details
- + Accurate 50mm shallow depth of field effect
- + Higher overall visual fidelity in the foreground
- − The red bicycle is partially cropped, reducing narrative impact
- − The light rain is barely visible in the atmosphere
Wan 2.6
- + Better 'candid' environmental composition showing the full scene
- + Strong execution of wet pavement reflections and light rain visible on clothing
- + Dynamic background with motion blur from passing cars
- − Noticeable anatomy issues with the hands appearing distorted
- − The raindrops on the jacket appear as static white dots rather than natural streaks
- − Lower level of fine facial detail compared to Image A
Verdict: GPT Image 1 Mini wins on realism and technical photography cues, providing a highly believable portrait with natural skin textures. However, Wan 2.6 does a better job of capturing the specific atmospheric elements requested, such as reflections and rain effects, though it suffers from significant anatomical errors in the hands and a more 'AI-rendered' look.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent depiction of warm torchlight reflecting off the metal surfaces.
- + Highly detailed and realistic skin texture with subtle scarring and dirt.
- + Intricate engraving on the plate armor that looks consistent and weighted.
- − The braids are very thin and subtle, almost blending into the hair.
- − Beads in the hair are missing or not clearly visible as requested.
Wan 2.6
- + Strong adherence to the 'beads' and 'braids' prompt with clearly visible decorations.
- + Excellent texture on the leather straps and frayed cloth underlayer.
- + Very lifelike, expressive eyes that convey the 'battle-worn' theme well.
- − The grime on the face looks somewhat like digital smudges rather than integrated dirt.
- − Some of the floating sparks/bokeh in the foreground look a bit artificial and distracting.
Verdict: Wan 2.6 followed the specific details of the prompt better, particularly regarding the hair beads and the texture of the cloth/leather clothing. While GPT Image 1 Mini produced a more cohesive lighting environment and superior skin realism, Wan 2.6 captured the 'battle-worn' aesthetic more aggressively and included all requested elements.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text legibility and clean sans-serif typography.
- + Perfect alignment between section headings and corresponding food types in the grid.
- + High-quality, isolated food photography that fits the minimalist aesthetic.
- − The menu content is empty, showing only headings without item descriptions or prices.
- − The layout is a bit overly simplistic, bordering on a template rather than a finished design.
Wan 2.6
- + Successfully incorporates vibrant color accents as requested in the corners and lines.
- + Includes realistic menu elements like prices, item names, and descriptions.
- + Features a greater variety of food photography within the grid.
- − Contains significant gibberish text and spelling errors (e.g., 'Resstaurantsr2eher').
- − Poor organization where 'Pizza' appears in the grid but the section header is below the photos, and 'Appetizers' contains photos of pizza.
Verdict: GPT Image 1 Mini produces a much cleaner and more professional-looking design that strictly adheres to the 'minimalist' and 'bold sans-serif' requirements, though it lacks filler text. Wan 2.6 attempts a more complex layout with vibrant accents but suffers from common AI text artifacts and a confusing logical flow between the images and section headers.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent and clean text rendering for all elements
- + Highly professional studio-style lighting on the burger components
- + Perfect adherence to the starburst and fiery glowing text requirements
- − Static composition that lacks the 'exploded' motion requested
- − The background is slightly too plain compared to the fiery request
Wan 2.6
- + Dynamic 'exploded' composition with a true sense of motion and flying ingredients
- + Very detailed fiery background with smoke and embers
- + High-quality textures on the meat and sauce
- − The 'MAGIC BURGER' text has slight irregularities in the 'B' and 'A'
- − The starburst is a generic cartoon graphic rather than the requested fiery glowing effect
Verdict: GPT Image 1 Mini renders cleaner, more professional typography and perfectly follows the graphic design instructions like the starburst, but it feels static. Wan 2.6 captures the 'dynamic' and 'exploded' energy of the prompt much better, creating a more exciting visual despite slightly less polished text. GPT Image 1 Mini is preferred for a finished ad look, while Wan 2.6 is better for its photographic energy.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text legibility and accuracy with zero spelling errors.
- + Stable and clean composition.
- + Consistent chalk texture across all lettering.
- − The 'cursive' request for the title was not followed, as it remains blocky.
- − The text looks slightly too uniform, bordering on a digital font look rather than true organic handwriting.
Wan 2.6
- + Captures the 'elegant cursive' and 'handwritten' request much more authentically.
- + The texture of the board with chalk dust and smudges provides superior realism and atmosphere.
- + Good adherence to the slanted handwriting requested in the prompt.
- − Small spelling error in the date ('APRIL' is misspelled as 'APRIL' but the 'L' is mangled and there is an extra 'I').
- − The composition is slightly tighter at the edges.
Verdict: GPT Image 1 Mini provides a very clean and perfectly legible menu, but fails to deliver the 'elegant cursive' requested for the title. Wan 2.1 captures the artistic soul of the prompt much better, featuring beautiful cursive and a realistic chalkboard texture, despite a minor character artifact in the date. Wan 2.1 is the preferred choice for its superior interpretation of the requested style and atmosphere.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent cinematic lighting and texture on the space suit.
- + Crisp resolution and realistic rendering of the horse's musculature.
- − Failed the negative constraint entirely; the astronaut is riding the horse.
- − The composition is a bit standard and lacks the requested surrealism.
Wan 2.6
- + High level of detail in the astronaut's gear and the cosmic background.
- + Dynamic composition with vibrant colors and lighting.
- − Failed the specific spatial instruction for the horse to be 'on top'.
- − The prompt requested surrealism, but this remains a literal interpretation of an astronaut riding a horse.
Verdict: Both models failed the specific prompt instruction to place the horse 'on top' of the astronaut, instead opting for the cliche image of an astronaut riding a horse. GPT Image 1 Mini has a more grounded, cinematic feel, while Wan 2.6 offers a more vibrant, hyper-detailed cosmic aesthetic, but both are fundamentally flawed regarding prompt adherence.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent adherence to the request for the 'exact elaborate outfit' including the watch and belt.
- + Successfully maintains the vitiligo pattern on the visible hands.
- + Provides a full-body view that demonstrates logical garment fitting.
- − Significantly alters the subject's facial features and hair, failing the 'completely unchanged' instruction.
- − The subject's skin tone appears more muted and less vibrant than the original source.
Wan 2.6
- + Near-perfect preservation of the source person's face, skin texture, hair, and sand on the cheek.
- + Accurately places the scarf and coat while matching the perspective of the original pose.
- + Background remains highly consistent with the original source image.
- − Missed several components of the outfit such as the watch, belt, and shoes by cropping the image.
- − Added sunglasses that were not part of the person's face in Image 1 and look different from the ones in Image 2.
Verdict: Wan 2.6 is the superior model for image editing because it flawlessly preserved the identity of the person from the source image, whereas GPT Image 1 Mini generated an entirely new face that only vaguely resembles the original. Although GPT Image 1 Mini was more thorough in including all requested clothing items (watch, shoes), its failure to maintain the base person makes it a poor choice for a person-centric edit.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent photorealism and cinematic lighting
- + High quality texture detail on the capybara's fur
- + Accurate moody atmosphere for a night taxi ride
- − The passenger is seated correctly in the back but slightly out of focus
- − Only one paw is clearly visible and positioned on the steering wheel
Wan 2.6
- + Follows the instruction for both front paws on the steering wheel better
- + Includes vibrant New York street lights in the background
- + Captures the bored expression of the businesswoman well
- − Layout error where the passenger appears to be in the front seat next to the driver rather than the back seat
- − The capybara's hat looks more like a police officer hat than a taxi driver cap
- − Visual artifacts on the car roof
Verdict: GPT Image 1 Mini creates a much more believable and photorealistic scene with superior lighting and texture, though it ignores the specific 'two paws' detail. Wan 2.6 attempts more of the literal prompt details but fails significantly on composition by placing the passenger in the front seat, which contradicts the prompt and realistic taxi layouts.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text legibility and alignment
- + Strong parchment texture that feels cohesive with the vintage theme
- + Balanced layout with clear hierarchy
- − Very dark color palette limits the visibility of the twisted trees
- − The 'webs and thorns' border is less intricate than requested
Wan 2.6
- + Ornate and detailed border with thorns and cobwebs
- + Highly cinematic lighting with a vibrant blue sky and orange glow contrast
- + Great attention to detail in the twisted tree textures
- − Text rendering is slightly inconsistent, with 'a' looking like 'o' in 'to'
- − The text lacks the perfect alignment of Model A
Verdict: GPT Image 1 Mini produces a more professional and legible invitation with superior layout and text clarity. However, Wan 2.6 offers a much more visually striking and detailed illustration that better captures the 'cinematic lighting' and 'thorns' requested in the prompt, despite being slightly less polished in its typography.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully added a thick head of hair with realistic texture
- + Maintains the lighting and environment of the original photo
- − Significantly altered the facial features, making the man unrecognizable compared to the source image
- − Changed the style of the glasses and smoothed out the skin texture
Wan 2.6
- + Excellent source preservation, keeping the facial features and glasses almost identical to the original
- + Hair density and texture look very natural and well-integrated
- − The hairline on the forehead is slightly too low, creating a slightly compressed facial appearance
Verdict: Wan 2.6 is the clear winner because it successfully performed the edit while preserving the identity of the person in the source image. GPT Image 1 Mini failed at source preservation, generating a completely different face that merely shared the same beard and jacket as the original.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent soft 3D textures and smooth lighting
- + Very clean and bold typography
- + Subtle but effective wood grain texture
- − Perspective is slightly lower than the requested 45° isometric angle
- − Chopsticks were not explicitly requested, though they fit the theme
Wan 2.6
- + Perfect adherence to the 45° isometric diorama request
- + Stronger miniature 'toy' aesthetic
- + Correct placement of 'JAPAN' above 'SUSHI'
- − The flag is placed to the left instead of being small/after text as in most clean graphic designs
- − The gray base feels slightly heavy compared to the soft aesthetic requested
Verdict: GPT Image 1 Mini produces a much more polished and commercially appealing visual with superior textures and lighting. However, Wan 2.6 followed the structural instructions more accurately, capturing the specific isometric 'diorama' look and the text layout. Overall, GPT Image 1 Mini is preferred for its significantly higher visual quality and cleaner graphic design.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent preservation of the subject's facial features and denim shirt within a caricature style.
- + Perfectly follows the caricature art style with hand-drawn textures and exaggerated proportions.
- + Includes all requested elements: anchor profession, dog, and hockey gear.
- − The hand holding the microphone is a bit simplified and claw-like.
Wan 2.6
- + Creative inclusion of multiple dogs and a complete TV studio set.
- + Clean, modern vector illustration style with vibrant colors.
- + Incorporates a hockey jersey on the dog for extra thematic flair.
- − The facial resemblance to the source image is significantly lower than Model A.
- − The caricature is less 'exaggerated' in the traditional sense and more of a generic cartoon character.
Verdict: GPT Image 1 Mini is the clear winner because it successfully converts the specific woman in the source image into a caricature while maintaining her likeness and her actual outfit. While Wan 2.6 provides a more detailed scene, the character looks like a generic cartoon and loses the personal connection to the source photo.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent fur texture and sharpness on all four animals.
- + Better composition with a clear sense of 'tumbling' and movement.
- + Natural lighting that enhances the golden retriever and kitten fur.
- − The fox's anatomy is slightly stiff compared to the others.
- − Butterfly wings have very basic patterns.
Wan 2.6
- + Stronger 'god rays' and atmosphere with beautiful dew sparkles.
- + Highly detailed butterfly wing patterns and more complex foreground floral elements.
- + Correct animal count and species depiction.
- − The golden retriever's face and paws look slightly distorted/melted into the grass.
- − The kitten's pose and anatomy look less natural compared to the others.
- − Significant blurring on the white bunny's face.
Verdict: GPT Image 1 Mini provides a higher quality of animal rendering, with sharp, well-defined fur and more expressive faces across all four requested species. Wan 2.6 succeeds in creating a more magical atmosphere with superior lighting and 'dew sparkles', but falls short on the physical coherence of the animals, particularly the golden retriever and the bunny.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully captures the specific facial expressions from the meme in an illustrative style.
- + Excellent use of textured, colored-pencil aesthetic that feels hand-painted.
- + Preserves the composition and character poses almost perfectly.
- − The color palette is a bit overly yellow/warm, losing some of the 'pastel' variety requested.
- − The background architectural details are very blurred compared to the original.
Wan 2.6
- + Captures the Studio Ghibli watercolor aesthetic more accurately with clean line work and soft washes.
- + Maintains the likeness of the original subjects very well while translating them to anime.
- + Includes 'dreamy' lighting elements like soft bokeh/sparkles that enhance the mood.
- − The man's facial expression is slightly less 'guilty' or exaggerated than the original meme.
- − The background characters are a bit simplified.
Verdict: Both models did an exceptional job at preserving the source image's layout and meaning. Wan 2.6 is the winner as its aesthetic much more closely aligns with the Studio Ghibli 'watercolor and ink' style, whereas GPT Image 1 Mini feels more like a colored pencil sketch. Wan 2.6 also balanced the pastel colors and dreamy lighting more effectively.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully added a large volume of flying leaves
- + Clearly modified the hair to flow outward
- + High level of energy conveyed through the environment
- − Significantly altered the woman's facial features and clothing details
- − Poor source preservation as it essentially redrew the characters
- − Lost the original dog's specific pose and look
Wan 2.6
- + Excellent source preservation of the woman's face and original clothing
- + Accurate hair motion that stays consistent with the original style
- + Correctly added the requested flying leaves while keeping the background intact
- − The green leaves look slightly artificial and flat compared to the scene
- − Motion is a bit more subtle than the 'energetic' prompt might suggest
Verdict: Wan 2.6 is the clear winner as it successfully follows the edit instructions while preserving the identity of the person and dog from the source image. In contrast, GPT Image 1 Mini creates a new image that resembles the source but changes the woman's facial structure and the dog's appearance, failing the primary goal of an image edit.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with perfect spelling and accents.
- + High-contrast, clean vector style that looks professional.
- + Accurate representation of the requested banner and text elements.
- − Ignored the 'light background' instruction, opting for black.
- − The texture is very subtle, almost appearing as noise rather than a vintage paper texture.
Wan 2.6
- + Followed the 'light background' and 'subtle texture' instructions perfectly.
- + Clean, modern interpretation of a minimalist logo.
- + Good spatial balance between the icon and text.
- − The banner for 'Est. 1720' is unusually small and poorly integrated.
- − The 'Est. 1720' text is slightly warped and less legible than Model A.
Verdict: Wan 2.6 followed the background and texture instructions much better, creating an authentic vintage paper look. However, GPT Image 1 Mini produced superior typography and a more cohesive emblem design, even though it failed the background color requirement. GPT Image 1 Mini is the better logo designer, while Wan 2.6 is better at following the overall scene description.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent adherence to the infographic structure, including all 6 requested steps.
- + Clean flat-vector style with a perfect NASA-inspired color palette.
- + Perfect text rendering for all step titles and the crew section.
- − The translunar trajectory line is a bit messy and over-loopy.
- − Minor cropping at the bottom of the image.
Wan 2.6
- + High resolution with a clean aesthetic.
- + Successfully included the names of the three crew members.
- − Completely failed to include the requested 6-step infographic content.
- − The background looks like a fabric texture rather than a vector graphic.
- − Composition is mostly empty space.
Verdict: GPT Image 1 Mini followed the complex prompt instructions perfectly, creating a functional and aesthetically pleasing infographic with all requested stages and iconography. Wan 2.6 failed to generate any of the actual mission steps, providing only a title and crew names on a textured background. GPT Image 1 Mini is the clear winner for its superior prompt adherence and design logic.
Explore each model
Alibaba's multimodal generation model from the Wan AI suite, supporting text-to-video, image-to-video, reference-to-video with audio, and text-to-image, in both Chinese and English