OpenAI's state-of-the-art image generation model with better instruction following and adherence to prompts
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
GPT Image 1.5
#7 of 62 in Text-to-Image
Wan 2.7 Pro
#38 of 62 in Text-to-Image
Where the votes landed
GPT Image 1.5
100.0%
win rate
Ties
0.0%
Wan 2.7 Pro
0.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1.5
- + Excellent photographic quality with realistic glass reflections and refraction.
- + Perfect lighting consistency between the window light and the shadows on the sphere.
- + Strong aesthetic balance and clean composition.
- − The sphere is quite large, deviating slightly from the 'small blue sphere' description.
Wan 2.7 Pro
- + Successfully includes the window in the frame to justify the light source.
- + Good plant variety that feels natural in the setting.
- + Accurate placement of all requested elements.
- − The glass cube has internal vertical lines that look like structural supports or seams rather than a solid glass object.
- − The red book looks slightly distorted where it meets the glass edge.
Verdict: GPT Image 1.5 is the winner because it achieves a much higher level of photorealism and renders the glass cube with superior clarity and natural refractions. While Wan 2.7 Pro successfully captures all prompt elements including the window frame, the technical execution of the glass material is less convincing than GPT Image 1.5.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1.5
- + Excellent preservation of the specific person's likeness and clothing
- + Accurate integration of the man's hand on the steering wheel
- + Very high quality California coastline background that matches the car's perspective
- − The car is slightly cropped compared to the original composition
- − Minor lighting mismatch between the man and the bright outdoor environment
Wan 2.7 Pro
- + Preserved the entire car from the source image accurately
- + Great lighting and shadow integration on the road
- − Complete failure to use the man from the source image
- − The driver is a generic placeholder with sunglasses
Verdict: GPT Image 1.5 successfully merged both source images, maintaining the specific identity of the man and placing him realistically behind the wheel of the correctly identified car. Wan 2.7 Pro followed the location instruction well but failed the core editing task of using the provided person, replacing him with a generic figure.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1.5
- + Excellent adherence to the 'candid' and 'imperfect framing' prompt with a tight, ground-level composition.
- + The rain droplets on the man's jacket and the reflections on the ground are highly realistic.
- + The shallow depth of field is well-executed, creating a high-quality cinematic look.
- − The car in the background lacks the requested 'motion blur' and looks like it is parked.
- − The anatomy of the bicycle frame becomes confusing near the rear wheel.
Wan 2.7 Pro
- + Successfully captured more of the urban Japanese environment in the background.
- + The man's skin texture and clothing are naturally rendered without oily AI sheen.
- − The bicycle rendering is physically impossible, with a frame that doesn't connect logically and spokes that are chaotic.
- − The composition feels like a posed portrait rather than the 'candid street photo' requested.
- − Failed to include any motion blur for passing cars.
Verdict: GPT Image 1.5 followed the stylistic cues and mood of the prompt much more effectively, delivering a gritty, candid street scene with beautiful lighting and textures. Wan 2.7 Pro struggled significantly with the structural integrity of the bicycle and provided a more generic composition that didn't feel like a spontaneous candid shot.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1.5
- + Exceptional rendering of skin texture, dirt, and realistic scar tissue
- + Intricate engraving and metallic luster on the armor with high-fidelity cloth textures
- + Perfectly executed bokeh sparks and warm lighting that integrates with the subject
- − The beads in the hair are somewhat uniform and less varied in material than they could be
Wan 2.7 Pro
- + Strong composition with visible torches providing context for the lighting
- + Creative interpretation of the braids and colorful beads
- + Good adherence to the 'battle-worn' descriptor with visible facial scars
- − The armor engraving looks slightly flatter and less detailed compared to Model A
- − Leather straps lack the high-detail grit and weathered texture requested in the prompt
Verdict: GPT Image 1.5 is the superior image due to its incredible attention to detail in the textures; the skin, metal, and cloth look tactile and lifelike. While Wan 2.7 Pro offers a great composition and interesting hair detail, it lacks the raw photographic realism and intricate surface rendering found in GPT Image 1.5.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1.5
- + Excellent text legibility and correct spelling.
- + Well-organized hierarchy with clear section headers.
- + High-quality, appetizing food photography that looks professional.
- − The layout is a bit generic for a 'modern' design.
- − Missing the 'Mains' category photo grid, focusing only on appetizers and pizza.
Wan 2.7 Pro
- + Beautiful aesthetically pleasing layout with a strong grid system.
- + Sophisticated branding elements including logo, QR code, and social icons.
- + Creative presentation with background props.
- − Text is largely nonsensical and illegible upon close inspection.
- − The food photos contain minor AI artifacts and inconsistencies.
- − Tiny text makes it non-functional as a real menu mockup.
Verdict: GPT Image 1.5 is the clear winner for any practical purpose because its text is perfectly readable, correctly spelled, and the hierarchy is logical for a menu. Wan 2.7 Pro produces a more visually stylish 'Dribbble-style' mockup with a superior grid, but fails completely on text generation and legibility, which is critical for a graphic design prompt.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1.5
- + Excellent typography with a fiery, glowing effect that perfectly matches the prompt.
- + Superior textures on the meat and bun create a highly realistic and appetizing look.
- + Dynamic composition with a strong sense of heat and energy.
- − The 'exploded' effect is slightly more compressed than Model B.
- − The bottom bun looks a bit over-charred.
Wan 2.7 Pro
- + Clearer separation of individual burger components in the 'exploded' view.
- + Clean, minimalist background that makes the ingredients pop.
- − Failed to include 'LIMITED TIME ONLY' and '€6.99' in a starburst as requested.
- − The 'MAGIC BURGER' text is plain orange and lacks the requested fiery, glowing effect.
- − The lighting on the ingredients feels a bit artificial compared to the environment.
Verdict: GPT Image 1.5 is the clear winner as it followed every part of the complex text prompt, including the secondary taglines and the specific starburst price element. While Wan 2.7 Pro provided a nice exploded view of the ingredients, it failed to render most of the required text and lacked the gritty, photorealistic texture of the first image.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1.5
- + Excellent chalk texture and smudging effects for a realistic feel.
- + Perfect spelling and completion of the unfinished prompt text.
- + Consistent and authentic handwriting style that looks truly manual.
- − Limited background context compared to the cozy café requested.
Wan 2.7 Pro
- + Beautiful composition including the café interior and framing.
- + Clean, legible text rendering with high contrast.
- − Text appears more like a digital font than natural chalk handwriting.
- − Severe repetition error with 'Grilled Octopus' line appearing twice.
- − Failed to produce the 'elegant cursive' style for the title.
Verdict: GPT Image 1.5 is the clear winner as it followed the technical requirements of the handwriting prompt perfectly, including the realistic chalk texture and the correct completion of the final menu item. Wan 2.7 Pro produced a more aesthetically pleasing background but failed significantly on the text by repeating a line and using characters that look like a digital font rather than natural handwriting.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
GPT Image 1.5
- + Successfully integrated the man from Image 2 into the scene.
- + Maintained character attributes including sunglasses, scarf, and specific clothing details.
- + Accurately replicated the yellow background and red ottoman from Image 1.
- − Failed to match the extreme 'bent over' torso angle and head tilt of the reference pose.
- − The anatomy of the feet is significantly distorted with too many toes.
- − The right hand contains an extra finger and lacks anatomical consistency.
Wan 2.7 Pro
- + Maintained the exact original image composition and pose.
- − Completely failed the edit instruction by returning a slightly blurred version of Image 1.
- − Failed to include any elements from the character reference (Image 2).
Verdict: GPT Image 1.5 successfully followed the complex instruction of combining a character and a pose, though it struggled with the fine details of hands, feet, and the extreme flexibility of the original pose. Wan 2.7 Pro failed the task entirely, providing an output that was essentially a duplicate of the pose reference image with no character changes.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1.5
- + Excellent cinematic lighting and texture on the horse's coat.
- + Rich background detail including a lunar lander and nebula-like starfields.
- + Dynamic composition with kicked-up lunar dust adding to the realism.
- − Failed the negative prompt instruction 'horse on top, not vice versa' by placing the astronaut on top.
- − Anatomical issues with the horse's front legs and hoof attachments.
Wan 2.7 Pro
- + Successfully interpreted the surreal nature of the prompt by letting the horse float in space.
- + Clean, high-resolution rendering with a polished digital art aesthetic.
- + Accurate astronaut suit details and a sense of scale relative to the Earth below.
- − Failed the specific positional instruction 'horse on top, not vice versa'.
- − The floating planets in the background feel a bit cluttered and copy-pasted.
Verdict: Both models failed the specific logic-trap in the prompt ('horse on top'), resulting in a standard 'astronaut riding horse' image. GPT Image 1.5 is the preferred choice due to its superior cinematic quality, more complex lighting, and the grounded, gritty atmosphere of the lunar surface compared to the flatter, more generic space background of Wan 2.7 Pro.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1.5
- + Excellent transfer of the specific outfit components including the coat, scarf, glasses, and watch.
- + Maintains the subject's unique vitiligo patterns and facial features accurately.
- + Perfectly aligns the pose and lighting of the clothing with the original subject.
- − Crop is tighter than the original source image, losing some background context.
- − The skin under the sunglasses doesn't perfectly match the original eye area's texture.
Wan 2.7 Pro
- + Preserves the full composition and framing of the original scene.
- + Maintains the full body pose including the placement of the feet on the sand.
- − Completely failed to use the correct outfit from Image 2, generating a generic gold-patterned blazer instead.
- − Artifacts on the hands such as merging fingers and inconsistent skin patterns.
- − The face is distorted and loses the specific likeness of the person in the source image.
Verdict: GPT Image 1.5 is the clear winner as it followed the complex instruction to transfer a specific outfit from one image to another with high fidelity. Wan 2.7 Pro completely ignored the visual reference of the outfit, providing a random suit instead, and also suffered from significant anatomical artifacts and loss of subject likeness.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1.5
- + Excellent texture on the capybara's fur and the taxi dashboard.
- + Perfect adherence to the requested composition with the passenger in the background.
- + Highly realistic lighting and cinematic bokeh.
- − The capybara's paws look slightly more like badger or otter paws than true capybara feet.
Wan 2.7 Pro
- + Natural-looking hands on the steering wheel.
- + Clear, high-resolution rendering of the urban background.
- − Failed to place the passenger in the back seat, putting her in the passenger seat instead.
- − The capybara's head is not naturally attached to the body, appearing like a mask or a floating element.
- − The composition feels less like a real interior shot due to the wide side-angle.
Verdict: GPT Image 1.5 is the clear winner as it correctly followed the spatial instructions to place the businesswoman in the backseat, creating a more cohesive and humorous scene. Wan 2.7 Pro failed on the composition by placing the passenger in the front seat and produced a disjointed image where the capybara's head does not convincingly align with the driver's body.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1.5
- + Excellent adherence to the 'vintage dark parchment' aesthetic
- + Text is perfectly rendered and integrated into the design
- + Cinematic lighting and moody atmosphere are very effective
- − The thorns and webs border is a bit messy and cluttered
Wan 2.7 Pro
- + Clean, illustrative style with good composition
- + Accurate text rendering for all requested fields
- + Creative border elements like skulls and roses
- − Lacks the requested 'dark parchment' and vintage gothic feel, appearing more like a modern digital illustration
- − The central pumpkin has a strange candle-nose artifact
Verdict: GPT Image 1.5 perfectly captures the requested vintage gothic mood and cinematic lighting, creating a cohesive and professional-looking invitation. While Wan 2.7 Pro followed all text instructions correctly, its bright, clean illustrative style misses the 'dark parchment' and 'spooky' atmosphere requested in the prompt.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1.5
- + Excellent texture and density integration with a natural-looking curly style.
- + Maintains very high fidelity to the original facial features and expression.
- + Seamless blend between the sideburns and the new hair.
- − Higher hairline than Model B, bordering on a slightly receded look despite being 'full'.
Wan 2.7 Pro
- + Successfully adds a full head of hair that matches the character's aesthetic.
- + Higher degree of source preservation for the background and jacket textures.
- + Good hair texture and realistic volume.
- − Slightly alters the eye/eyebrow area, making the subject look a bit younger or different than the source.
- − Hair texture feels slightly more 'painted' compared to the coarse realism in Model A.
Verdict: Both models successfully added hair while respecting the source image. GPT 1.5 is the preferred winner because it manages to add dense, realistic hair while perfectly preserving the unique character of the man's face and the specific texture of his beard and skin. Wan 2.7 Pro is also very strong but slightly softens the facial features during the generation process.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1.5
- + Excellent PBR materials with realistic textures on the wood and ceramics
- + Follows the diorama base instruction perfectly with a multi-layered square block
- + Clear, bold text rendering and accurate flag placement
- − The scale of the tea set and soy sauce bottle makes the 'miniature' diorama feel a bit crowded
Wan 2.7 Pro
- + Captures the 'soft refined textures' and cartoon aesthetic very effectively
- + Clean, minimalist composition that feels very high-clarity
- + Modern typography and layout
- − Missed the 'diorama base' requirement, placing the plate directly on the floor
- − The flag icon is placed to the right instead of at the top-center as requested
Verdict: GPT Image 1.5 adhered better to the technical requirements of the prompt, specifically the diorama base and the exact text layout. Wani 2.7 Pro produced a very clean and aesthetically pleasing '3D cartoon' style, but missed several spatial instructions. GPT Image 1.5 is the winner for its superior adherence to the scene's structural details.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1.5
- + Excellent character likeness preserved while applying the caricature style
- + Very high visual quality with 3D-painted textures
- + Clever integration of hobbies like the dog wearing a hockey helmet
- − The fingers on the hand holding the microphone are slightly warped
Wan 2.7 Pro
- + Successfully incorporates all prompt elements including speech bubbles
- + The facial expression is appropriately exaggerated for the request
- − The art style is more generic 2D clip-art compared to Model A
- − Likeness to the source image is significantly weaker
- − The 'hockey stick' held in the hand is illogical, merging into a microphone stand
Verdict: GPT Image 1.5 is the clear winner as it maintains a much stronger resemblance to the woman in the source image while delivering a professional, high-quality 3D caricature. Wan 2.7 Pro produces a more generic cartoon style with several anatomical and logical errors, such as the strange hybrid hockey stick/microphone stand.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1.5
- + Perfectly captures the 'tumbling together' aspect of the prompt with a dynamic and cuddly composition.
- + Superior fur texture and lighting, creating a warm, golden glow with atmospheric god rays.
- + Excellent expressions on all animals that convey a truly joyful and wholesome vibe.
- − The kitten has an extra toe/pad visible on its raised paw.
- − The fox's anatomy is slightly merged with the puppy on the right side.
Wan 2.7 Pro
- + Better individual limb definition and separation between the animals.
- + Realistic inclusion of all requested animals in a full-body action pose.
- + Clean butterfly renders with varied colors.
- − The animals appear more like they are standing next to each other rather than 'tumbling together'.
- − The kitten's facial structure is slightly distorted and less 'cute' than the others.
- − The lighting feels a bit more flat and less atmospheric than the requested '8K masterpiece' glow.
Verdict: GPT Image 1.5 wins by capturing the emotional heart of the prompt with its warm lighting and the adorable, huddled composition of the tumbling animals. While Wan 2.7 Pro handles individual animal anatomy with more clarity, it lacks the 'ultra-detailed soft fur' and magical atmosphere provided by GPT Image 1.5.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1.5
- + Captures the characteristic Ghibli 'dreamy' lighting and soft color palette perfectly.
- + Successful re-interpretation of the faces into a hand-drawn anime style while keeping the expressions recognizable.
- + Excellent use of soft textures and a warm, nostalgic mood as requested.
- − The foreground character is quite blurry, making it feel less like a sharp illustration and more like a filter.
Wan 2.7 Pro
- + Excellent preservation of the source image's composition and minor details like the plaid pattern on the shirt.
- + Very high clarity and sharpness for a watercolor-style illustration.
- + Accurately conveys the hand-painted texture mentioned in the prompt.
- − The art style is more generic Western watercolor/sketch rather than specific Studio Ghibli anime style.
- − The man's face retains too many realistic features (like stubble detail) to feel like a Ghibli character.
Verdict: GPT Image 1.5 wins on creative interpretation of the 'Studio Ghibli' prompt, providing the characteristic soft lighting, simplified character features, and nostalgic atmosphere typical of the studio's work. Wan 2.7 Pro preserves the source image details much better and has higher clarity, but its art style leans more toward a generic watercolor illustration than the specific anime aesthetic requested.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1.5
- + Successfully added a large volume of falling leaves
- + Effectively captures the wind blowing the hair outwards
- + Maintains high color consistency with the original photo
- − The leaf placement looks a bit flat and like a multi-layered overlay
- − Slightly alters the facial features of the woman compared to the source
Wan 2.7 Pro
- + Natural integration of leaves with varied motion blur
- + Hair movement looks realistic and follows a consistent wind direction
- + Excellent preservation of the woman's face and the dog's appearance
- − Fewer leaves than Model A might feel less 'energetic' to some users
Verdict: Both models followed the instructions well, but Wan 2.7 Pro is the superior edit due to its superior source preservation and more realistic motion effects. GPT Image 1.5 added more leaves, but they appear as a static overlay, whereas the leaves in Wan 2.7 Pro have depth and speed-blur that better matches the background.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1.5
- + Excellent typography with correct capitalization and accent mark
- + Strong vector-style illustration with high-quality stippled shading
- + Centralized and balanced composition for a logo
- − Failed the instruction for a light background, delivering a black one instead
- − The Cloche has a slightly modern/glossy lighting effect rather than purely minimalist vintage
Wan 2.7 Pro
- + Successfully followed the requirement for a light background with subtle texture
- + Exceptional vintage aesthetic with delicate line-art illustration
- + Comprehensive emblem design with extra details like geographical location
- − Spelling error in the primary brand name (Florion instead of Florian)
- − Clutter at the edges of the frame distracts from the central logo emblem
Verdict: GPT Image 1.5 perfectly rendered the text and created a punchy, high-contrast vector logo, but failed the background color requirement. Wan 2.7 Pro captured the vintage 'light background' aesthetic much better and provided a more sophisticated illustration, but failed on the most critical element: the spelling of the brand name. GPT Image 1.5 is the preferred winner because it follows the primary text prompt accurately, which is essential for branding.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1.5
- + Excellent adherence to the color palette and flat-vector style.
- + Clear, large-scale iconography that is easily readable.
- + Perfectly legible and accurate spelling for all step titles and astronaut names.
- − The layout feels a bit cramped within the individual panels.
- − The Saturn V rocket design is slightly generic and lacks the iconic black/white pattern.
Wan 2.7 Pro
- + Superior infographic layout with a logical vertical flow and mission data.
- + Includes extra details like dates, sites, and mission duration that add to the poster feel.
- + Creative use of minimalist icons that fit the modern vector aesthetic well.
- − Multiple spelling errors in the text (e.g., 'DESCRIPT', 'Tranquilicy', 'minures').
- − The NASA logo in the corner is distorted and contains gibberish.
Verdict: GPT Image 1.5 provides a very clean and visually consistent set of illustrations with perfect text rendering, making it a reliable choice for a simple infographic. However, Wan 2.7 Pro better understands the 'poster' aspect of the prompt, creating a more sophisticated vertical layout with technical data points. GPT Image 1.5 is the winner due to the absence of the significant spelling and logo artifacts present in Wan 2.7 Pro.
Explore each model
Alibaba's Wan 2.7 Pro image generation and editing model with higher-quality outputs and support for 4K image generation