Head to head
Esc

Models · slot A

to navigate to pick

Vidu Q2 ShengShu Technology Wan 2.6 Alibaba

Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.

Vidu Q2

19.8 arena score

#42 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.6

23.3 arena score

#27 of 62 in Text-to-Image

Top 2 in Image-to-Video
Vote tally

Where the votes landed

Vidu Q2

0%

win rate

Ties

0%

Wan 2.6

0%

win rate

Shared challenges 19

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent photographic quality and sharp focus on the glass cube.
  • + Accurate and high-quality rendering of the red book's texture and gold leaf details.
  • + Very realistic lighting and shadow play on the wooden table.
  • The plant is positioned slightly above/behind rather than being clearly visible through the glass.

Wan 2.6

  • + Strong composition with a depth of field that emphasizes the subject.
  • + Faithful adherence to the lighting direction requested.
  • + Distinct glass-on-glass reflections on the blue sphere.
  • The glass cube has distorted internal edges and alignment issues.
  • The book is missing the front corner where it overlaps with the glass edge.
  • Lower overall clarity compared to the other image.

Verdict: Vidu Q2 is the superior model here, providing much higher image resolution, realistic textures, and structural consistency. While Wan 2.6 captures the soft window lighting well, it suffers from significant geometry errors in the glass cube and the book's shape.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent preservation of the specific car model features from the source image.
  • + Accurate facial likeness and hair style of the man from the source image.
  • + Dynamic composition that conveys a sense of speed along a coastline.
  • The steering wheel placement looks slightly unnatural relative to the driver's hands.
  • Small artifacts on the road ahead of the car.

Wan 2.6

  • + Successfully captures the man's distinct plaid coat and scarf from the source image.
  • + High-quality background environment that strongly evokes the California coastline.
  • + Good motion blur on the wheels and road.
  • Alters the front-end design of the car significantly compared to the source image.
  • The man's facial features are less recognizable as the specific individual from the source.

Verdict: Vidu Q2 is the winner because it identifies and preserves the specific characteristics of both subjects better than its counterpart. While Wan 2.6 does a great job with the man's clothing, it loses the specific design of the Rolls-Royce and the man's facial identity; Vidu Q2 maintains the car's integrity and the person's likeness while placing them perfectly in the requested environment.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent detail in the bicycle's mechanical components
  • + Strong pavement reflections and realistic wet surfaces
  • + Effective use of 'imperfect framing' requested in the prompt
  • The anatomy of the hands and how they grip the bike is physically confused
  • The car in the background lacks the requested motion blur
  • The perspective of the bicycle frame is distorted

Wan 2.6

  • + Stronger cinematic composition with a clear subject and background
  • + Excellent portrayal of age and skin texture on the man
  • + Good atmosphere with visible rain droplets and background bokeh
  • The bicycle appears to be two different bikes merged together (front and back halves)
  • The raindrops on the jacket look like frozen beads or glass rather than liquid
  • Minor artifacting around the man's hands where they meet the chain

Verdict: Wan 2.6 captures the 'cinematic but realistic' atmosphere much better than Vidu Q2, providing a more evocative character study and better environment. While Vidu Q2 followed the 'imperfect framing' instruction well, its anatomical errors and lack of motion blur make it less successful overall.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent engraving detail on the plate armor with very clean lines.
  • + Very high-quality texture on the leather straps and underlayer fabric.
  • + Strong, clear lighting that makes the character stand out against the background.
  • The character appears a bit too clean and 'pretty' for the 'battle-worn' description.
  • The hair beads are very subtle and almost unnoticeable.
  • The skin texture is somewhat smooth, bordering on a digital render look.

Wan 2.6

  • + Perfectly captures the 'battle-worn' aesthetic with realistic dirt, grime, and sweat.
  • + Excellent interpretation of the hair beads, making them a distinct feature as requested.
  • + More natural skin texture and lifelike eyes that convey a sense of exhaustion.
  • The engraving on the armor is slightly less defined compared to the other model.
  • Some of the leather strap detailing is a bit muddy or lacks sharp contrast.
  • The lighting is a bit more diffused, which slightly reduces the metallic sheen reflection.

Verdict: Wan 2.6 is the clear winner for its superior storytelling and adherence to the 'battle-worn' and 'braided hair with beads' aspects of the prompt, creating a much more authentic character. While Vidu Q2 produces exceptionally sharp and clean armor engravings, the overall presentation feels too polished and lacks the gritty realism requested in the prompt.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Includes distinct sections for Appetizers, Pizza, and Mains as requested.
  • + Vibrant food photography with professional lighting.
  • + Effective use of white space and colorful accents.
  • Text is highly distorted and contains many nonsensical 'AI-style' characters.
  • Layout feels a bit cluttered and lacks a clear focal point.
  • Image borders and elements feel slightly cut off or unevenly spaced.

Wan 2.6

  • + Superior grid layout for food photos that feels more like a modern design.
  • + Text rendering is significantly cleaner and closer to real English characters.
  • + Stronger adherence to the 'minimalist' aesthetic with a balanced composition.
  • The photos primarily show pizza even in sections labeled 'Appetizers' or 'Mains'.
  • Includes some repetitive layout elements like double headings.

Verdict: Wan 2.6 is the clear winner for design quality, offering a much more professional and realistic grid layout with cleaner, legible typography. While Vidu Q2 followed the content categories better (showing different food types for different sections), its text is extremely messy and the overall composition lacks the 'modern minimalist' feel requested in the prompt.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent typography rendering with almost perfect text accuracy.
  • + Bright and vibrant colors that make the food look appealing.
  • + Strong adherence to the starburst and glowing effect requirements.
  • The currency symbol in the starburst is a gibberish character rather than a Euro sign.
  • The burger layers are slightly less 'exploded' and more just stacked with gaps.

Wan 2.6

  • + Superior sense of motion and 'exploded' dynamics with sauce droplets and flying ingredients.
  • + Correct rendering of the Euro (€) currency symbol.
  • + Highly atmospheric background with realistic smoke and glowing embers.
  • The glowing fire effect on the word 'BURGER' makes the text slightly harder to read than model A.
  • The bottom bun looks a bit flat and less photorealistic compared to the top halves.

Verdict: Both models followed the complex prompt exceptionally well, but Wan 2.6 is the winner due to its superior composition and correct rendering of the currency symbol. While Vidu Q2 had cleaner typography, Wan 2.6 captured the 'dynamic, exploded' request more effectively with better motion blur and ingredient separation.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent authentic chalk texture with realistic smudges and strokes.
  • + Accurately captures the 'handwritten' feel with varying line weights.
  • Poor spelling and character rendering throughout the menu items.
  • Incorrect price for the first item ($34 instead of $24).

Wan 2.6

  • + Perfect text rendering and spelling for the entire requested prompt.
  • + Highly realistic composition with chalk dust on the ledge and environmental lighting.
  • + Followed all specific price and item instructions accurately.
  • The handwriting style is slightly more uniform than the highly varied style of Model A.

Verdict: Wan 2.6 is the clear winner as it successfully rendered all text with perfect spelling and followed every detail of the prompt, including the specific prices. Vidu Q2 captured a more authentic chalk texture, but suffered from significant spelling errors and hallucinatory text at the bottom.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent integration of the horse and galaxy texture for a surreal effect.
  • + High-quality, cinematic color grading and lighting.
  • + Correctly interprets 'astronaut riding horse' despite the confusing negative prompt phrasing.
  • The astronaut's hand/reins connection is a bit messy and floating.
  • The hooves have some anatomical distortions.

Wan 2.6

  • + Very realistic rendering of the spacesuit and horse tack details.
  • + Dynamic composition with effective use of light rays and nebulae.
  • + Strong sense of motion and physical presence.
  • The horse appears to be galloping on a rocky surface, which slightly reduces the 'in space' surrealism compared to the other.
  • Minor artifacts in the reins where they meet the astronaut's hands.

Verdict: Both models reasonably ignored the confusing contra-instruction 'horse on top, not vice versa' to deliver the logical 'astronaut riding horse' scene. Vidu Q2 is preferred for its superior surrealism, as the horse itself is made of starlight/galaxies, whereas Wan 2.6 leans toward a more literal, grounded interpretation that places the horse on a solid planetary surface.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Successfully transferred the pea coat, plaid scarf, sunglasses, watch, and rings.
  • + Included more of the body and background by zooming out the frame.
  • + Maintained the distinct skin patterns and added a creative touch of sand on the clothes.
  • Significantly altered the facial features and added a mustache contrary to the prompt.
  • The 'sand' on the jacket looks more like a digital transparency artifact than real sand.
  • Changed the aspect ratio and expanded the background instead of keeping it completely unchanged.

Wan 2.6

  • + Maintained the original subject's facial structure and features more accurately than the competitor.
  • + Kept the background framing and cropping consistent with Image 1.
  • + Accurately rendered the specific plaid pattern and coat material.
  • Missed many peripheral accessories like the watch, jewelry, and shoes.
  • Cropped off the bottom half of the person's body.
  • Simplified the sunglasses style compared to Image 2.

Verdict: Both models failed the instruction to keep the face 'completely unchanged,' as both added sunglasses and minor adjustments to the skin/facial hair. Vidu Q2 succeeded in transferring more of the outfit's complex layers and accessories but failed significantly on face preservation and source framing. Wan 2.6 is the likely winner because it preserved the subject's identity and the original image's composition much better, even though it omitted the lower-body accessories.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent interior depth and realistic lighting from city signs.
  • + Strong adherence to the 'both paws on steering wheel' and capybara clothing requirements.
  • + Clean, modern photographic look with good bokeh.
  • The passenger's hands and phone interaction are slightly distorted.
  • The perspective from the side window slightly obscures the 'inside at night' mood compared to a frontal shot.

Wan 2.6

  • + Perfectly captures the 'bored' expression of the businesswoman.
  • + Great atmospheric detail with rain on the windshield and classic taxi elements.
  • + Composition effectively shows both characters clearly in one frame.
  • Anatomical issues with the capybara's paws, which appear fused and glove-like.
  • The passenger is sitting in the front seat instead of the back seat as requested.

Verdict: Vidu Q2 followed the spatial instructions more accurately by placing the passenger in the back seat and correctly rendering the capybara's paws on the wheel. Wan 2.6 created a more atmospheric and expressive scene with excellent textures, but it failed the positioning prompt by placing the businesswoman in the passenger seat.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Strong parchment texture aesthetic
  • + Clear and bold gothic font style
  • + Excellent jack-o-lantern rendering
  • Multiple spelling errors in the title and banner
  • Incorrect date and time numbers
  • Less atmospheric composition compared to the other model

Wan 2.6

  • + Perfect text accuracy for all requested fields
  • + Superior atmospheric lighting and depth
  • + Intricate border integration with thorns and webs
  • The 'golden' texture on the font is slightly less classic gothic
  • Slightly busier composition

Verdict: Wan 2.6 is the clear winner as it followed all specific text instructions perfectly, including the exact date and location. While Vidu Q2 has a nice parchment style, it failed significantly on text legibility and accuracy, containing multiple typos and numerical errors.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
Vidu Q2
Before After
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent preservation of the original facial structure and features
  • + Natural integration of the hair with the existing beard and sideburns
  • + Maintains consistent lighting across the new head shape
  • The hairline is a bit high and stiff
  • The hair texture looks slightly more digital than the rest of the image texture

Wan 2.6

  • + Provides a very thick, full volume of hair as requested
  • + Hair texture is highly realistic with varied strands
  • + Good color matching to the subject's existing beard
  • Slightly alters the eye shape and facial expression compared to the original
  • The forehead area looks a bit compressed due to the volume of the hair

Verdict: Both models successfully added a full head of hair while preserving the background and clothing. Vidu Q2 is the better choice because it perfectly preserves the subject's identity and facial features with zero distortion, whereas Wan 2.6 slightly modifies the eye and brow area.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent typography with a 3D effect tailored to the 'cartoon' prompt
  • + Vibrant colors and high-quality rendering of textures
  • + Creative integration of the flag icon
  • The 'diorama base' is a bit small and less distinct compared to the plate
  • Garnish looks slightly cluttered compared to the 'minimal' request

Wan 2.6

  • + Perfect adherence to the isometric perspective and diorama base description
  • + Clean, minimalist composition that looks very professional
  • + Accurate rendering of the requested text and flag icon
  • The texture of the sushi rice is a bit repetitive and less 'soft refined'
  • Overall lighting feels slightly flatter than Model A

Verdict: Both models followed the prompt very well, but Model B achieved a more professional isometric 'miniature' look with its distinct diorama base and clean layout. Model A had superior text rendering and more appealing food textures, but it felt less like an isometric diorama and more like a standard character-style render.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent preservation of the subject's facial features in a caricature style.
  • + Cleverly integrates the hockey theme directly into the studio desk/ice rink hybrid setting.
  • + High-quality illustrative style with clean lines and vibrant colors.
  • The fingers on the left hand are anatomically incorrect and poorly rendered.
  • The microphone 'NEWS' tag is floating awkwardly.

Wan 2.6

  • + Successfully incorporates all elements: tv anchor, dogs, and hockey gear.
  • + Clear and consistent cartoon art style.
  • The facial likeness to the original subject is almost entirely lost.
  • The hockey stick is being held by a dog's paw in a physically impossible way.
  • The overall composition feels more generic and less like a personalized caricature.

Verdict: Vidu Q2 is the clear winner because it maintains a recognizable facial likeness to the original source image while translating it into a caricature style. While Wan 2.6 followed the prompt instructions, it produced a generic cartoon character that no longer resembles the user, whereas Vidu Q2 successfully integrated the hockey rink, news desk, and dogs into a cohesive and personalized scene.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Successfully included all four animal types plus an extra puppy
  • + Bright and clear colors with high energy
  • + Butterflies are well-formed and logically placed throughout the meadow
  • The fox kit has a slightly cartoonish aesthetic compared to the other animals
  • Background lighting is a bit flat despite the prompt for god rays

Wan 2.6

  • + Excellent interpretation of 'god rays' and morning dew sparkles
  • + Highly realistic fur textures and expressive facial features
  • + Superior composition with animals gathered closely to imply 'tumbling together'
  • Floating debris or thistle seeds appearing as artifacts near the butterflies
  • The rabbit's ear anatomy is slightly merged with the puppy's fur

Verdict: Wan 2.6 is the winner as it perfectly captures the requested atmosphere, including atmospheric god rays and sparkling dew that Vidu Q2 missed. While Vidu Q2 followed the count of animals well, Wan 2.6 provided a more emotive and hyper-photorealistic scene that feels more like a coherent, high-quality masterpiece.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent character consistency, maintaining the original poses and expressions perfectly
  • + Stronger adherence to the specific 'Studio Ghibli' line art and character design style
  • + Preserves the precise background architecture and street layout
  • Lighting is a bit flat compared to the requested 'dreamy' atmosphere

Wan 2.6

  • + Beautiful watercolor-like hand-painted textures
  • + Soft, dreamy lighting with nostalgic sparkles/bokeh effects
  • + Effective use of soft pastel colors
  • Character expressions are slightly softened, losing some of the 'distracted boyfriend' meme's intensity
  • Slightly less 'Ghibli' in line work, leaning more toward general watercolor anime art

Verdict: Both models performed exceptionally well at preserving the source image while applying the requested style. Vidu Q2 captured the iconic Ghibli character design and sharp line work more accurately, while Wan 2.6 better captured the 'hand-painted' texture and dreamy mood requested in the prompt. Vidu Q2 is the winner for better preserving the comedic tension of the original expressions within the new art style.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
Vidu Q2
Before After
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Excellent adherence to the 'flying leaves' request with numerous, high-contrast autumn leaves.
  • + Significant hair motion that feels natural to the scene's wind.
  • + Good preservation of the background elements and the dog.
  • The warm orange color of the leaves contrasts slightly with the vibrant green trees in the background, creating a minor seasonal inconsistency.
  • One leaf appears to be growing out of the dog's back.

Wan 2.6

  • + Subtle, realistic hair motion that matches the breeze.
  • + The green leaves more naturally match the existing foliage in the background.
  • + Strong preservation of the woman's face and the dog's features.
  • The amount of flying leaves is sparse compared to Model A.
  • The 'dynamic motion' effect feels a bit timid and less energetic than requested.

Verdict: Both models successfully added the requested elements while perfectly preserving the woman, dog, and background. Vidu Q2 followed the prompt more literally by adding a large volume of dynamic, flying leaves which creates a stronger sense of energy, despite the seasonal color shift. Wan 2.6 is more subtle and realistic with its green leaves, but it captures less of the 'energetic and lively' feel requested in the edit instructions.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Clean vector-style execution of the cloche dome design.
  • + Good adherence to the requested warm brown and cream color palette.
  • Significant text errors including 'Caffe Farmiin', 'Esttt', and 'Caffce Fopli20'.
  • Redundant and cluttered text placement below the banner.
  • The steam effects look like floating shapes rather than integrated design elements.

Wan 2.6

  • + Perfect text rendering of 'Caffè Florian' and 'Est. 1720'.
  • + Stronger 'minimalist' aesthetic more suitable for a modern-vintage logo.
  • + Excellent use of subtle texture on the background edges to create a vintage feel.
  • The 'Est. 1720' banner is slightly detached from the cloche and lacks a bit of visual balance.
  • The steam graphic is a bit simplistic compared to the rest of the emblem.

Verdict: Wan 2.6 is the clear winner because it correctly renders the requested text with zero spelling errors, which is critical for a logo design. While Vidu Q2 has a nice cloche illustration, its text is completely garbled and redundant, failing to follow the core requirements of the prompt.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

Vidu Q2
Wan 2.6

AI Judge Analysis

Vidu Q2

  • + Successfully visualizes the flat-vector infographic style requested.
  • + Follows the multi-step structural requirement with icons for Earth, Moon, and the Lunar Module.
  • + Captures the NASA-inspired aesthetic and iconography accurately.
  • The text is largely gibberish and does not follow the requested step labels.
  • The sequencing is confusingly numbered, jumping between 'Step 1' and 'Step 2' in a non-linear way.

Wan 2.6

  • + Excellent text rendering for the title and the names of the astronauts.
  • + Adheres strictly to the navy and white color palette.
  • Completely fails to include the infographic steps, icons, or specific mission phases requested.
  • The layout resembles a poster or a towel rather than a data-driven infographic.
  • Lacks the Saturn V and Earth/Moon icons specified in the prompt.

Verdict: Vidu Q2 is the superior choice because it actually attempts to create the requested infographic, providing a complex layout with the specific icons (rocket, orbit, lunar module) asked for, even though the text it generated is garbled. Wan 2.6 produced a very clean aesthetic with perfect text, but it completely ignored the core instructional components of the prompt, failing to provide any of the numbered steps or mission icons.

Next steps

Explore each model