Head to head
Esc

Models · slot A

to navigate to pick

Imagen 4.0 Fast Generate 001 Google Wan 2.5 (Preview) Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

Imagen 4.0 Fast Generate 001

17.7 arena score

#52 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.5 (Preview)

23.4 arena score

#27 of 62 in Text-to-Image

Vote tally

Where the votes landed

Imagen 4.0 Fast Generate 001

0%

win rate

Ties

0%

Wan 2.5 (Preview)

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent reflection and refraction through the glass cube.
  • + High photorealism with professional-looking book textures.
  • + Accurate interpretation of 'small' sphere.
  • The glass cube has an illogical mirrored base plate that wasn't requested.
  • The plant is placed behind and to the side, rather than directly behind the cube.

Wan 2.5 (Preview)

  • + Perfect adherence to the spatial prompt with the plant directly behind the cube.
  • + Strong natural lighting effects with visible dust motes.
  • + Clear and consistent geometry for the glass cube.
  • The blue sphere appears quite large, bordering on filling the cube rather than being 'small'.
  • The book's edges are slightly blurred and lack the crisp detail found in the other image.

Verdict: Both models followed the complex spatial instructions well, but Imagen 4.0 Fast Generate 001 produced a more aesthetically pleasing and photorealistic result with superior glass physics. Wan 2.5 (Preview) followed the placement of the plant more accurately, but the blue sphere's scale was less 'small' than requested.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent depiction of surface reflections and wet pavement textures.
  • + Successfully captures motion blur from passing cars as requested.
  • + Effective use of 'imperfect framing' with the foreground occlusion.
  • The subject's face is partially cut off at the top of the frame.
  • The bicycle's front wheel is floating slightly off the ground.

Wan 2.5 (Preview)

  • + Clearer focus on the repair activity with tools visible on the ground.
  • + Strong adherence to 'natural skin texture' on the hands and face.
  • + Better lighting coherence on the subject.
  • Lacks the requested motion blur from passing cars.
  • Visible artifacts where the bicycle's kickstand and tools meet the pavement.

Verdict: Imagen 4.0 Fast Generate 001 followed the technical aspects of the prompt better, specifically capturing the motion blur and the 'imperfect framing' of a candid street photo. Wan 2.5 (Preview) produced a more descriptive scene with better facial detail, but it failed to include the requested motion blur and had more structural errors in the shadows and contact points on the ground.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + High resolution and realistic skin textures
  • + Lush garden environment and natural lighting
  • Completely failed to follow the prompt's subject matter
  • Shows a man in a leather jacket instead of a paladin in engraved armor
  • Lacks the torchlight, braids, beads, and battle-worn details requested

Wan 2.5 (Preview)

  • + Excellent adherence to every specific element of the text prompt
  • + Impressive detail on the ornate engraved plate armor and tattered cloth layers
  • + Strong atmosphere with warm torchlight reflections and bokeh sparks
  • The braids appear slightly stiff where they meet the hairline
  • The bokeh sparks are a bit large and potentially distracting

Verdict: Imagen 4.0 Fast Generate 001 suffered a total failure in prompt adherence, producing a modern man in a garden instead of the requested fantasy paladin. Wan 2.5 (Preview) followed the prompt perfectly, capturing the engraved armor, braided hair with beads, and battle-worn aesthetic with high cinematic quality.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Features a professional grid layout with structured price points and clear typographic hierarchy.
  • + Effectively uses vibrant color block accents to separate sections.
  • + Accurately represents the specific section headers requested: Appetizers, Pizza, and Main Courses.
  • Repeats the 'PIZZA' heading multiple times unnecessarily.
  • Visual variety in the food photos is low, as all images appear to be very similar pizzas.

Wan 2.5 (Preview)

  • + Excellent food photography variety including salads, mains, and pizzas with high visual appeal.
  • + Strong prompt adherence regarding the colorful grid and modern minimalist aesthetic.
  • + Typography is more legible and cleaner at a glance compared to Model A.
  • Included the prompt text 'Modern minimalist' and 'Menue' directly into the design.
  • Layout is slightly less functional for a real menu, lacking visible pricing for most items.

Verdict: Both models successfully captured the modern aesthetic, but they differed in functional design. Wan 2.5 (Preview) produced much better food photography and variety, though it mistakes the prompt for literal text to include in the design. Imagen 4.0 Fast Generate 001 provides a more realistic menu structure with prices and clear sections, but fails on photo variety by using pizza for every section.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent typography with a glowing neon texture
  • + Clean and professional composition suitable for a high-end menu
  • + Highly legible price starburst that perfectly follows the prompt
  • The burger feels less 'exploded' and more like it is simply stacked with gaps
  • The lettuce is only at the bottom rather than being part of the mid-air suspension

Wan 2.5 (Preview)

  • + Superb sense of motion with ingredients scattered dynamically in 3D space
  • + Creative melting effect on the 'Magic Burger' text that fits the fiery theme
  • + More realistic patty texture with grill marks and better ingredient variety
  • The price text is slightly less cohesive with the rest of the fiery effect
  • The bottom bun includes extra tomato/sauce that wasn't explicitly requested

Verdict: Both models followed the complex prompt very well, but Wan 2.5 captured the 'exploded' and 'dynamic' motion aspects much more effectively than Imagen 4.0. While Imagen 4.0 produced a cleaner typographic design, Wan 2.5's composition feels like a genuine high-action food advertisement with superior lighting and atmospheric effects.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent text legibility and accuracy
  • + Followed the date prompt exactly
  • + Clean and centered composition
  • Text looks like a digital font rather than authentic hand-drawn chalk
  • Spelling error on 'Octuphus' and 'Cookes'
  • The background looks like a generic flat image rather than a 3D cafe environment

Wan 2.5 (Preview)

  • + Text has a very realistic chalk texture with natural smears and dust
  • + Captures the 'cozy cafe' atmosphere with depth of field and warm lighting
  • + Very natural handwriting style that feels authentic to a person writing
  • Includes some redundant text for the cookie item
  • Slightly less legible than the clean lines of Model A
  • The 'Herbs' portion of the second menu item is missing

Verdict: Imagen 4.0 Fast Generate 001 provides much clearer text and adheres better to the exact spelling of the date, although the lettering feels more like a digital font than real chalk. Wan 2.5 (Preview) produces a significantly more artistic and realistic image that captures the requested 'chalk texture' and 'cozy café' atmosphere perfectly, despite some minor text omissions. Wan 2.5 is preferred for its superior visual quality and authentic texture.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + High resolution and clear facial details.
  • + Artistic cinematic lighting and nebula effects.
  • Failed the spatial reasoning constraints by putting the astronaut on top.
  • The horse's anatomy becomes distorted at the bottom with too many legs/joints.

Wan 2.5 (Preview)

  • + Excellent surface detail on the spacesuit and horse fur.
  • + Good use of planetary background elements.
  • Failed the spatial reasoning constraints by putting the astronaut on top.
  • Visible artifacting where the horse's front hoof meets the dust.

Verdict: Both Imagen 4.0 Fast Generate 001 and Wan 2.5 (Preview) failed the core challenge of the prompt, which was to invert the typical relationship and put the horse on top of the astronaut. Because both models defaulted to the standard 'astronaut on horse' trope, the evaluation rests on visual quality, where Wan 2.5 slightly edges out with better suit textures and background composition despite minor artifacts.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + High-quality texturing on the capybara's fur and the leather jacket.
  • + Dynamic composition that cleanly displays the interaction between the passenger and driver space.
  • + Excellent rendition of the bored businesswoman's expression.
  • The capybara's paws lack clear definition on the steering wheel.
  • The background lacks the characteristic bright neon of Manhattan at night.

Wan 2.5 (Preview)

  • + Captures the iconic New York night atmosphere with vibrant lights and rain effects.
  • + Strong adherence to the requirement for both front paws to be positioned on the steering wheel.
  • + Good inclusion of external taxi details like the roof sign.
  • The businesswoman in the background is slightly out of focus and less detailed.
  • The steering wheel geometry appears slightly distorted near the paws.

Verdict: Imagen 4.0 Fast Generate 001 excels in interior clarity and character expressions, providing a more intimate and photorealistic feel to the subjects. Wan 2.5 (Preview) offers a more atmospheric and cinematic rendition of Manhattan, better capturing the specific lighting and external taxi branding requested. While both followed the prompt well, Imagen 4.0 is preferred for its superior texture work and character detail.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Includes the deckled edge parchment paper effect which adds to the vintage feel.
  • + Good layout balance between the illustration and the text sections.
  • + Clean rendering of the event details at the bottom.
  • Significant spelling error in the main title ('IINVIITATION').
  • Characters on the scroll banner are distorted and difficult to read.
  • Lacks the 'thorns' requested in the border description.

Wan 2.5 (Preview)

  • + Perfect text rendering for all requested strings including the banner and title.
  • + Stronger adherence to the prompt with the inclusion of thorns and a highly detailed pumpkin.
  • + Excellent cinematic lighting and a more intricate gothic aesthetic.
  • The transition between the central blue circle and the background parchment is a bit abrupt.
  • The border composition is slightly cluttered compared to the more open layout of Model A.

Verdict: Wan 2.5 (Preview) is the clear winner because it followed all text instructions perfectly, whereas Imagen 4.0 Fast Generate 001 had a major spelling error in the primary title and illegible text on the scroll. Additionally, Wan 2.5 captured the specific 'thorns' detail from the prompt which Model A omitted.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent adherence to the 45° top-down isometric perspective
  • + Perfect rendering of text and the flag icon
  • + High clarity on the raised diorama base and miniature elements
  • The salmon texture looks slightly more plastic than realistic PBR
  • A thin white border is visible around the edge of the square frame

Wan 2.5 (Preview)

  • + Beautiful, high-quality PBR lighting and textures for the salmon and rice
  • + Vibrant colors and professional 3D cartoon finish
  • + Accurate text and icon rendering
  • Failed the isometric perspective, utilizing a lower eye-level angle instead
  • The sushi is slightly too large for the 'miniature' diorama aesthetic requested

Verdict: Imagen 4.0 Fast Generate followed every part of the prompt, including the specific isometric viewing angle and diorama layout, though its textures feel slightly more synthetic. Wan 2.5 (Preview) produced a more visually striking and polished 3D render with superior materials, but it failed to capture the requested isometric perspective. Imagen 4.0 is the winner for its superior prompt adherence regarding composition.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent fur texture rendering and realistic lighting
  • + High physical coherence in how the animals are posed together
  • + Subtle and natural-looking golden hour atmosphere
  • Missed the butterfly prompt element entirely
  • Animals are sitting still instead of 'playfully chasing and tumbling'
  • The cat is solid black rather than the requested tabby

Wan 2.5 (Preview)

  • + Successfully included all prompt elements including butterflies and movement
  • + Accurately rendered specific breeds like the golden retriever and tabby kitten
  • + Captures the 'joyful' and 'tumbling' energy requested in the prompt
  • The fox's eyes appear unnaturally blue and slightly distorted
  • Floating water droplets look like digital artifacts rather than natural dew
  • Composition is a bit cluttered with many competing effects like god rays and pollen

Verdict: Wan 2.5 (Preview) is the winner because it successfully captured the active, joyful energy of the prompt, including the chasing behavior and specific animal markings. While Imagen 4.0 Fast Generate 001 produced a higher-quality, more realistic static portrait, it failed to include the butterflies or the movement requested.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent typography rendering with the correct 'È' accent.
  • + Very clean, professional minimalist layout.
  • + Subtle, high-quality background texture that feels like premium paper.
  • Includes small, nonsensical filler text above the main title.
  • The 'Est. 1720' text is slightly cramped within the banner.

Wan 2.5 (Preview)

  • + Beautiful vector illustration of the cloche and steam.
  • + Stronger 'vintage' aesthetic with the ornate border and crumpled paper texture.
  • + Followed the banner layout more dynamically with arched text.
  • Missed the grave accent on the 'e' in 'Caffè'.
  • The steam feels a bit thick compared to the minimalist request.

Verdict: Both models followed the prompt well, but Imagen 4.0 Fast Generate 001 produced a more authentic logo design with accurate spelling, including the required accent mark. While Wan 2.5 (Preview) offered a more evocative vintage paper background and illustration style, its failure to include the accent in 'Caffè' makes it less precise for the specific brand request.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

Imagen 4.0 Fast Generate 001
Wan 2.5 (Preview)

AI Judge Analysis

Imagen 4.0 Fast Generate 001

  • + Excellent adherence to the requested color palette
  • + Clean vector aesthetic and consistent iconography
  • + Accurate spelling of 'Apollo' in most instances and clear section titles
  • Confused terminology like 'Saturn Vicon' and 'Moor'
  • The logical flow of the infographic is difficult to follow and nonsensical
  • Literal interpretation of prompt phrasing resulted in text artifacts like '+ ON +'

Wan 2.5 (Preview)

  • + Strong composition that flows logically from launch to landing
  • + High-quality vector illustrations of the Saturn V and Lunar Module
  • + Perfect spelling of mission names, crew members, and locations
  • Contains a secondary space shuttle-like craft that is historically inaccurate
  • Faces of the astronauts have slightly distorted or uncanny features
  • Missing specific icons for the 'Descent' and 'Landing' text blocks

Verdict: Wan 2.5 (Preview) produced a far more cohesive and readable infographic that follows the mission timeline logically from top to bottom. While Imagen 4.0 Fast Generate 001 captured the requested color palette well, its layout is chaotic and its text rendering includes many errors based on a literal misinterpretation of the prompt's instructions.

Next steps

Explore each model