Head to head
Esc

Models · slot A

to navigate to pick

Imagen 4.0 Generate 001 Google Wan 2.7 Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

Imagen 4.0 Generate 001

17.0 arena score

#55 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.7

20.5 arena score

#39 of 62 in Text-to-Image

Vote tally

Where the votes landed

Imagen 4.0 Generate 001

0%

win rate

Ties

0%

Wan 2.7

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent photographic quality with a clean, modern aesthetic.
  • + Highly accurate rendering of caustics and reflections on the sphere and cube.
  • + The soft window light is realistically depicted across all surfaces.
  • The plant is barely visible and does not appear 'behind' the cube through the glass as clearly as Model B.
  • The cube reflects its sides internally in a way that looks slightly more like a mirror than transparent glass.

Wan 2.7

  • + Perfect adherence to the plant placement, clearly visible through the glass as requested.
  • + Realistic material textures on the wooden table and the book spine.
  • + Includes the reflection of the sphere on the bottom of the glass accurately.
  • The glass cube has strange internal framing/edges that don't match a solid glass object.
  • The sphere appears to be floating slightly above the bottom surface without a clear shadow contact point.

Verdict: Both models followed the complex spatial instructions well, but Wan 2.7 is the winner for better literal adherence to the plant and glass transparency requirements. While Imagen 4.0 Generate 001 produced a more aesthetically pleasing and high-end photograph, it failed to clearly show the plant through the glass, whereas Wan 2.7 successfully integrated all elements within a realistic scene.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent skin texture and realistic wet clothing details.
  • + Effective use of shallow depth of field and bokeh.
  • + Strong color contrast and cinematic lighting.
  • The bicycle frame geometry is slightly distorted near the hands.
  • The framing feels a bit too balanced for the 'imperfect framing' prompt.

Wan 2.7

  • + Successfully captures a more 'candid' and imperfect composition.
  • + Good reflection work on the wet asphalt pavement.
  • + Atmospheric rainy day color palette.
  • Missing the requested motion blur from passing cars.
  • The hands and bicycle handlebars are significantly mangled and anatomically incorrect.
  • The depth of field is deeper than requested, losing the 50mm lens look.

Verdict: Imagen 4.0 significantly outperforms Wan 2.7 in terms of anatomical accuracy and texture. While Wan 2.7 captures a more natural candid composition, its failure to render the hands and bicycle parts correctly makes it a lower-quality image compared to the sharp and professional-looking output from Imagen 4.0.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent intricate engraving on the plate armor
  • + Strong warm lighting with clear reflections
  • + High contrast and sharp details on the facial texture
  • The 'beads' in the hair look more like metal tech components than traditional beads
  • Skin scars look somewhat artificial like drawn lines
  • Composition feels slightly more 'digital art' and less cinematic

Wan 2.7

  • + Very realistic skin texture with lifelike scars and dirt
  • + Authentic braided hair with small beads as requested
  • + Superb 'battle-worn' aesthetic with realistic scuffs on the metal
  • The engraving on the armor is less intricate than in the other model
  • The torch in the background is slightly distracting and less refined

Verdict: Wan 2.7 delivers a more grounded and realistic interpretation of a battle-worn paladin, with superior skin textures and authentic hair braiding. While Imagen 4.0 has more impressive technical detail in the armor engravings, it feels more like a stylized digital render, whereas Wan 2.7 feels like a lifelike cinematic photograph.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent adherence to the 'grid' prompt requirement
  • + Highly legible title fonts
  • + True minimalist aesthetic with balanced negative space
  • Internal category text is nonsensical gibberish
  • The colorful accents feel a bit random and messy compared to a professional layout

Wan 2.7

  • + Professional and realistic commercial layout
  • + Superior text rendering with mostly legible item names and prices
  • + Includes logical branding elements like a logo, social handles, and QR code
  • Failed the white background requirement by including a textured table and props
  • Layout is more of a standard vertical list than the requested grid of sections

Verdict: Imagen 4.0 followed the specific structural instructions much better, delivering a clean grid and a white background, though the text is nonsensical. Wan 2.7 produced a much more realistic and usable menu design with high-quality food photography and branding, but it ignored the white background constraint and the specific geometric grid layout requested.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent photorealistic texture on the meat and bun
  • + Clean and readable text layout
  • + High fidelity and sharpness throughout the image
  • The 'MAGIC BURGER' text lacks the 'fiery' effect requested in the prompt
  • Composition is a bit static compared to the dynamic intent

Wan 2.7

  • + Strong creative interpretation of the 'fiery' text effect and background
  • + Dynamic sense of motion with splashes and flying ingredients
  • + Very coherent stylized aesthetic
  • The burger ingredients look slightly more illustrative and less photorealistic than Image A
  • Random nut-like shapes floating in the air are out of place

Verdict: Both models followed the prompt well, but Imagen 4.0 provides a much more photorealistic and professional-looking food advertisement. While Wan 2.7 excelled at the creative 'fiery' typography, the realism and clarity of the burger in Imagen 4.0 make it a better overall marketing image.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Successfully rendered the requested menu items with correct pricing.
  • + The text style has a believable chalk texture and slight slant as requested.
  • Includes significant hallucinated text and 'meta' instructions like 'Tittle', 'Menu', and 'Footer' written on the board.
  • Contains several spelling errors like 'Berbs' instead of 'Herbs' and nonsensical filler text at the bottom.

Wan 2.7

  • + Excellent text legibility and accuracy, following the prompt's menu items exactly.
  • + Superior background composition and lighting, creating a realistic 'cozy café' atmosphere.
  • + Perfectly captures the chalk texture and smudge marks on the blackboard.
  • The 'natural variations' in handwriting are subtle; the text looks slightly more like a digital font than Model A's version.

Verdict: Wan 2.7 is the clear winner as it produces a professional, aesthetically pleasing image that adheres perfectly to the requested menu list without hallucinating meta-tags or gibberish. While Imagen 3.0 captures a more authentic 'hand-drawn' feel, it fails significantly on prompt instruction by writing the word 'Footer' and other instructional phrases directly onto the chalkboard.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent cinematic lighting with creative rainbow and nebula effects
  • + High level of detail on the space suit and horse's flowing mane
  • + Dynamic and dramatic composition that feels truly surreal
  • Failed the negative constraint; the astronaut is riding the horse

Wan 2.7

  • + Clean, high-resolution rendering of the astronaut and planets
  • + Realistic textures on the horse's coat and leather saddle
  • + Good cosmic background with clear galaxies and planets
  • Failed the negative constraint; the astronaut is riding the horse
  • Composition feels a bit static and less surreal than the rival model

Verdict: Both Imagen 4.0 and Wan 2.7 failed the specific prompt instruction to place the horse on top of the astronaut, instead defaulting to the common trope of an astronaut riding a horse. Imagen 4.0 is the preferred choice as its artistic interpretation is more cinematic and detailed, whereas Wan 2.7 feels like a more standard stock-style image.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent adherence to the perspective prompt, placing the woman in the back seat and the capybara in the driver's seat.
  • + Clean and legible text on the taxi sign and driver's cap.
  • + Very clear, photorealistic rendering with a strong cinematic feel.
  • The capybara's claws are slightly stylized and somewhat bird-like.

Wan 2.7

  • + Natural textures on the capybara's fur and the jacket.
  • + Effective use of street lighting and reflections on the car window.
  • Failed the spatial instructions by placing the passenger in the front seat next to the driver.
  • The hands/claws of the capybara are messy with anatomical artifacts.
  • Illegible symbols on the taxi roof sign.

Verdict: Imagen 4.0 followed the prompt instructions much more accurately, correctly placing the businesswoman in the back seat to create the requested 'taxi ride' dynamic, whereas Wan 2.7 placed her in the front passenger seat. Imagen 4.0 also produced sharper details and cleaner text, making it the clear winner for both composition and technical quality.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent typography and layout design that feels like a professional digital poster
  • + Strong cinematic lighting with high-contrast color palette
  • + Accurate rendering of the requested text elements with clear legibility
  • The parchment/scroll element on the left side is abstract and poorly integrated
  • The composition feels a bit more modern/digital than 'vintage gothic'

Wan 2.7

  • + Perfectly captures the 'vintage gothic' aesthetic with an aged parchment look
  • + Highly detailed illustrative style with additional atmospheric elements like ravens and skulls
  • + Intricate border design that follows the prompt's request for webs and thorns
  • The text is a bit small and less dominant compared to the illustration
  • The font choice for the event details is more standard and less 'gothic' than Model A

Verdict: Both models followed the prompt exceptionally well, but Wan 2.7 better captures the 'vintage' and 'parchment' aesthetic requested, creating a cohesive piece of art. While Imagen 4.0 has superior typography and a more polished cinematic look, Wan 2.7's intricate illustrative border and classic gothic vibe feel more authentic to the theme.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent photorealistic textures and PBR materials.
  • + Soft, realistic lighting that enhances depth and form.
  • + Clean, high-quality rendering of individual food components.
  • Completely failed to include the requested text ('JAPAN', 'SUSHI') and flag icon.
  • Background is a dull grey/white instead of the requested solid light blue.
  • Lacks the 3D cartoon miniature style, leaning too heavily into realism.

Wan 2.7

  • + Perfect adherence to text instructions with clean, bold typography and flag.
  • + Successfully captured the 3D cartoon miniature diorama aesthetic on a light blue background.
  • + Accurate isometric perspective as requested in the prompt.
  • Textures are slightly more 'plastic' than 'refined' or 'realistic'.
  • Chopsticks are disproportionately small compared to the sushi pieces.

Verdict: While Imagen 4.0 produces a much more realistic and appetizing food render, it fails almost all the specific layout and text instructions. Wan 2.7 perfectly follows the prompt's structural requirements, including the diorama base, background color, and specific text elements, making it the superior choice for the creative brief.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent character interaction with the puppy holding the kitten
  • + Extremely vibrant colors and detailed dew drops on grass
  • + Full compliance with all requested animals and lighting effects like god rays
  • The style leans into an illustrative or 'digital art' look rather than hyper-photorealistic
  • Anatomical oddity with the kitten's leg count and paws

Wan 2.7

  • + Achieves a much more photorealistic texture and depth of field
  • + Natural lighting and realistic fur rendering
  • + Dynamic poses that truly suggest a chase in a meadow
  • The fox's facial structure is slightly distorted and looks less like a 'kit'
  • Minimal interaction between the animals compared to model a

Verdict: While Imagen 4.0 captures a charming, whimsical scene with great character interaction, it feels more like a 3D digital illustration. Wan 2.7 adheres much closer to the 'hyper-photorealistic' part of the prompt, providing lifelike textures and a more convincing outdoor environment. Wan 2.7 is the preferred choice for its realistic interpretation of the animals and lighting.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Perfectly adheres to the minimalist aesthetic with a clean vector style.
  • + Accurate text rendering for both the name and the date.
  • + Composition is balanced and professionally centered.
  • The 'subtle texture' requested for the background is extremely faint.
  • The steam lines are very basic compared to the rest of the illustration.

Wan 2.7

  • + Excellent 'vintage' feel with an emblem-style circular layout.
  • + Great texture on the background giving it a paper-like feel.
  • + Sophisticated use of tones and decorative elements like stars and leaves.
  • Spelling error in the main text, changing 'Florian' to 'Florion'.
  • Less 'minimalist' than the prompt requested, leaning more towards ornate.

Verdict: Imagen 4.0 delivers a much more accurate result by following the minimalist constraint and correctly spelling the brand name. While Wan 2.7 has a more convincing vintage texture and complex composition, the misspelling of 'Florian' as 'Florion' makes it less usable for a real logo design.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

Imagen 4.0 Generate 001
Wan 2.7

AI Judge Analysis

Imagen 4.0 Generate 001

  • + Excellent vector illustration of the Saturn V rocket.
  • + Clean, modern aesthetic with high visual clarity.
  • + Perfectly follows the specified NASA-inspired color palette.
  • Missed the final two steps of the requested infographic (Descent and Landing).
  • The logical flow of the diagram is confusing, with icons placed out of chronological order.

Wan 2.7

  • + Successfully included all six requested steps in the correct chronological order.
  • + Incredible attention to detail with relevant factual data like dates and altitudes.
  • + Creative inclusion of the crew members at the bottom.
  • Includes several spelling errors such as 'Tranquiliry' and 'Descript'.
  • Iconography is slightly less refined and consistent compared to the other model.

Verdict: Wan 2.7 is the clear winner for its superior prompt adherence, successfully depicting all six mission stages and adding impressive supporting details. While Imagen 4.0 Generate 001 produced a more polished Saturn V illustration, it failed to follow the instructional sequence of the prompt and omitted the final two steps.

Next steps

Explore each model