Head to head
Esc

Models · slot A

to navigate to pick

Imagen 3.0 Generate 002 Google Wan 2.7 Alibaba

Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.

Imagen 3.0 Generate 002

21.0 arena score

#36 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.7

20.5 arena score

#38 of 62 in Text-to-Image

Vote tally

Where the votes landed

Imagen 3.0 Generate 002

0%

win rate

Ties

0%

Wan 2.7

0%

win rate

Shared challenges 6

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Imagen 3.0 Generate 002
Wan 2.7

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Strong minimalist aesthetic with a clean grid layout.
  • + Consistent photography style across all images.
  • + Bold, readable headings that fit the modern design.
  • Text is largely gibberish and unreadable under the main headings.
  • The grid layout feels a bit repetitive with multiple pizza images despite labels for 'Mains'.
  • Lacks specific pricing and detailed item names requested for a functional menu.

Wan 2.7

  • + Excellent typography and legible text throughout.
  • + Better logical structure with clear sections for Appetizers, Pizza, and Mains.
  • + Higher photographic variety showing a complete dining experience.
  • Includes some unnecessary background props (pen, rosemary) that weren't requested.
  • Text rendering on smaller descriptions is slightly blurred but still superior to the competitor.

Verdict: Imagen 3.0 provides a visually striking minimalist layout that feels more like a professional design board, but it fails significantly on text legibility and logical content. Wan 2.1 creates a much more functional and realistic menu with readable text, logical pricing, and distinct food categories, making it the superior version for a dining context.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Imagen 3.0 Generate 002
Wan 2.7

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent photorealistic texture on the bun and patty.
  • + Complex layering of ingredients that feels cohesive to the burger structure.
  • + Text integration is clean and well-positioned at the bottom.
  • The 'exploded' effect is quite stacked and lacks a strong sense of dynamic motion.
  • There is some nonsensical gibberish text inside the starburst and at the bottom.
  • The lighting on the ingredients is a bit flat compared to the fiery background.

Wan 2.7

  • + Highly dynamic 'exploded' composition with ingredients scattered in a more interesting way.
  • + The fiery glowing effect on the title text is very well-executed and matches the prompt perfectly.
  • + Crisp, clean text rendering for all requested elements with no spelling errors.
  • The food textures look slightly more illustrative/rendered than photorealistic.
  • Some ingredients like the pickles and loose seeds feel a bit disconnected from the main burger.
  • The starburst design is slightly cluttered with the extra sun-like rays.

Verdict: Wan 2.7 is the clear winner for its superior text rendering, adherence to the 'fiery glowing effect' in the typography, and a more dynamic sense of motion. While Imagen 3.0 has slightly more realistic food textures, its composition is static and the text includes significant artifacts and gibberish.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Imagen 3.0 Generate 002
Wan 2.7

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Very realistic, hand-drawn chalkboard texture with authentic smudges
  • + Captures the requested 'variations in letter size' and 'slight slant' perfectly
  • Numerous spelling errors including 'Risottto', 'Octpsuin', and 'Choouip'
  • Redundant text at the bottom repeats sentences awkwardly

Wan 2.7

  • + Perfect text accuracy and spelling for all menu items
  • + Clean, legible composition with a warm, cozy café background atmosphere
  • + Beautifully rendered shadows and lighting on the chalkboard surface
  • Text looks slightly more like a digital font than authentic chalk handwriting
  • Missing the 'elegant cursive' style requested for the title

Verdict: Imagen 3.0 provides a more authentic chalkboard feel with natural handwriting imperfections, but it fails significantly on spelling and text coherence. Wan 2.7 delivers perfect text rendering and a better overall aesthetic, even though the lettering looks slightly too uniform for a hand-drawn board.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Imagen 3.0 Generate 002
Wan 2.7

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent cinematic lighting with a soft nebula glow behind the subject.
  • + Highly detailed textures on the astronaut suit and horse's coat.
  • + Well-balanced composition against a deep space background.
  • The astronaut is notably smaller in scale compared to the horse.
  • The reigns appear to be floating or disconnected from the astronaut's hands.

Wan 2.7

  • + Stronger adherence to the 'horse on top' spatial relationship, positioning the group higher over a planet.
  • + Realistic proportions between the astronaut and the horse.
  • + Clear, sharp details on the astronaut's life support pack and the horse's tack.
  • The background composition is a bit cluttered with multiple planets and galaxies.
  • The lighting is somewhat flat and less cinematic than the competitor.

Verdict: Both models failed the 'trick' prompt instruction to have the horse on top of the astronaut, instead providing the literal 'horse-riding astronaut'. Imagen 3.0 provides a much more cinematic and atmospheric image with superior lighting, while Wan 2.7 offers better anatomical proportions and clearer technical details.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Imagen 3.0 Generate 002
Wan 2.7

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Clean 3D render with a consistent toy-like aesthetic
  • + Excellent isometric perspective accuracy
  • + Minimalist layout that adheres well to the 'small raised diorama' instruction
  • Text is placed in the top-left rather than the requested top-center
  • Lack of variety in sushi types compared to Model B

Wan 2.7

  • + Text is perfectly centered as requested in the prompt
  • + Superior material sub-surface scattering and textures for the fish and rice
  • + Included a nice variety of sushi types and garnishes
  • The shrimp sushi has a slightly strange, stylized face looking element on the end
  • The text 'SUSHI' is slightly misaligned with the flag icon

Verdict: Wan 2.7 followed the specific layout instructions much better by placing the text at the top-center, whereas Imagen 3.0 justified it to the left. Additionally, Wan 2.7 provided higher visual fidelity with its PBR materials, making the miniature scene look more premium and detailed.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Imagen 3.0 Generate 002
Wan 2.7

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent fur texture rendering and soft depth of field.
  • + Wholesome and cohesive composition with animals naturally interacting.
  • + Subtle, realistic lighting that enhances the 'hyper-photorealistic' request.
  • The bunny's anatomy is slightly distorted in its rolling pose.
  • The fox kit has a slightly generic feline/canine hybrid appearance.

Wan 2.7

  • + Stronger adherence to the 'god rays' and 'butterfly' elements of the prompt.
  • + Dynamic poses showing active movement across the meadow.
  • + Very clear distinction between the four specified animal types.
  • The kitten's facial expression is slightly unnerving and less 'cute'.
  • The lighting looks more artificial and overly sharpened compared to Model A.
  • Some anatomical issues where the kitten is touching the bunny.

Verdict: Imagen 3.0 Generate 002 creates a more believable, high-quality image with superior fur textures and a softer, more professional lighting style. While Wan 2.7 captures the action of 'chasing and tumbling' more literally, its execution of the animals' faces—particularly the kitten—is less aesthetically pleasing and feels more like a standard AI generation than a masterpiece.

Next steps

Explore each model