Head to head
Esc

Models · slot A

to navigate to pick

Imagen 3.0 Generate 002 Google Qwen Image 2512 Alibaba

Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.

Imagen 3.0 Generate 002

21.0 arena score

#36 of 62 in Text-to-Image

Skill signature · Text-to-Image

Qwen Image 2512

22.9 arena score

#30 of 62 in Text-to-Image

Vote tally

Where the votes landed

Imagen 3.0 Generate 002

0%

win rate

Ties

0%

Qwen Image 2512

0%

win rate

Shared challenges 6

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Imagen 3.0 Generate 002
Qwen Image 2512

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent grid layout that integrates photos and text seamlessly.
  • + Clean, professional typography that aligns with the minimalist request.
  • + High-quality, realistic food photography with consistent lighting.
  • Small text blocks are mostly illegible placeholder characters.
  • The layout is square rather than the typical rectangular menu format.

Qwen Image 2512

  • + Strong use of vibrant color accents through graphic icons.
  • + Clear sectioning of the menu that is easy to navigate visually.
  • + Good adherence to the 'casual dining' vibe with bold header fonts.
  • Text rendering is very poor with significant gibberish and typos.
  • Food images have slightly more AI-generated artifacts compared to Model A.

Verdict: Imagen 3.0 Generate 002 produces a much more professional and realistic output with a sophisticated grid system and high-quality photography. While Qwen Image 2512 captures the 'vibrant accents' request well with colorful icons, its text rendering is significantly worse, and the layout feels more cluttered and less modern than Imagen 3.0.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Imagen 3.0 Generate 002
Qwen Image 2512

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent exploded view with clear separation of every layer
  • + Superior lighting on the food items making them look appetizing
  • + Accurate inclusion of all requested elements including the starburst
  • Includes gibberish text inside the starburst above the price
  • The 'MAGIC BURGER' text uses a gold texture rather than the requested fiery glowing effect

Qwen Image 2512

  • + Perfectly executed fiery, glowing text for the main title
  • + High energy 'motion' feel with flying crumbs and droplets
  • + Correct text rendering for 'MAGIC BURGER' and 'LIMITED ONLY'
  • The burger is not fully 'exploded' as several layers are still touching and messy
  • Missing the word 'TIME' in the secondary 'LIMITED ONLY' message

Verdict: Imagen 3.0 Generate 002 is the overall winner because it captures the 'exploded burger' concept much more clearly, showing each individual ingredient suspended in air as requested. While Qwen Image 2512 has a more impressive fiery text effect, its burger layout is cluttered and it failed to include the word 'TIME' in the required secondary text.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Imagen 3.0 Generate 002
Qwen Image 2512

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent chalk texture and board smears for realism
  • + Accurate date rendering
  • + Contains very believable handwriting variations
  • Significant spelling errors and repetitive lines (e.g., 'Risotttto', 'Octpsuin & ebbs')
  • Bottom text repeats almost the exact same sentence twice
  • Title font is a hollow outline style rather than the requested elegant cursive

Qwen Image 2512

  • + Near-perfect spelling of complex menu items like 'Grilled Octopus with Lemon & Herbs'
  • + Beautiful and consistent cursive handwriting that matches the prompt's tone
  • + Superior composition with a more natural 'cozy café' background
  • Slight misspelling of 'Risotto' as 'Risitto'
  • Text appears slightly too clean, appearing more like a digital brush than physical chalk

Verdict: Qwen Image 2512 is the clear winner as it successfully follows the complex text-heavy prompt with minimal spelling errors and beautiful cursive handwriting. In contrast, Imagen 3.0 suffers from major legibility issues, including repeating words and garbled text across the middle and bottom sections of the board.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Imagen 3.0 Generate 002
Qwen Image 2512

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent cinematic lighting and texture on the horse's coat.
  • + Strong composition with a beautiful nebula background creating a sense of depth.
  • Failed the negative constraint; the astronaut is riding the horse instead of the horse riding the astronaut.
  • Missing the 'surreal' inversion requested in the prompt.

Qwen Image 2512

  • + High level of detail on the space suit and horse's bridle.
  • + Realistic rendering of the horse and astronaut textures.
  • Failed the negative constraint; completely ignored the instruction for the horse to be on top.
  • The composition feels slightly cramped compared to the other image.
  • Visual artifacting where the astronaut's right arm meets the horse's neck.

Verdict: Both Imagen 3.0 and Qwen Image 2512 completely failed the core logical challenge of the prompt, which specifically requested the horse to be on top of the astronaut. Instead, both models produced a standard 'astronaut riding a horse' image. Imagen 3.0 is slightly better overall due to its superior cinematic lighting and more aesthetically pleasing background.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Imagen 3.0 Generate 002
Qwen Image 2512

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent 3D miniature clay-like aesthetic
  • + Strict adherence to the 45-degree isometric perspective
  • + Clean and professional font rendering
  • Text and icon are placed in the top-left rather than the requested top-center
  • Simplified textures lean more towards cartoonish than realistic PBR

Qwen Image 2512

  • + Perfect text placement at top-center as requested
  • + Incorporates realistic PBR-style textures into the cartoon scene
  • + Better central composition of the diorama
  • Perspective is slightly more tilted than a true 45-degree isometric view
  • Includes extra greenery and garnish that was requested to be minimal

Verdict: Qwen Image 2512 followed the layout instructions more accurately, specifically placing the text and flag in the top-center. While Imagen 3.0 has a cleaner isometric look, Qwen Image 2512 captures the 'miniature 3D' diorama feel and material richness more effectively.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Imagen 3.0 Generate 002
Qwen Image 2512

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent depiction of playful movement with the bunny tumbling.
  • + Natural lighting with subtle bloom and environmental integration.
  • + High-quality fur textures and realistic anatomy for all four animals.
  • The fox kit's face is slightly less detailed compared to Northern European red fox features.
  • The god rays are very subtle and almost indistinguishable from general morning haze.

Qwen Image 2512

  • + Stronger adherence to the 'god rays' element of the prompt with distinct light beams.
  • + Very expressive 'big eyes' on all animals as requested.
  • + Crisp details on the butterflies and flowers.
  • Anatomical errors, specifically the puppy appears to have five legs/paws visible in the foreground cluster.
  • The composition feels more like a static portrait than the requested 'playfully chasing' scene.

Verdict: Imagen 3.0 Generate 002 is the superior image because it successfully captures the energy of animals 'tumbling together' with realistic anatomy and a high level of photographic polish. While Qwen Image 2512 has very cute expressions and bold lighting, it suffers from a significant anatomical hallucination involving extra paws under the puppy's face.

Next steps

Explore each model