Head to head
Esc

Models · slot A

to navigate to pick

Imagen 3.0 Generate 002 Google Imagen 4.0 Generate 001 Google

Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.

Imagen 3.0 Generate 002

21.0 arena score

#36 of 62 in Text-to-Image

Skill signature · Text-to-Image

Imagen 4.0 Generate 001

17.1 arena score

#55 of 62 in Text-to-Image

Vote tally

Where the votes landed

Imagen 3.0 Generate 002

0%

win rate

Ties

0%

Imagen 4.0 Generate 001

0%

win rate

Shared challenges 6

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Imagen 3.0 Generate 002
Imagen 4.0 Generate 001

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Strict adherence to the 4x4 grid layout requested in the prompt.
  • + Professional typography and layout that resembles a real menu poster.
  • + Consistent food photography style across all tiles.
  • The text is largely gibberish despite having a good visual weight.
  • Low variety in the food itself, as almost every photo features a pizza.

Imagen 4.0 Generate 001

  • + Excellent typography with legible, clean labels for sections.
  • + Better variety of food photography including salads and plated mains to match the headers.
  • + Effective use of 'vibrant accents' through the geometric border elements.
  • The layout is not a strict grid, which slightly deviates from the prompt's 'grid' request.
  • Less 'minimalist' than the other model due to many overlapping geometric shapes.

Verdict: Imagen 3.0 Generate 002 creates a more cohesive grid design that perfectly matches a minimalist aesthetic, but fails to provide food variety. Imagen 4.0 Generate 001 is preferred because it delivers much higher quality text rendering and shows a relevant variety of food (appetizers, mains, and pizzas) that actually corresponds to the section headers.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Imagen 3.0 Generate 002
Imagen 4.0 Generate 001

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent integrated lighting that reflects the fiery background onto the food.
  • + Higher photorealistic quality in the texture of the meat and bun.
  • + Very clean starburst element for the price tag.
  • Included significant gibberish text inside the starburst badge.
  • The composition feels slightly static and vertically compressed.

Imagen 4.0 Generate 001

  • + Perfect text rendering for all requested strings without any gibberish.
  • + Stronger sense of motion and dynamic layering with the angled burger components.
  • + Correctly applied the fiery, glowing effect to the starburst and primary text.
  • The burger patty texture looks a bit more plasticky compared to Model A.
  • The lighting on the food is a bit flat and does not fully interact with the background embers.

Verdict: Imagen 3.0 Generate 002 produces a more lifelike food image with superior internal lighting, but fails on text accuracy by adding nonsensical words. Imagen 4.0 Generate 001 provides a much more dynamic composition and perfect text integration, making it the more effective advertisement, despite slightly lower texture realism on the meat.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Imagen 3.0 Generate 002
Imagen 4.0 Generate 001

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent realistic chalk texture with dusty smudges and varying pressure.
  • + The chalkboard itself looks authentic with wood grain and board wear.
  • + Achieved a nice cursive-style script for the main menu items.
  • Several spelling errors like 'Risottto' and 'Octpsuin'.
  • Included duplicate lines for the same items instead of moving to the next requested item.

Imagen 4.0 Generate 001

  • + Successfully spelled most menu items correctly, including the full 'Brown Butter Chocolate Chip Cookies'.
  • + Very clean and legible handwriting style.
  • + Followed the prompt's structural requirements for the footer better.
  • Included meta-commentary text from the prompt (e.g., 'Tittle Menu', 'Footer') directly onto the board.
  • Text looks a bit more like a digital font than actual chalk on a board.
  • Spelling error in the word 'Berbs' instead of 'Herbs'.

Verdict: Imagen 3.0 (Image A) captures the authentic aesthetic of a chalk board much better, with realistic textures and smudges, but fails significantly on spelling and content repetition. Imagen 4.0 (Image B) has much higher prompt adherence regarding the specific text requested, but it mistakenly treats the prompt's instructions as text to be written on the board (rendering words like 'Title' and 'Footer'). While Imagen 4.0 is more accurate with the specific menu items, Imagen 3.0 is a better visual representation of the requested medium.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Imagen 3.0 Generate 002
Imagen 4.0 Generate 001

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent anatomical rendering of the horse and lighting integration with the nebula.
  • + Highly cinematic composition with a strong sense of floating in deep space.
  • Fails the specific logic prompt (horse on top of astronaut).

Imagen 4.0 Generate 001

  • + Beautiful creative touches like the nebulas within the horse's fur and the visor reflection.
  • + Dynamic composition with the curvature of the Earth and light streaks.
  • Fails the specific logic prompt (horse on top of astronaut).
  • Minor anatomical artifacts on the horse's rear hooves/legs.

Verdict: Both Imagen 3.0 and Imagen 4.0 failed to follow the specific spatial instruction to place the horse on top of the astronaut, instead defaulting to the standard interpretation of the phrase. However, Imagen 4.0 is slightly more visually impressive due to its creative use of color, light streaks, and surreal textures within the horse's coat.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Imagen 3.0 Generate 002
Imagen 4.0 Generate 001

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Perfect adherence to the requested isometric 3D cartoon style.
  • + Flawless text rendering and placement as specified in the prompt.
  • + Accurate solid light blue background and small raised diorama base.
  • The textures are more plastic-like than realistic PBR textures.
  • The sushi shapes are slightly stylized and simplified compared to real food.

Imagen 4.0 Generate 001

  • + Excellent realistic PBR materials and lighting on the food and plate.
  • + Very high clarity and level of detail in the rice and fish textures.
  • Completely failed to include the requested 'JAPAN' and 'SUSHI' text and flag icon.
  • Ignored the '3D cartoon' and 'isometric' style instructions in favor of realism.
  • Background is a gradient off-white/grey rather than the requested solid light blue.

Verdict: Imagen 3.0 (Model A) followed every specific layout and text instruction, delivering a clean isometric cartoon scene that perfectly matches the prompt's aesthetic. Imagen 4.0 (Model B) ignored almost all stylistic and textual constraints, producing a realistic photo of sushi instead of a stylized 3D diorama with text. Consequently, Imagen 3.0 is the clear winner for prompt adherence.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Imagen 3.0 Generate 002
Imagen 4.0 Generate 001

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent photorealistic fur texture and lighting.
  • + Natural, grounded interaction between the four distinct animals.
  • + Beautifully soft bokeh and realistic dew drop effects.
  • The fox kit lacks some of the characteristic fiery red color requested.
  • The 'tumbling' action is a bit static compared to the high-energy request.

Imagen 4.0 Generate 001

  • + Stronger adherence to the 'tumbling' and 'chasing' action described in the prompt.
  • + Vibrant color palette and very clear 'god rays' lighting.
  • + Accurate species identification and distinct coloration for each animal.
  • Leans more towards digital illustration or an AI-generated greeting card style rather than 'hyper-photorealistic'.
  • Anatomical oddity with the kitten having five limbs/paws visible in the central cluster.
  • The scale of the flowers relative to the animals feels slightly inconsistent.

Verdict: Imagen 3.0 Generate 002 is the superior image due to its impressive photorealism and coherent anatomy, capturing a sense of genuine nature photography. While Imagen 4.0 Generate 001 interprets the 'tumbling' action better, it fails the realism criteria and contains a significant anatomical error with the kitten's paws.

Next steps

Explore each model