Head to head
Esc

Models · slot A

to navigate to pick

Imagen 3.0 Generate 002 Google LongCat-Image Meituan

Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.

Imagen 3.0 Generate 002

21.0 arena score

#36 of 62 in Text-to-Image

Skill signature · Text-to-Image

LongCat-Image

9.6 arena score

#62 of 62 in Text-to-Image

Vote tally

Where the votes landed

Imagen 3.0 Generate 002

0%

win rate

Ties

0%

LongCat-Image

0%

win rate

Shared challenges 6

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Imagen 3.0 Generate 002
LongCat-Image

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Strict adherence to the 4x4 grid layout requested in the prompt.
  • + Professional typography and layout that feels like a real, balanced menu design.
  • + High-quality food photography that fits the grid boxes perfectly.
  • Text is mostly gibberish/placeholder characters despite looking geometrically correct.
  • Limited variety in food types, leaning very heavily into pizza images.

LongCat-Image

  • + Dynamic use of bold color accents as requested.
  • + Good resolution on the individual food photos.
  • Chaotic and unbalanced layout that does not follow the grid request.
  • Significant text distortion and illegible characters.
  • Does not effectively include the requested categories in a logical order.

Verdict: Imagen 3.0 provides a much more professional and coherent layout that strictly follows the prompt's request for a grid design and specific menu sections. While both models failed at generating legible text, Imagen 3.0 maintained a sophisticated aesthetic suitable for a real restaurant, whereas LongCat-Image produced a cluttered and disorganized layout with poor typographic control.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Imagen 3.0 Generate 002
LongCat-Image

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent adherence to the 'exploded' layout with all components suspended.
  • + Highly photorealistic textures on the lettuce, sesame bun, and patty.
  • + Excellent inclusion and placement of the required text in a cohesive ad layout.
  • Includes some gibberish text inside the starburst and above the price.
  • The cheese looks a bit plastic-like compared to the other realistic textures.

LongCat-Image

  • + Strong 'fiery and glowing' effect on the typography.
  • + Sharp rendering of the main logo text.
  • + Atmospheric foreground with burning charcoal.
  • Failed the 'exploded' burger requirement as components are mostly stacked.
  • The sauce on top of the cheese looks somewhat digital and unnatural.
  • Lacks the sense of dynamic motion requested in the prompt.

Verdict: Imagen 3.0 Generate 002 is the clear winner as it followed the 'exploded' burger instruction perfectly, creating a dynamic and professional ad layout. LongCat-Image failed to separate the burger components and produced a more static image that didn't fully capture the requested energy, despite having better glowing text effects.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Imagen 3.0 Generate 002
LongCat-Image

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent text legibility and spelling accuracy.
  • + The chalk texture and handwriting style look authentic and varied.
  • + Follows the specific date and price instructions almost perfectly.
  • Repeats similar menu items (e.g., two versions of Grilled Octopus/Brown Butter) rather than listing them distinctly.
  • Spells 'Risotto' as 'Risotlto'.

LongCat-Image

  • + Effective cafe environment background for context.
  • + Bold chalk stroke aesthetics.
  • Severe spelling hallucinations and garbled text (e.g., 'Toays Gtays', 'Arlil').
  • Fails to render the specific menu items requested in the prompt.
  • Poor layout with centered prices breaking up the flow of the text.

Verdict: Imagen 3.0 provides a much more functional and accurate response, correctly rendering the majority of the complex requested text, including the specific date and menu items. LongCat-Image fails significantly on prompt adherence, producing mostly nonsensical gibberish instead of the requested phrases.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Imagen 3.0 Generate 002
LongCat-Image

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent cinematic lighting and atmosphere with a high-quality nebula background.
  • + Coherent anatomy for both the horse and the astronaut.
  • + Superior visual clarity and detail in the space suit and horse's coat.
  • Fails the specific negative constraint requiring the horse to be on top of the astronaut.

LongCat-Image

  • + Includes more background elements like planetary bodies and space hardware.
  • + Correctly interprets the space setting.
  • Fails the negative constraint requiring the horse to be on top of the astronaut.
  • Major anatomical issues with the horse's legs appearing to morph or float incorrectly.
  • Contains nonsensical artifacts like the distorted ship in the upper right corner.

Verdict: Both models failed the specific prompt instruction to place the horse on top of the astronaut (a reversal of the typical scene), with both instead generating an astronaut riding a horse. However, Imagen 3.0 Generate 002 is the superior image due to its professional cinematic quality, better anatomy, and lighting, whereas LongCat-Image suffers from significant anatomical distortions and messy background artifacts.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Imagen 3.0 Generate 002
LongCat-Image

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent 45° isometric perspective following true grid lines.
  • + Perfect adherence to text placement instructions.
  • + Clean, consistent 3D rendering with soft clay-like textures.
  • Text and icon are slightly offset to the left rather than 'top-center'.
  • The rice grain texture is a bit large and repetitive.

LongCat-Image

  • + Strong text rendering and better vertical alignment of 'top-center' text.
  • + Beautiful material work on the salmon and tuna textures.
  • + Dynamic interpretation of the flag icon with a flagpole.
  • Perspective is more of a front-view 3D than a true isometric overhead.
  • Cropped slightly at the bottom, losing the diorama effect.
  • The flag icon is partially cutting into the 'I' in 'SUSHI'.

Verdict: Imagen 3.0 provides a superior isometric layout that perfectly captures the 'miniature diorama' request with a consistent 45-degree angle. While LongCat-Image has more interesting lighting and finer material textures on the food, it fails to achieve the requested isometric perspective and crops the base of the scene.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Imagen 3.0 Generate 002
LongCat-Image

AI Judge Analysis

Imagen 3.0 Generate 002

  • + Excellent anatomical accuracy for all four animals.
  • + Beautiful lighting with visible dew drops and soft bokeh.
  • + Realistic fur textures and natural color palette.
  • The fox kit's size relative to the cat and puppy is a bit large.
  • The lighting is slightly hazy across the entire frame.

LongCat-Image

  • + Strong 'god rays' lighting effects that match the prompt.
  • + Vibrant and cheerful color palette.
  • Serious anatomical failure where the cat has long bunny ears.
  • The bunny mentioned in the prompt is missing as a separate animal.
  • The butterflies look like flat stickers rather than integrated 3D objects.

Verdict: Imagen 3.0 successfully rendered all four requested animals with high anatomical realism and a cohesive, high-quality aesthetic. LongCat-Image failed significantly on the prompt details, merging the cat and bunny into a single creature and missing one animal entirely, while also displaying lower visual realism.

Next steps

Explore each model