Google's Imagen 3.0 text-to-image generation model, producing high-quality images with improved detail and lighting
Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.
Imagen 3.0 Generate 002
#36 of 62 in Text-to-Image
OmniGen v2
#57 of 62 in Text-to-Image
Where the votes landed
Imagen 3.0 Generate 002
0%
win rate
Ties
0%
OmniGen v2
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Features a very clean and professional grid layout that feels like a real modern menu.
- + The food photography is high quality and consistent in style.
- + Text hierarchy and spacing are well-executed for a minimalist design.
- − The text content is largely gibberish despite the professional look.
- − Repetitive use of pizza images limits the variety of the 'mains' section.
OmniGen v2
- + Uses vibrant color accents as requested in the prompt.
- + Provides a good variety of food types including soup and appetizers.
- + Creative use of a two-page spread layout.
- − Very poor spelling and font rendering compared to the other model.
- − The grid layout is inconsistent and less professional looking.
- − Visual artifacts are present in the food imagery.
Verdict: Imagen 3.0 Generate 002 is the clear winner for its superior professional layout and high-quality food photography that perfectly captures the 'modern minimalist' aesthetic. While OmniGen v2 attempted more vibrant color accents, its poor text rendering and disjointed composition make it less effective as a design piece.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent adherence to the 'exploded' suspended animation request.
- + Highly photorealistic textures on the lettuce, bun, and beef patty.
- + Superior integration of all three required text elements with correct spelling.
- − Includes some gibberish text inside the starburst above the price.
- − The flying sauce droplets look slightly artificial compared to the burger.
OmniGen v2
- + Strong graphic design feel with clear focal points.
- + Vibrant fiery lighting effect that pop against the dark background.
- − Completely ignored the request for an 'exploded' burger with components in mid-air.
- − The 'LIMITED TIME ONLY' text is partially cut off on the left side.
- − Missing the Euro symbol (€) requested in the prompt.
Verdict: Imagen 3.0 follows the complex structural requirements of the prompt perfectly, delivering a high-quality 'exploded' burger with impressive realism. OmniGen v2 fails the core composition request by providing a static, assembled burger and has significant issues with text rendering and completeness.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent chalk texture with natural-looking smudges and varied pressure.
- + Accurately rendered the specific date requested in the prompt.
- + Superior typography that maintains readability despite minor spelling errors.
- − Includes several spelling errors like 'Risotttto' and 'Octpsuin'.
- − Duplicates menu items rather than listing them uniquely as intended.
OmniGen v2
- + Successfully captured a bold, handwritten chalk style.
- + Maintains a clean and centered composition.
- − Severe spelling and legibility issues in almost every word.
- − The text looks more like a digital font overlay than natural chalk on a board.
- − Failed to correctly format the menu items, leading to jumbled pricing and text.
Verdict: Imagen 3.0 Generate 002 produces a far more realistic image, correctly capturing the texture and smudging characteristic of actual chalkboard writing. While it struggles with some spelling and duplicate lines, OmniGen v2 fails significantly on legibility, spelling, and logical item placement.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent cinematic lighting and realistic nebulae in the background.
- + High level of detail on the space suit and the horse's musculature.
- + Successful interpretation of 'horse riding astronaut' in a surreal space environment.
- − Technically failed the specific logic check 'horse on top, not vice versa' requested in the prompt.
- − The reins are clipping through the astronaut's hand.
OmniGen v2
- + Satisfactory horse anatomy and clean lines.
- + Dynamic use of multiple moons in the background.
- − Failed the specific logic check 'horse on top, not vice versa' requested in the prompt.
- − The visual style is more illustrative and flat rather than the requested 'cinematic' and 'highly detailed'.
- − The shadows on the astronaut and horse do not match the background lighting sources.
Verdict: Both models failed the negative constraint/inverted logic of the prompt ('horse on top, not vice versa'), which was a trick challenge to see if they could place a horse on an astronaut's back. However, Imagen 3.0 Generate 002 is the superior image as it adhered much better to the stylistic descriptors of 'cinematic' and 'highly detailed' with beautiful rendering of space and realistic textures, whereas OmniGen v2 produced a more basic, clip-art style image.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent 3D miniature clay-like aesthetic
- + Accurate rendering of the Japanese flag
- + Clean typography integrated into the scene top-left
- − Text is placed in the top-left corner instead of the requested top-center
- − The color palette is slightly muted compared to the vibrant request
OmniGen v2
- + Perfect top-center text alignment with requested bold font
- + Vibrant colors and high-quality 3D shading
- + Accurate isometric diorama perspective
- − The flag icon is incorrect, featuring a yellow/red horizontal stripe instead of a red circle
- − The sushi design is slightly repetitive with only two identical pieces
Verdict: Imagen 3.0 provides a more authentic and varied sushi platter with a correct flag icon, but fails the specific layout instruction for top-center text. OmniGen v2 hits the layout and typography requirements perfectly and offers higher visual contrast, though it fails significantly on the flag icon detail.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent photorealistic textures of fur and grass
- + Captures all four requested animals correctly
- + Natural lighting and realistic god rays with dew sparkles
- − The fox kit has a slightly awkward extra leg or paw appearing behind it
OmniGen v2
- + Strong expressive eyes as requested in the prompt
- + Vibrant colors and clear sunburst effect
- − Failed to include the rabbit, only showing three animals
- − Very stylized/cartoonish appearance rather than hyper-photorealistic
- − Poor anatomy on animals, especially the cat's face and the fox's body
Verdict: Imagen 3.0 successfully adhered to the complex prompt's detail requirements, including all four specific animals in a highly realistic style. OmniGen v2 failed on multiple counts: missing one animal entirely and producing a cartoon-like 3D render style instead of the requested photorealism.
Explore each model
Unified multimodal model for text-to-image generation, instruction-guided image editing, personalized generation, and virtual try-on