Google's Imagen 3.0 text-to-image generation model, producing high-quality images with improved detail and lighting
Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.
Imagen 3.0 Generate 002
#43 of 62 in Text-to-Image
Qwen Image 2.0
#34 of 62 in Text-to-Image
Where the votes landed
Imagen 3.0 Generate 002
0%
win rate
Ties
0%
Qwen Image 2.0
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent grid layout with clear typographic hierarchy
- + Includes specific sections for pizza, mains, and appetizers as requested
- + Professional use of white space and minimalist graphic design
- − Text is mostly illegible gibberish
- − Food photos are somewhat repetitive with too many pizza shots
Qwen Image 2.0
- + Clean, modern aesthetic with consistent image styling
- + Better text rendering for headers
- + Each photo is high-quality and unique
- − Layout is a bit repetitive compared to Model A
- − Text below items is garbled
- − Headers are floating at the top rather than integrated into a complete menu structure
Verdict: Imagen 3.0 Generate 002 creates a more realistic menu layout with distinct boxes for different sections and a professional typographic grid, whereas Qwen Image 2.0 provides better individual food photography but feels more like a simple photo gallery. Imagen 3.0 is preferred for capturing the complexity and structural requirements of the 'restaurant menu' prompt.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent layout with a centralized, clearly layered 'exploded' view of the ingredients.
- + Vibrant color palette and high-quality rendering of food textures.
- + Perfect adherence to all text requirements including price and secondary message.
- − Includes some gibberish text inside the starburst above the price.
- − The background fire feels a bit more like a static texture compared to the burger.
Qwen Image 2.0
- + Stronger 'fiery glowing effect' applied to the text as requested.
- + Realistic moisture on the burger patty and dripping sauce.
- + Good sense of heat and atmosphere with smoke and embers.
- − The burger is not as effectively 'exploded'; many ingredients are still touching each other.
- − The composition feels a bit cramped at the top with the text very close to the bun.
Verdict: Imagen 3.0 Generate 002 is the preferred result because it better captures the 'exploded' nature of the burger, showing each component suspended clearly in mid-air. While Qwen Image 2.0 did a better job with the fiery text effect and realistic lighting on the meat, its composition was less dynamic and didn't follow the 'exploded' instruction as literally.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent chalk texture and board residue for realism.
- + Clear, legible handwriting style that remains consistent.
- − Significant spelling errors throughout the menu items, including 'Risottto' and 'Octpsuin'.
- − Failed the cursive requirement for the title, using a block-style outline font instead.
- − The footer text is repetitive and nonsensical.
Qwen Image 2.0
- + High prompt adherence, correctly using cursive for the title as requested.
- + Nearly perfect spelling across even complex menu items.
- + Superior composition that places the menu board within a 'cozy café' environment.
- − Slightly less 'chalky' texture in the letter strokes compared to the other model.
Verdict: Qwen Image 2.0 is the clear winner as it followed all specific prompt instructions, including the cursive title requirement and accurate spelling of the menu items. While Imagen 3.0 captured a very realistic chalk texture, its frequent spelling errors and failure to use cursive for the heading made it less successful overall.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent cinematic lighting and nebula background.
- + Consistent anatomical detail on both the horse and the spacesuit.
- + Captures a calm, surreal atmosphere with high resolution.
- − The horse's legs have slightly muffled hoof shapes.
Qwen Image 2.0
- + Creative use of texture with scale-like patterns on the horse's neck.
- + Good sense of depth with the Earth in the background and floating droplets.
- + Strong adherence to the surreal requirement.
- − Anatomical error with the horse having five legs visible.
- − The reins are tangled and clipping through the horse's neck and the astronaut's hand.
Verdict: While both models failed the logic trap to place the horse on top of the astronaut, Imagen 3.0 (Generate 002) is the superior image due to its technical coherence and cinematic quality. Qwen Image 2.0 includes major anatomical errors, specifically an extra leg on the horse, and messy rendering of the reins.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent adherence to the '3D cartoon' and 'miniature' style requests.
- + The isometric perspective and composition are very precise.
- + Text rendering is clean and well-integrated into the design.
- − The text is placed in the top-left rather than the requested top-center.
- − Colors are slightly muted compared to the vibrancy usually expected in PBR renders.
Qwen Image 2.0
- + Perfect text and icon placement at the top-center as requested.
- + Realistic materials and textures are very high quality.
- + Central composition is strong and balanced.
- − Failed to follow the 'cartoon' style request, opting for a photorealistic look instead.
- − The perspective is more of a standard low-angle shot than the requested 45-degree isometric view.
Verdict: Imagen 3.0 followed the stylistic 'cartoon miniature' instructions much better, resulting in a cohesive diorama that looks like a designed asset. Qwen Image 2.0 followed the layout and text placement instructions more accurately but ignored the specific art style, producing a realistic photo instead of the requested 3D cartoon scene.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent fur texture rendering and overall sharpness.
- + Better adherence to the 'dew sparkles' requirement with visible bokeh effects in the grass.
- + Cohesive lighting and facial expressions that convey a wholesome vibe.
- − The 'tumbling' action is a bit static compared to the prompt.
- − The bunny has five legs/limbs visible due to anatomical confusion under the dog.
Qwen Image 2.0
- + Dynamic 'tumbling' and 'chasing' action that better captures the playful intent.
- + Strong execution of 'god rays' and backlighting from the sunrise.
- + All four requested animals are present with distinct, energetic poses.
- − The fox kit's face is somewhat distorted as it rolls over.
- − The kitten's paws have anatomical issues with extra digits and merged fur textures.
- − The butterflies appear a bit flat and pasted-on compared to the animals.
Verdict: Imagen 3.0 Generate 002 produces a cleaner, more aesthetically pleasing 'portrait' with superior fur textures and lighting, though it suffers from a significant anatomical error in the bunny's limbs. Qwen Image 2.0 captures the action of the prompt much better, showing the animals actively tumbling and interacting, even if the fine details are slightly less refined. Imagen 3.0 is preferred for its overall visual clarity and professional finish despite the limb error.
Explore each model
Alibaba's Qwen Image 2.0 model with enhanced text rendering, supporting both Chinese and English prompts with up to 6 images per request