Google's Imagen 3.0 text-to-image generation model, producing high-quality images with improved detail and lighting
Settled by community votes across 3 shared challenges, with an AI judge weighing in on each.
Imagen 3.0 Generate 002
#36 of 62 in Text-to-Image
Qwen Image Max
#35 of 62 in Text-to-Image
Where the votes landed
Imagen 3.0 Generate 002
0%
win rate
Ties
0%
Qwen Image Max
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent 'exploded' layout showing all individual layers as requested.
- + Clean typography that is well-integrated with the ground plane.
- + High photorealistic detail on the textures of the bun and meat.
- − Includes some gibberish text inside the starburst above the price.
- − The lighting on the burger feels slightly static compared to the background.
Qwen Image Max
- + Outstanding fiery text effects that perfectly match the prompt's aesthetic.
- + Dynamic sense of motion with debris and diagonal composition.
- + Perfect text rendering without additional gibberish.
- − The 'exploded' effect is less pronounced, as several ingredients are still touching.
- − The bun looks slightly compressed or skewed at the top.
Verdict: Both models followed the prompt well, but Imagen 3.0 provided a superior 'exploded' view that clearly showcased every layer of the burger. However, Qwen Image Max captured the requested 'fiery, glowing effect' for the text much more effectively and avoided the small artifacts and gibberish text found in the first image.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent adherence to the cartoon 3D miniature aesthetic
- + Perfectly clean layout and text positioning
- + Consistent 45-degree isometric projection
- − Text layout is top-left rather than the requested top-center
- − Simplified textures lack the 'realistic PBR' quality requested
Qwen Image Max
- + Successfully combines realistic textures with a 3D cartoon style
- + Follows text placement and layout instructions perfectly
- + High material quality on the glass base and sushi toppings
- − The perspective is slightly lower than a true 45-degree isometric top-down view
- − Small artifacts present in the text rendering (minor shadow issues)
Verdict: Qwen Image Max followed the spatial instructions more closely, placing the text at the top-center and achieving a better balance between cartoon style and realistic PBR materials. While Imagen 3.0 captured the 'miniature 3D' look very well, it failed to center the text and the textures were a bit too flat compared to the prompt's request for realism.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
Imagen 3.0 Generate 002
- + Excellent anatomical accuracy for all four distinct animals.
- + Naturally integrated lighting and realistic dew effects.
- + Highly detailed fur texture and believable expressive eyes.
- − The bunny is upside down which looks a bit awkward in its placement.
Qwen Image Max
- + Dymanic composition with many butterflies and clear god rays.
- + Vibrant color palette that emphasizes the 'joyful wholesome' theme.
- − Failed to include the requested baby bunny, instead adding a second golden retriever puppy.
- − The fox has anatomical issues with its paws and leg joints.
- − The kitten's tail is unnaturally thick and striped more like a raccoon.
Verdict: Imagen 3.0 successfully included all four requested animals with high photorealism and anatomical accuracy. Qwen Image Max failed the prompt by omitting the baby bunny and produced several anatomical artifacts, such as the fox's distorted front paws and the kitten's strange tail.
Explore each model
The Max series of Tongyi Qwen’s image generation model excels across a wide range of generation tasks. Compared with the Plus series, it significantly reduces the “AI-like” feel in generated images, enhancing their realism. It delivers more lifelike material textures for human subjects, finer and more detailed natural textures, and more visually appealing text rendering.