Black Forest Labs' state-of-the-art image generation model with maximum quality and speed, supporting text-to-image and multi-reference image editing with up to 4MP output
Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.
FLUX.2 [pro]
#8 of 62 in Text-to-Image
Imagen 3.0 Generate 002
#36 of 62 in Text-to-Image
Where the votes landed
FLUX.2 [pro]
0%
win rate
Ties
0%
Imagen 3.0 Generate 002
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent font legibility with high contrast and readable prices.
- + Distinct color-coded sections (red, yellow, green) for different food types.
- + Photographs look appetizing and are well-integrated into a clean layout.
- − Text duplication issues, repeating 'Pizza' sections and item names.
- − The grid of photos contains some logic errors, such as a steak under the 'Pizza' header.
- − The currency formatting is slightly non-standard with commas in odd places.
Imagen 3.0 Generate 002
- + Beautiful photography with a professional, high-end feel.
- + Consistent grid layout that creates a strong visual rhythm.
- + Accurately represents a modern minimalist aesthetic with effective use of white space.
- − Text is largely gibberish and very difficult to read compared to the other model.
- − The 'grid' format results in a layout where text and photos are scattered, making it less functional as a menu.
- − Poor adherence to section headers, mixing 'Mains' and 'Pizza' randomly throughout.
Verdict: FLUX.2 [pro] is more successful as a menu because its text is legible and its sections are clearly defined by color, despite some repetitive naming errors. Imagen 3.0 Generate 002 produces more artistic photography and a cleaner grid, but fails as a functional design due to illegible text and a confusing organizational structure. FLUX.2 [pro] is the preferred choice for a practical design application.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent typography with perfect prompt adherence and glowing effects.
- + Superior lighting and texture on the patty, resembling professional food photography.
- + Clean composition that balances the text and the product well.
- − The 'exploded' effect is slightly less dynamic than requested, with components appearing neatly stacked.
Imagen 3.0 Generate 002
- + Great sense of motion and 'exploded' dynamics with components flying in different directions.
- + Vibrant color palette and intense fiery background.
- − Failed to render the starburst text correctly, including gibberish characters.
- − Lower photorealism with some food elements appearing slightly plastic or illustrative.
- − Text at the bottom is less integrated into the 'fiery' theme compared to the other model.
Verdict: FLUX.2 [pro] is the clear winner as it followed every instruction, including specific text and pricing, with perfect spelling and high-end photorealism. While Imagen 3.0 exhibited a more dynamic 'exploded' layout, it failed significantly on text accuracy and realism, producing artifacts in the starburst element.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent text accuracy with perfect spelling of all menu items.
- + Features a very realistic chalk texture with dusty specks and natural smudging on a slate-like surface.
- + Consistent handwritten cursive style throughout the entire board.
- − The 'slate' board has slightly rough, irregular edges that might not fit some 'cozy cafe' aesthetics.
Imagen 3.0 Generate 002
- + Clean, well-framed composition within a classic wooden board frame.
- + Good use of layout and spacing for the prices.
- − Significant spelling errors and repetitive lines (e.g., 'Risotttto', 'Octpsuin & ebbs').
- − The text style looks more like a digital chalk font rather than authentic, messy handwriting.
- − Includes repetitive and nonsensical footer text.
Verdict: FLUX.2 [pro] followed the prompt almost perfectly, delivering highly realistic chalk textures and accurate spelling for the complex menu items. In contrast, Imagen 3.0 struggled with basic spelling and generated repetitive lines of text that failed the adherence check. FLUX.2 is the clear winner for its superior text rendering and authentic materials.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.2 [pro]
- + Successfully followed the specific 'horse on top' prompt instruction
- + Highly detailed and cinematic background showing Earth and various galaxies
- + Creative and surreal composition that defies standard tropes
- − Anatomical weirdness with an extra horse torso emerging from the astronaut’s back
Imagen 3.0 Generate 002
- + High visual clarity and clean rendering
- + Good cinematic lighting along the horse's mane
- − Failed the negative constraint; the astronaut is riding the horse instead of vice versa
- − Standard interpretation lacking the requested surreal reversal
Verdict: FLUX.2 [pro] is the clear winner as it successfully interpreted the challenging logic of the prompt, placing the horse on top of the astronaut. In contrast, Imagen 3.0 Generate 002 ignored the specific instruction to reverse the positions, providing a generic astronaut riding a horse.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent typography with perfect centering and large, bold weights
- + Superior 3D modeling feel with more realistic PBS lighting and shadows
- + Clean minimalist composition that strictly adheres to the 'small raised diorama' request
- − One piece of sushi features a scalloped texture on the rice that looks less realistic than the others
Imagen 3.0 Generate 002
- + Good variety of sushi types including nigiri and gunkan
- + Accurate 45-degree isometric perspective
- + Pleasant soft pastel color palette
- − Text is not perfectly centered as requested, ending up in the top left corner
- − Chopsticks were not requested and slightly clutter the 'minimal' scene
- − The text size is smaller than prompted
Verdict: FLUX.2 [pro] followed the layout instructions much more closely, particularly regarding the text placement, size, and centering. While Imagen 3.0 Generate 002 produced a charming illustration, it failed to center the text and included unrequested elements like chopsticks, whereas FLUX.2 [pro] captured the premium 3D diorama aesthetic much more effectively.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent interaction between the animals, such as the kitten reaching out.
- + Incredible fur texture and realistic anatomy for all four animals.
- + Butterflies are integrated naturally into the scene with shadows and realistic placement.
- − The fox looks slightly more like an adult than a 'kit'.
- − The sunrise light is more of a soft haze than distinct 'god rays'.
Imagen 3.0 Generate 002
- + Successfully captures the 'tumbling' and 'playful' aspect of the prompt with the bunny on its back.
- + Includes more visible dew sparkles and more distinct light rays.
- + Animals are positioned more dynamically within the meadow.
- − The bunny has five limbs or poorly defined under-paws.
- − The kitten's facial structure is slightly less realistic compared to the others.
- − The lighting on the animals' fur is a bit flat compared to the complex environment.
Verdict: Both models followed the prompt well, including all four specific animals and the meadow setting. FLUX.2 [pro] produced a more polished and anatomically correct image with superior fur detail, while Imagen 3.0 better captured the requested 'tumbling' action and atmospheric effects like dew and god rays. However, FLUX.2 [pro] is the winner due to the lack of physiological artifacts seen in the bunny of the Imagen 3.0 output.
Explore each model
Google's Imagen 3.0 text-to-image generation model, producing high-quality images with improved detail and lighting