Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 [schnell] Black Forest Labs Imagen 3.0 Generate 002 Google

Settled by community votes across 6 shared challenges, with an AI judge weighing in on each.

FLUX.1 [schnell]

18.7 arena score

#48 of 62 in Text-to-Image

Skill signature · Text-to-Image

Imagen 3.0 Generate 002

21.0 arena score

#36 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 [schnell]

0%

win rate

Ties

0%

Imagen 3.0 Generate 002

0%

win rate

Shared challenges 6

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 [schnell]
Imagen 3.0 Generate 002

AI Judge Analysis

FLUX.1 [schnell]

  • + Strong minimalist aesthetic with clean white space.
  • + Legible main category headers.
  • + Good use of bold sans-serif typography.
  • Smaller body text is garbled and illegible.
  • Food photos are repetitive and look slightly low-resolution.
  • Includes a nonsensical category header 'ORFEFUS'.

Imagen 3.0 Generate 002

  • + Professional grid layout that feels like a real casual dining menu.
  • + Diverse and high-quality food photography.
  • + Excellent use of the grid system to balance text and images.
  • Several spelling errors in headers like 'APPETIEES' or 'MANS/PIZZA'.
  • Text within sections is largely decorative gibberish.
  • The layout is quite dense, arguably pushing the boundaries of 'minimalist'.

Verdict: Imagen 3.0 Generate 002 is the superior choice because its composition and image quality far exceed FLUX.1 [schnell], creating a layout that feels like a genuine professional menu template. While both models struggle with the specific text content, Imagen 3.0 provides a much more vibrant and appetizing visual presentation that better fits the 'casual dining' requirement.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 [schnell]
Imagen 3.0 Generate 002

AI Judge Analysis

FLUX.1 [schnell]

  • + Excellent photorealistic lighting and texture on the burger bun and melting cheese.
  • + Dynamic sense of motion with flying croutons or ingredients.
  • + High-quality rendering of the fiery background embers.
  • Significant text failures including a misspelling ('AGIC') and redundant, incorrect prices.
  • Horizontal burger alignment doesn't fully capture an 'exploded' view of all components.
  • The starburst element is poorly integrated and lacks the fiery effect.

Imagen 3.0 Generate 002

  • + Perfectly executed 'exploded' view showing every requested layer clearly.
  • + Flawless text rendering for 'MAGIC BURGER', 'LIMITED TIME ONLY', and the correct price.
  • + Strong artistic coherence with the fiery theme applied to both text and background.
  • The food textures look slightly more digital/artificial compared to Model A.
  • The background floor lacks the high-definition detail seen in the embers of the competitor.

Verdict: Imagen 3.0 Generate 002 is the clear winner as it followed every instruction, including the complex exploded layout and specific text requirements, with perfect accuracy. FLUX.1 [schnell] produced a more realistic-looking burger patty and bun, but it failed significantly on the text rendering and the vertical separation of the ingredients.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 [schnell]
Imagen 3.0 Generate 002

AI Judge Analysis

FLUX.1 [schnell]

  • + Features a clear chalk-like texture on the board surface.
  • + The handwriting has natural variations and looks authentically like chalk.
  • Numerous spelling errors including 'Pril', 'Taffle', and 'Octtoopus'.
  • Fails the cursive instruction for the title.
  • The text becomes nonsensical at the bottom of the board.

Imagen 3.0 Generate 002

  • + Much better spelling accuracy including the full date 'APRIL 30, 2026'.
  • + Follows the cursive instruction for the menu items effectively.
  • + Excellent chalk texture and realistic wood grain on the frame.
  • The title is outlined/shadowed rather than simple elegant cursive.
  • Some minor repetition and spelling artifacts in the lower items like 'Octpsuin' and 'Choouip'.

Verdict: Imagen 3.0 Generate 002 is the clear winner as it correctly rendered the full date and maintained much higher spelling accuracy for the specific menu items requested. While FLUX.1 [schnell] has a more traditional 'hand-printed' chalk look, its inability to spell basic words like 'April' or 'Truffle' makes it less useful for this specific prompt.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 [schnell]
Imagen 3.0 Generate 002

AI Judge Analysis

FLUX.1 [schnell]

  • + Perfectly follows the specific instruction to have the horse on top of the astronaut.
  • + High cinematic quality with dramatic lighting and a detailed planetary backdrop.
  • + Captures the 'surreal' aspect of the prompt effectively.
  • The horse has severe anatomical glitches, including two heads and a missing back leg.
  • The astronaut's backpack/head structure is a bit confusingly rendered.

Imagen 3.0 Generate 002

  • + Superior anatomical rendering of the horse and the astronaut.
  • + Beautifully detailed space background with nebulae and stars.
  • + Excellent clarity and sharp details throughout.
  • Completely failed the negative constraint to put the horse on top.
  • Produces a generic 'astronaut on a horse' image seen many times before.

Verdict: FLUX.1 [schnell] is the clear winner for its superior prompt adherence, successfully interpreting the unusual spatial requirement of having the horse on top of the astronaut. While Imagen 3.0 has much better technical execution and anatomical correctness, it ignored the core instruction of the prompt, resulting in a standard composition rather than the requested surreal scene.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 [schnell]
Imagen 3.0 Generate 002

AI Judge Analysis

FLUX.1 [schnell]

  • + Excellent high-clarity rendering with realistic PBR materials on the salmon.
  • + Clean, minimalist aesthetic that feels very professional.
  • + Includes the flag icon as requested.
  • Completely missed the word 'SUSHI' in the text overlay.
  • The text 'JAPAN' is somewhat low contrast against the light blue background.
  • The sushi piece is a strange hybrid between a roll and nigiri that looks slightly anatomical.

Imagen 3.0 Generate 002

  • + Followed all text instructions perfectly including 'JAPAN', 'SUSHI', and the flag icon.
  • + Beautiful soft-textured 3D cartoon style that matches the 'miniature' and 'diorama' prompts.
  • + Excellent composition with a wider variety of sushi and accessories on the raised base.
  • The chopsticks are merging into each other at the tips.
  • The text placement is slightly off-center to the left rather than top-center.

Verdict: Imagen 3.0 Generate 002 is the winner as it adhered to all text prompts, whereas FLUX.1 [schnell] failed to include the word 'SUSHI'. Additionally, Imagen 3.0 captured the 'miniature 3D cartoon' aesthetic much more effectively with a charming diorama layout, while FLUX.1 [schnell] produced a more clinical/isolated image.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 [schnell]
Imagen 3.0 Generate 002

AI Judge Analysis

FLUX.1 [schnell]

  • + Vibrant color palette and warm lighting
  • + Very expressive facial features and eyes
  • + Excellent texture on the puppy's fur
  • Failed to include a rabbit; instead generated two kitten-like hybrids
  • The fox kit has unusual, slightly distorted facial features

Imagen 3.0 Generate 002

  • + Perfectly included all four requested animals: puppy, kitten, bunny, and fox
  • + Excellent capture of the 'tumbling' and 'playful' aspect of the prompt
  • + Beautifully rendered dew drops and god rays contributing to the atmosphere
  • The fox kit lacks some of the characteristic 'kit' features, looking more like a small fox
  • The butterflies are less integrated into the composition than in Model A

Verdict: Imagen 3.0 Generate 002 is the clear winner as it successfully rendered all four specific animals requested, whereas FLUX.1 [schnell] missed the bunny entirely and replaced it with a second cat-like creature. Imagen 3.0 also better captured the playful movement described in the prompt, with a very high level of photorealism and detail in the environment.

Next steps

Explore each model