Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [dev] Black Forest Labs LongCat-Image Meituan

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.2 [dev]

24.5 arena score

#18 of 62 in Text-to-Image

Skill signature · Text-to-Image

LongCat-Image

12.9 arena score

#61 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [dev]

0%

win rate

Ties

0%

LongCat-Image

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent adherence to lighting instructions from the left
  • + High realism in textures, especially the wood grain and glass reflections
  • + Natural plant placement visible through the glass volume
  • The sphere is slightly smaller than what might be expected for 'small' vs 'tiny' in comparison to the cube size
  • The top glass edge under the book looks slightly merged

LongCat-Image

  • + Strong geometric clarity in the glass cube's construction
  • + Very vibrant blue sphere with good internal light refraction
  • + Clean and simple composition
  • The book appears to be floating slightly above the glass rather than resting on it
  • The lighting is more frontal than from the left window
  • The plant's occlusion through the glass is less realistic, appearing more like a backdrop than behind the object

Verdict: Both models followed the prompt's spatial instructions perfectly. FLUX.2 [dev] is the winner due to its superior handling of physical realism, particularly how the book weight rests on the glass and how the lighting consistently hits the scene from the window. LongCat-Image is visually appealing but has minor physics issues, such as the book appearing to hover slightly.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent photo-realism with natural skin textures and greasy hands.
  • + Perfect execution of motion blur from the passing car while keeping the subject sharp.
  • + Strong adherence to the 'imperfect framing' and 'candid' aspects of the prompt.
  • The anatomy of the elderly man's left hand is slightly mangled around the fingers.
  • The bicycle handlebars are overly complex and physically confusing.

LongCat-Image

  • + Great composition and use of reflections on the wet pavement.
  • + Accurate street atmosphere with recognizable background storefronts.
  • + Matches the 'shallow depth of field' request well.
  • Failed to include 'motion blur' on the passing cars, which appear sharp.
  • The red bicycle is nonsensical with extra wheels and floating parts.
  • The man's feet and the ground relationship appear slightly 'photoshopped' or floaty.

Verdict: FLUX.2 [dev] followed the prompt much more accurately, specifically succeeding where LongCat-Image failed by correctly adding motion blur to the background traffic. While both models struggled with the complex geometry of the bicycle, FLUX.2 [dev] provided a more convincing 'candid' and realistic texture for the character and the environment.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent depiction of the 'close portrait' instruction with expressive facial details.
  • + Highly intricate engraving on the armor and realistic leather grain.
  • + Superior handling of lighting, with a plausible torch source and naturalistic reflections.
  • The scars appear slightly more like recent wounds than faint, battle-worn scars.

LongCat-Image

  • + Beautifully braided hair with clearly defined beads and textures.
  • + Strong overall composition with atmospheric sparks and secondary lighting.
  • + Armor design is very ornate and fits the 'paladin' aesthetic well.
  • The framing is more of a medium shot than a 'close portrait'.
  • Metal textures look a bit smoother and more 'CG' compared to the grittier feel of the competitor.

Verdict: FLUX.2 [dev] followed the 'close portrait' framing more accurately, providing a much higher level of detail in the skin pores, leather straps, and metal engravings. LongCat-Image produced a visually striking image with great hair detail, but the lighting and textures felt less grounded and lifelike than those in FLUX.2.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Features a very clean and professional multi-column layout.
  • + Includes clear headings for Appetizers, Pizza, and Mains as requested.
  • + High-quality, realistic food photography and consistent design accents.
  • Text is mostly gibberish with several horizontal lines intersecting the text boxes.
  • Headers have some spelling errors like 'PIZZAU'.

LongCat-Image

  • + Vibrant use of yellow and blue accents fulfills the 'colorful' part of the prompt.
  • + Contains several food photos arranged in a grid-like style.
  • Text rendering is extremely poor with distorted and illegible characters.
  • Layout feels cluttered and less like a professional 'minimalist' menu.
  • The grid is inconsistent and chaotic compared to the organized prompt requirement.

Verdict: FLUX.2 [dev] is the clear winner as it successfully produces a professional, minimalist menu layout that adheres to all specific section requests. While its text is not perfectly legible, it maintains a much higher level of visual coherence and aesthetic quality than LongCat-Image, which suffers from severe text distortion and a messy composition.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent adherence to the 'exploded' and 'suspended' layout requested.
  • + Uniform and coherent fiery, glowing text effects across all elements.
  • + Superior photorealistic quality in the texture of the meat and lettuce.
  • The starburst shape is somewhat irregular and rough.

LongCat-Image

  • + Strong presence of burning coals and embers in the foreground.
  • + Clear and legible typography with high contrast.
  • + Crisp lighting on the burger surface.
  • Failed to follow the 'exploded' instruction, providing a mostly assembled burger instead.
  • The starburst is a flat vector-style graphic that clashes with the 3D scene.
  • The sauce texture looks somewhat plastic and artificial.

Verdict: FLUX.2 [dev] is the clear winner as it precisely followed the complex request for an 'exploded' view of the burger, whereas LongCat-Image provided a standard stacked burger. FLUX.2 [dev] also achieved a more professional advertisement look with consistent fiery glow effects across the text and a more realistic food presentation.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent text legibility with perfect spelling for all menu items.
  • + Realistic chalk texture and smudging on the chalkboard surface.
  • + Matches the requested elegant cursive and handwritten variations seamlessly.
  • The word 'with' in the second item has some faint ghosting/overlap artifact.

LongCat-Image

  • + Natural-looking cafe background with depth of field.
  • + Good chalk dust effects at the bottom of the board.
  • Poor spelling and garbled text across the entire board.
  • Text size and layout are inconsistent with the prompt's logical structure.
  • Failed to render the full names of the dishes correctly.

Verdict: FLUX.2 [dev] significantly outperforms LongCat-Image by following the complex text requirements perfectly, maintaining correct spelling, and capturing the tactile feel of a chalkboard. LongCat-Image struggled with the text generation, resulting in numerous typographical errors and incoherent words like 'Toays Stays'.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent cinematic lighting and texture on the horse's coat.
  • + Strong anatomical consistency for both the horse and the astronaut's gear.
  • + Superior background rendering with realistic nebulas and planetary curvature.
  • Failed to follow the specific spatial instruction 'horse on top, not vice versa'.

LongCat-Image

  • + Includes interesting background elements like lunar modules and spacecraft.
  • + Good color contrast and sharp focus.
  • Failed to follow the specific spatial instruction 'horse on top, not vice versa'.
  • Anatomical issues with the horse's lower legs and hoof placement.
  • The composition feels cluttered with disparate elements like planes and multiple moons.

Verdict: Both FLUX.2 [dev] and LongCat-Image failed the negative constraint/spatial instruction to place the horse on top of the astronaut, instead defaulting to the common trope of an astronaut riding a horse. FLUX.2 [dev] is the superior image due to its significantly higher visual quality, realistic lighting, and better anatomical rendering compared to the cluttered and anatomically incorrect output from LongCat-Image.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent fur texture and capybara facial expression following the 'calm, professional' description.
  • + Superior background realism with authentic New York bokeh through the windows.
  • + Strong composition that clearly shows both the driver and the passenger as requested.
  • The capybara's hands look slightly more human/primate-like than actual capybara paws.
  • The interior framing feels a bit tight, cutting off the top of the taxi frame.

LongCat-Image

  • + Includes more external taxi details like the roof light and door markings.
  • + Good rendering of the capybara's profile and clothing.
  • + Crisp lighting on the steering wheel and dashboard.
  • Contains a significant anatomical error with a second 'ghost' passenger appearing behind the main passenger.
  • The passenger's hands interacting with the phone are distorted with too many fingers.
  • The placement of the capybara's arm relative to its body looks anatomically disconnected.

Verdict: FLUX.2 [dev] is the clear winner as it successfully follows all prompt instructions without the significant structural errors found in the competitor. While LongCat-Image attempted a more complex wide shot, it suffered from major AI artifacts, including a duplicated passenger and distorted limbs. FLUX.2 [dev] produced a more coherent, photorealistic image with a much more convincing integration of the capybara into the scene.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent typography with 100% accuracy on provided text.
  • + Perfect composition with a dark, atmospheric vintage aesthetic.
  • + High-quality rendering of the thorny border and moody lighting.
  • The parchment effect is slightly less 'paper-like' in the center compared to the edges.

LongCat-Image

  • + Strong parchment texture and creative border design.
  • + Good implementation of the requested banner and jack-o-lantern.
  • Several spelling errors in the event details including 'The Armiees' and a random '3ulie'.
  • The composition feels a bit cluttered with the way the background image is 'punched through' the parchment.
  • Text rendering on the banner is slightly distorted.

Verdict: FLUX.2 [dev] significantly outperforms LongCat-Image by delivering a professional-grade invitation with perfect spelling and a cohesive, moody aesthetic. LongCat-Image struggles with text accuracy and has a more cluttered, less polished composition.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent adherence to the square isometric diorama layout.
  • + Clean, professional typography and accurate flag representation.
  • + Highly realistic textures on the sushi and rice.
  • The lighting is a bit flat compared to the other model.

LongCat-Image

  • + Vibrant colors and high-quality subsurface scattering effects on the fish.
  • + Dynamic composition with a nice 3D finish on the text and flag.
  • + Detailed clay-like 'cartoon' aesthetic that fits the miniature prompt well.
  • The flag is slightly distorted and has an unnecessary flagpole.
  • The rice grains look like uniform beads rather than varied sushi rice.

Verdict: FLUX.2 [dev] followed the prompt more precisely, particularly with the 'raised diorama base' and centered layout, resulting in a cleaner overall graphic. LongCat-Image provided better stylized textures and more vibrant lighting, but it struggled slightly with the flag icon and the specific isometric base requested. FLUX.2 [dev] is the winner for its superior composition and perfect text rendering.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Exhibits superior fur texture and realistic anatomy for all animals.
  • + Excellent lighting with natural god rays and dew sparkles that feel integrated into the scene.
  • + Included all four requested animals (dog, cat, rabbit, fox) plus an extra rabbit.
  • The animals are sitting relatively still rather than 'playfully chasing' or 'tumbling' as requested.

LongCat-Image

  • + Successfully captured a more active, playful motion with the golden retriever.
  • + Bright, vibrant colors that fit the 'joyful vibe' well.
  • Significant anatomical errors, most notably a kitten with rabbit ears.
  • Failed to include the baby bunny as a distinct animal in the scene.
  • The dew sparkles look like floating glass orbs rather than water droplets.

Verdict: FLUX.2 [dev] is the clear winner as it successfully rendered all four requested animals with high photorealism and beautiful lighting. LongCat-Image failed the prompt by omitting the bunny and creating a 'cat-rabbit' hybrid, while the digital artifacts for dew and lighting felt artificial.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent typography and spelling of the restaurant name.
  • + Clean, vector-style emblem layout that feels professional.
  • + Perfectly captures the minimalist vintage aesthetic requested.
  • The steam is very small and lacks a bit of stylistic impact.

LongCat-Image

  • + Dynamic steam illustration that fills the space well.
  • + Good texture on the background that fits the vintage theme.
  • Repetitive text with 'Caffè' appearing twice inside the cloche.
  • Messy typography with artifacts and inconsistent letter weighting.
  • Does not feel like a 'minimalist' logo, appearing cluttered and busy.

Verdict: FLUX.2 [dev] followed all instructions perfectly, producing a professional-looking vector logo with accurate spelling and a clean layout. LongCat-Image struggled with the text, repeating words and creating a cluttered composition that failed the minimalist requirement.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [dev]
LongCat-Image

AI Judge Analysis

FLUX.2 [dev]

  • + Excellent adherence to the vector icon style requested.
  • + Follows the specific mission steps (Launch, Orbit, etc.) with recognizable imagery.
  • + Captures the clean, modern aesthetic with sharp typography.
  • Includes several extra, unrequested icons that disrupt the sequential flow.
  • Some text labels are garbled or nonsensical (e.g., 'VALSIST ORDIT').

LongCat-Image

  • + Adheres well to the requested NASA-inspired color palette.
  • + Includes creative iconography variations like flags and map pins.
  • Failed to provide the list of 6 distinct mission steps requested.
  • The vector style is less 'clean' with irregular line weights and poor text legibility.
  • The layout is cluttered and does not function well as an infographic.

Verdict: FLUX.2 [dev] is the clear winner as it successfully interprets the 'infographic' request by providing a clear layout of sequential steps with high-quality vector illustrations. While it has some issues with text accuracy, LongCat-Image fails the core prompt by missing most of the specific mission steps and producing much lower quality, inconsistent vector art.

Next steps

Explore each model