Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 [dev] Black Forest Labs LongCat-Image Meituan

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 [dev]

24.6 arena score

#16 of 62 in Text-to-Image

Skill signature · Text-to-Image

LongCat-Image

9.8 arena score

#62 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 [dev]

100.0%

win rate

Ties

0.0%

LongCat-Image

0.0%

win rate

100.0% 0.0% ties 0.0%
Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent photorealistic rendering of light and reflections.
  • + Creative interpretation of the glass cube with thick, beveled edges.
  • + Accurate placement of the red book and blue sphere.
  • The glass cube is rectangular rather than a perfect cube.
  • The plant behind the cube is less clearly visible through the glass due to the thickness of the material.

LongCat-Image

  • + Perfectly adheres to the 'cube' geometry.
  • + Clearly shows the plant partially visible through the glass as requested.
  • + Realistic material textures for the wooden table and book cover.
  • The scale of the sphere is slightly off relative to the size of the cube.
  • Minor distortion in the reflection on the bottom surface of the cube.

Verdict: Both models followed the prompt instructions accurately, including the complex request for lighting and visibility through glass. FLUX.1 [dev] produced a more artistically striking image with superior lighting, while LongCat-Image provided a better literal interpretation of a 'cube' and more clearly depicted the plant behind the glass.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent skin texture and realistic material rendering on the jacket.
  • + Photorealistic bokeh and DOF centered on the person.
  • + Clean, logical bicycle structure.
  • The man is holding the handlebars rather than repairing the bike.
  • Lacks the requested 'imperfect framing'—the composition is very centered and tidy.

LongCat-Image

  • + Better adherence to the 'repairing' action with the kneeling pose.
  • + Captures the 'imperfect framing' and 'candid' street vibe more effectively.
  • + Good use of wet pavement reflections.
  • Severe anatomy and object hallucinations, including a third bicycle wheel.
  • The man's right hand is mangled and lacks realistic finger detail.
  • Rain effect looks like static vertical lines rather than realistic droplets.

Verdict: FLUX.1 [dev] produced a much higher quality image in terms of technical realism and anatomical correctness, though it missed the specific 'repairing' action by having the man simply stand with the bike. LongCat-Image captured the requested candid mood and pose better, but failed significantly on a technical level with distorted hands and an extra wheel appearing on the bicycle. FLUX.1 [dev] is the clear winner for its professional-grade visual coherence.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent skin texture with realistic pores and fine hair details.
  • + Cinematic lighting with effective shallow depth of field and bokeh.
  • + Striking, lifelike eye rendering.
  • Missed the 'braided with small beads' instruction.
  • The character appears more like a clean fashion model than a 'battle-worn' paladin.
  • Lacks the 'ornate engraved' detail on the armor, which looks more hammered than engraved.

LongCat-Image

  • + Perfectly captures all prompt elements including beads in hair and ornate engraving.
  • + Excellent representation of 'battle-worn' with visible scars, grime, and blood.
  • + Highly detailed textures on leather straps and the chainmail/cloth underlayer.
  • The bokeh sparks are a bit harsh and less integrated into the background.
  • Face skin texture is slightly smoother and less realistic than in Model A.

Verdict: LongCat-Image is the clear winner for its superior prompt adherence, successfully including the beads, ornate engravings, and leather straps that FLUX.1 [dev] omitted. While FLUX.1 [dev] has slightly more realistic skin texture, LongCat-Image better captures the 'battle-worn' narrative and the specific stylistic details requested.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent professional layout that truly feels like a printable menu.
  • + Clean white background and minimal aesthetic follow the prompt closely.
  • + Text is well-spaced and rendered in a clear, consistent typeface.
  • Missed the request for a 'grid' of food photos, opting for a decorative placement instead.
  • Missing the specific 'Pizza' section heading, though pizza is pictured.

LongCat-Image

  • + Successfully followed the 'grid' instruction for the food photos.
  • + Included sections for pizza and appetizers as requested.
  • The layout is cluttered and chaotic, lacking the 'modern minimalist' feel.
  • Text rendering is poor with significant artifacts and illegible characters.
  • Color blocks are distracting and reduce the professional quality of the design.

Verdict: FLUX.1 [dev] produced a much more realistic and professional menu that adheres to the minimalist aesthetic, despite missing the grid layout for images. LongCat-Image attempted to follow more of the specific content instructions, but the visual quality is low, the fonts are distorted, and the overall composition is too cluttered for the requested style.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent photorealistic texture on the burger components
  • + Atmospheric use of fire and 'spark' effects for a magical feel
  • + Perfectly centered and clean 'exploded' composition
  • Missed the primary 'MAGIC BURGER' text requirement
  • Missing the requested starburst for the price
  • Small sparkling shapes look more like floating debris than magic

LongCat-Image

  • + Successfully included all requested text and specific elements like the starburst
  • + Text is rendered with high-quality glowing and fiery effects
  • + Excellent adherence to the 'fiery background' and embers prompt
  • The burger is not 'exploded' or separated into mid-air components as requested
  • Composition feels a bit cluttered with the oversized text overlapping the scene

Verdict: LongCat-Image adheres much better to the complex text requirements of the prompt, including the specific 'MAGIC BURGER' title and the starburst. However, FLUX.1 [dev] produced a superior visual representation of the 'exploded burger' concept with much higher realism in textures, despite failing to include the main text.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent text accuracy, rendering all requested menu items and the date correctly.
  • + Very realistic chalk texture with slight smudging and opacity variations in the strokes.
  • + Clear, legible handwriting that matches the 'cozy café' aesthetic perfectly.
  • Minor spelling errors in the bottom footer text ('drish' instead of 'fresh', 'aur' instead of 'our').
  • The framing of the chalkboard is slightly tight at the top.

LongCat-Image

  • + Good background depth and wider composition showcasing the café environment.
  • + Realistic chalk dust accumulation at the bottom of the board.
  • Significant text rendering failures, resulting in garbled or nonsensical words.
  • Failed to follow the specific menu item text requested in the prompt.
  • The handwriting style is messy and inconsistent compared to the prompt's request for elegant cursive/handwriting.

Verdict: FLUX.1 [dev] is the clear winner as it successfully rendered nearly all the complex menu text requested with high legibility and realistic chalk textures. LongCat-Image struggled significantly with the text, producing gibberish and failing to adhere to the specific content of the prompt.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent lighting and cinematic color grading
  • + Anatomically clean horse depiction with high-resolution textures
  • + Strong visual coherence and elegant composition
  • Failed the negative prompt constraint; the astronaut is riding the horse instead of the horse riding the astronaut

LongCat-Image

  • + Richly populated background with interesting sci-fi elements
  • + Sharp details on the astronaut suit and horse mane
  • Failed the negative prompt constraint; horse is on bottom, not top
  • Poor perspective; the horse appears to be walking on a planetary surface while also being scaled to planetary height against the horizon
  • Nonsensical objects in the sky with distorted shapes

Verdict: Both FLUX.1 [dev] and LongCat-Image struggle with the complex spatial logic of the prompt, failing to put the horse 'on top' of the astronaut. However, FLUX.1 [dev] is significantly superior in visual quality, offering a clean, cinematic aesthetic, whereas LongCat-Image contains messy background artifacts and disjointed perspective.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent photorealism in various textures like fur and skin
  • + Logical spatial arrangement with the passenger clearly in the back seat
  • + Subtle, accurate phone screen lighting on the passenger's face
  • The passenger appears to be in the front passenger seat rather than the back seat
  • Includes two steering wheels or a confusing dashboard perspective

LongCat-Image

  • + Correctly places passenger in the back seat as requested
  • + More detailed taxi driver hat and professional uniform attire
  • + Sharp, clear background that captures the New York nighttime aesthetic
  • The capybara's hand/paw anatomy is distorted and unnatural
  • The passenger's head is strangely transparent against the capybara's fur
  • The taxi topper contains gibberish text and distorted shapes

Verdict: FLUX.1 [dev] produces a much more realistic and aesthetically pleasing image, though it fails on the spatial instruction by placing the passenger in the front seat. LongCat-Image follows the positioning instructions better but suffers from significant anatomical glitches and ghosting artifacts where the passenger and driver overlap. FLUX.1 [dev] is the preferred choice for its superior texture quality and lighting.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent graphic design and composition.
  • + Clean, thorned border that matches the spooky aesthetic.
  • + High visual clarity and stylized illustrations.
  • Failed significantly on text rendering, with scrambled title and duplicate time info.
  • Did not properly include the 'parchment' texture requested.
  • Missing the cobweb detail mentioned in the prompt.

LongCat-Image

  • + Excellent text rendering for the main title and sub-banner.
  • + Perfectly captures the 'dark parchment' texture and spiderweb corners.
  • + Strong cinematic lighting with a moody night sky and twisted trees.
  • Minor spelling errors in the bottom event details (e_g_, 'Armiees').
  • The transition between the parchment and the central scene is slightly brusque.

Verdict: LongCat-Image is the clear winner as it successfully follows the complex text-heavy prompt and incorporates all visual elements like parchment and cobwebs. FLUX.1 [dev] produced a more polished graphic illustration, but it failed to render the requested title text correctly and ignored several key descriptors in the prompt.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent 3D miniature aesthetic with realistic depth of field
  • + Sophisticated PBR lighting and soft clay-like textures
  • + Accurate 45-degree isometric composition
  • Failed the text prompt, displaying one incorrect character and misspelling 'SUSHI' as 'SUSH CATON'
  • The flag icon is stylized incorrectly

LongCat-Image

  • + Perfect adherence to text instructions with 'JAPAN' and 'SUSHI' rendered correctly
  • + Clean and vibrant colors that pop against the light blue background
  • + Very crisp rendition of the flag icon
  • The rice texture appears like large beads rather than realistic soft grains
  • Lighting is a bit flat compared to the soft shadows in the other model

Verdict: While FLUX.1 [dev] produced a more artistically pleasing render with superior textures and lighting, it completely failed to follow the text generation instructions. LongCat-Image followed all text prompts perfectly and delivered a clean, professional-looking graphic, making it the more functional choice for the specified requirements.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Uniform lighting and composition.
  • + Captures a very cute, whimsical aesthetic consistent with the 'joyful wholesome' prompt.
  • Failed the prompt's species requirements by showing two dogs and zero cats.
  • The style is 3D-animation/Disney-esque rather than the requested 'hyper-photorealistic'.
  • The animals appear as cartoon characters with stylized human-like eyes.

LongCat-Image

  • + Closer to the requested 'hyper-photorealistic' textures in fur and environment.
  • + Accurately includes the fox, puppy, and tabby kitten.
  • + Excellent lighting effects with distinct god rays and dew sparkles.
  • Created a 'cat-bunny' hybrid with long ears on the kitten rather than two separate animals.
  • The scale between the fox and the puppy is slightly inconsistent.

Verdict: LongCat-Image is the preferred output because it adheres more closely to the 'hyper-photorealistic' style and correctly identifies most of the requested species, whereas FLUX.1 [dev] produced a 3D-stylized image that failed to include the tabby kitten. While LongCat-Image struggle with the kitten's ears, its overall detail in texture and lighting better matches the 8K masterpiece requirement.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 [dev]
LongCat-Image

AI Judge Analysis

FLUX.1 [dev]

  • + Clean vector emblem style with professional spacing
  • + Correct implementation of the cloche dome and steam iconography
  • + High-quality typography rendering for the 'Est. 1720' text
  • Significant spelling errors in the brand name ('Flariláan') and secondary text ('Reseaurant')
  • Addition of random, hallucinated numbers like '11011' and '1941'

LongCat-Image

  • + Accurate spelling of the brand name 'Caffè Florian'
  • + Excellent paper texture on the background as requested
  • + Stronger retro artistic character and line work
  • Redundant text with 'Caffè' appearing twice
  • Slightly cluttered composition with the sunburst-style rays

Verdict: LongCat-Image wins this comparison by successfully spelling the brand name 'Caffè Florian' correctly, whereas FLUX.1 [dev] produced several typos and hallucinated extra numbers. While FLUX.1 [dev] had a cleaner minimalist aesthetic, LongCat-Image better captured the 'subtle texture' and 'vintage' requirements of the prompt.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 [dev]
LongCat-Image
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent adherence to the 'modern vector infographic' aesthetics with clean, crisp lines.
  • + Successfully captures the specific color palette requested (navy, white, muted red, light gray).
  • + Layout captures the logical flow of a mission trajectory across a central arc.
  • Text consists of nonsensical gibberish despite the header being clear.
  • Includes irrelevant planets like Saturn which were not part of the prompt.

LongCat-Image

  • + Features a recognizable lunar module on the surface for the final step.
  • + Correctly incorporates the flags and astronaut silhouettes as supporting details.
  • Fails to follow the 'modern vector' style, appearing more like a rough comic or hand-drawn illustration.
  • Does not provide the sequential 6-step infographic flow requested.
  • Poor text rendering and inconsistent line weights.

Verdict: FLUX.1 [dev] is the clear winner for its superior professional infographic design, which perfectly matches the aesthetic style and color palette described in the prompt. While the text is mostly illegible, the overall composition and vector quality are exactly what was requested, whereas LongCat-Image failed significantly on the stylistic requirements and sequential layout.

Next steps

Explore each model