Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [pro] Black Forest Labs LongCat-Image Meituan

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [pro]

20.3 arena score

#41 of 62 in Text-to-Image

Skill signature · Text-to-Image

LongCat-Image

9.8 arena score

#62 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [pro]

100.0%

win rate

Ties

0.0%

LongCat-Image

0.0%

win rate

100.0% 0.0% ties 0.0%
Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent structural geometry of the cube
  • + Realistic soft lighting and shadows on the table surface
  • + The blue sphere has a unique felt-like texture that provides tactile interest
  • The plant is barely visible through the cube, lacking the refractive presence requested
  • The blue sphere appears to be floating rather than resting on the bottom

LongCat-Image

  • + Superior handling of refraction and transparency through the thick glass
  • + High-quality realistic materials for the sphere and glass cube
  • + Accurate placement of the plant behind and through the glass elements
  • The perspective of the cube's top edge is slightly distorted under the book

Verdict: Both models followed the prompt instructions perfectly, including all requested objects and lighting. LongCat-Image is the winner because it handled the optical properties of the glass much more realistically, showing refraction of the plant and sphere, whereas FLUX.1 Kontext [pro] produced a cube that looks more like a hollow wireframe box with no refractive depth.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent natural skin texture and facial realism
  • + Realistic depiction of rain and wet clothing
  • + Effective shallow depth of field with 50mm characteristics
  • The subject appears to be riding or holding the bike rather than actively repairing it
  • Missing the requested motion blur from passing cars

LongCat-Image

  • + Stronger adherence to the 'repairing' action with a squatting pose
  • + Excellent reflections on the wet pavement
  • + Good inclusion of background street elements and cars
  • Anatomical failure with the bicycle geometry and additional floating wheels
  • The lighting on the subject feels slightly artificial compared to the environment
  • Lacks the requested motion blur for the cars

Verdict: FLUX.1 Kontext [pro] creates a much more convincing and photorealistic image with superior skin textures and lighting, though it fails to clearly show the 'repairing' aspect of the prompt. LongCat-Image better captures the requested action (repairing) and the street atmosphere, but it suffers from significant structural hallucinations in the bicycle's design. FLUX.1 Kontext [pro] is the preferred choice for its high visual fidelity and lack of AI artifacts.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [pro]
LongCat-Image
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent skin texture with realistic fine pores and faint scarring
  • + Subtle and sophisticated torchlight reflections on the plate armor
  • + Coherent hair braiding and overall mood
  • Missed the request for beads in the hair braids
  • The framing is slightly too close, obscuring some of the ornate armor details

LongCat-Image

  • + Successfully included small colored beads in the braids
  • + Highly detailed engraving on the plate armor and visible chainmail underlayer
  • + Dynamic lighting with clear fire and bokeh sparks in the background
  • The facial scars look a bit like graphic decals rather than physical texture
  • The lighting on the face is a bit flat compared to the dramatic environment

Verdict: FLUX.1 Kontext [pro] creates a more lifelike and cinematic portrait with superior skin and metal textures, though it missed the specific detail of beads in the hair. LongCat-Image adhered better to all prompt instructions, including the beads and complex underlayers, but at the cost of some photorealistic facial quality. FLUX.1 Kontext [pro] is the winner for its more natural and professional aesthetic.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with legible bold sans-serif headers
  • + Clean and professional layout that follows the requested sections
  • + High-quality food photography that feels integrated into the design
  • Nonsense filler text for menu descriptions
  • Inconsistent pricing values

LongCat-Image

  • + Includes a grid of many colorful food photos as requested
  • + Uses vibrant pink and orange accents
  • Garbled and illegible text throughout the design
  • Chaotic layout that lacks the requested minimalism
  • Low visual fidelity and artifacting in the graphics

Verdict: FLUX.1 Kontext [pro] successfully creates a professional-grade menu with high-quality headers and a clean, minimalist aesthetic that directly follows the prompt. LongCat-Image fails to produce legible text and the design is cluttered and visually unappealing. FLUX.1 Kontext [pro] is the clear winner for its superior composition and font rendering.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography style with high-quality glowing effect.
  • + Highly photorealistic textures on the burger ingredients.
  • + Dynamic sense of motion with embers and floating debris.
  • Repeated the price text twice in the layout.
  • Failed to include the requested starburst element for the price.
  • Burger components are less 'exploded' than Model B, appearing mostly stacked.

LongCat-Image

  • + Included the starburst element as requested in the prompt.
  • + Better adherence to the 'exploded' burger request with visible spacing between all layers.
  • + Accurate text integration without unnecessary repetition.
  • The burger ingredients look slightly more synthetic and less realistic than Model A.
  • The 'MAGIC BURGER' text has slight irregularities in the 'U' and 'R' letter shapes.
  • The glowing effects look a bit like digital overlays rather than lighting integrated into the scene.

Verdict: FLUX.1 Kontext [pro] produces a more photorealistic image with superior lighting and atmospheric effects, but it fails on specific prompt instructions like the starburst and includes redundant text. LongCat-Image follows the prompt layout more accurately, including the starburst and a better 'exploded' view of the burger, even if the overall image quality is slightly lower. LongCat-Image is the winner for better prompt adherence regarding composition and specific design elements.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent text rendering with nearly perfect spelling of the complex menu items.
  • + Realistic chalk texture and natural variations in the wood frame.
  • + High prompt adherence, correctly including the specific date and price points.
  • Some minor gibberish in the supplementary 'fresh daily' text at the bottom.
  • The cursive title style is less 'elegant' and more functional than requested.

LongCat-Image

  • + Good aesthetic background composition showing a cozy café interior.
  • + Captures the 'chalk dust' messy feel on the board and edges well.
  • Very poor text legibility with numerous spelling errors and garbled words.
  • Failed to follow the specific menu item text requested in the prompt.
  • The 'handwriting' looks more like a digital font effect than authentic chalk.

Verdict: FLUX.1 Kontext (pro) is the clear winner for its superior ability to render specific text accurately and follow the detailed menu instructions. While LongCat-Image creates a nice atmosphere, its failure to spell the menu items correctly and its reliance on distorted lettering makes it unusable for a text-heavy prompt.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Perfectly follows the specific instruction of 'horse on top'
  • + High cinematic quality with realistic space lighting
  • + Creative interpretation where the astronaut's feet are horse hooves
  • The small astronaut figure on top of the horse is somewhat redundant
  • Minor anatomical blending issues between the horse chest and astronaut's back

LongCat-Image

  • + Crisp details on the spacesuit and horse tack
  • + Vibrant colors and a wide, busy composition
  • Fails the core prompt instruction of 'horse on top'
  • Physical presence of dust on a surface conflicts with the 'in space' setting
  • Noticeable anatomical errors on the horse's legs

Verdict: FLUX.1 Kontext [pro] successfully followed the difficult spatial constraints of the prompt, creating a surreal and cinematic image where a horse is literally riding an astronaut. LongCat-Image completely ignored the negative constraint ('horse on top, not vice versa'), producing a standard astronaut-on-horse image with significant anatomical errors in the horse's legs.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent photorealistic texture on the capybara's fur.
  • + Accurate lighting that matches a nighttime city environment.
  • + Strong composition focusing on the capybara's professional expression.
  • The passenger is holding a phone to her ear like a call instead of looking at it as requested.
  • The capybara's cap looks a bit small and flimsy.

LongCat-Image

  • + Follows the prompt about the passenger looking at her phone perfectly.
  • + Included two passengers which adds to the busy NYC feel.
  • + The capybara's outfit, including a shirt and jacket, is more detailed.
  • The capybara's paws look more like human-monkey hybrid hands, which is uncanny.
  • The text on the taxi and the roof sign is nonsensical.
  • Noticeable duplicate faces for the two passengers in the back.

Verdict: FLUX.1 Kontext [pro] creates a much more believable and high-quality image with superior fur textures and realistic lighting. While LongCat-Image followed the specific interaction with the phone better, it suffered from significant anatomical issues with the paws and repetitive facial features on the passengers.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with a clean, legible gothic font.
  • + High-quality atmospheric lighting and composition that feels polished.
  • + Accurate text rendering for the primary title and location.
  • Includes a line of gibberish text in the event details section.
  • The background is more of a black void than dark parchment.

LongCat-Image

  • + Stronger adherence to the 'dark parchment' and 'thorn border' descriptors.
  • + Good use of color contrast between the orange pumpkin and blue night sky.
  • + Accurately represents all visual elements like webs and twisted trees.
  • Significant spelling errors and hallucinations in the event details text.
  • The composition feels slightly cluttered and less 'cinematic' than the competitor.
  • The thorn border looks somewhat artificial against the parchment edge.

Verdict: FLUX.1 Kontext [pro] creates a much more professional and aesthetically pleasing invitation with superior typography, though it failed to incorporate the parchment texture requested. LongCat-Image captured the requested individual elements better—such as the thorn border and parchment—but failed significantly on text accuracy and overall layout polish. FLUX.1 Kontext [pro] is the winner for its usability and high visual quality.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with a cohesive 3D 'bubble' style
  • + Beautifully rendered PBR textures on the salmon and rice
  • + Superior composition with a clean, centered layout
  • The flag icon is stylized as a square/box rather than a realistic flag circle

LongCat-Image

  • + Accurate rendering of a realistic Japanese flag icon
  • + Good use of multiple sushi pieces and garnish to fill the scene
  • Text placement is slightly cramped and less integrated than Model A
  • Texture on the red tuna piece looks slightly plastic and less organic

Verdict: FLUX.1 Kontext [pro] produced a more polished and professional-looking graphic with exceptional 3D text and soft, appealing lighting. While LongCat-Image handled the flag icon more accurately, its overall composition and material quality were slightly less refined than the competition.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [pro]
LongCat-Image
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent soft lighting and atmosphere with convincing god rays.
  • + Very natural fur textures and anatomy for all animals.
  • + Beautifully integrated background that feels lush and dreamy.
  • The fox looks a bit more like a kitten/fox hybrid.
  • Limited variety in the flower types compared to the prompt's request for a wildflower meadow.

LongCat-Image

  • + Distinct animal types that are easily identifiable.
  • + Strong literal interpretation of 'god rays' and 'dew sparkles'.
  • + Bright, vibrant colors that pop.
  • One animal is a bizarre cat-rabbit hybrid with two sets of ears.
  • The lighting feels more artificial and 'Photoshopped' rather than hyper-photorealistic.
  • Anatomy issues like the fox having five legs/paws visible beneath its body.

Verdict: FLUX.1 Kontext [pro] creates a much more cohesive and professional image with superior lighting and realistic fur textures. While LongCat-Image attempted to include more specific details like dew sparkles, it suffered from significant anatomical errors, most notably a cat with bunny ears and extra limbs on the fox.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with perfect spelling and accent marks.
  • + Clean, minimalist composition that looks like a professional logo.
  • + Accurate adherence to the color palette and vector emblem style.
  • The 'steam' element is a bit thick and looks slightly more like a flame.

LongCat-Image

  • + Dynamic steam illustration and a more ornate vintage feel.
  • + Good use of the 'Est. 1720' banner with classic styling.
  • Poor text rendering with repetitive and overlapping words.
  • The 'minimalist' aspect of the prompt was ignored in favor of a cluttered design.
  • Visible artifacts and messy lines within the cloche dome.

Verdict: FLUX.1 Kontext [pro] successfully created a clean, professional-grade logo that follows every aspect of the prompt, including correct typography and a minimalist aesthetic. LongCat-Image struggled with text legibility, producing repetitive words and a cluttered composition that deviated from the requested minimalist style. FLUX.1 Kontext [pro] is the clear winner for its clarity and brand-ready quality.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [pro]
LongCat-Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography rendering for the main title and most labels.
  • + Clean, professional flat-vector aesthetic that matches the requested NASA palette.
  • + Includes a complex but legible trajectory map connecting the elements.
  • Confuses the Saturn V rocket with Saturn the planet by adding rings to the nose cone.
  • The numerical steps are logically disorganized and scientific labels like 'Earth' are misidentified.
  • Includes nonsensical text in the bottom right corner.

LongCat-Image

  • + Strong iconography for the lunar module and Earth icons.
  • + Good layout balance between the list view and the featured landing illustration.
  • + Color palette is very consistent with the 'muted red and navy' instructions.
  • Significant spelling errors throughout all text elements including the main title.
  • Fails to follow the specific 6-step sequence outlined in the prompt.
  • Iconography for the rocket looks like a shuttle rather than a Saturn V.

Verdict: FLUX.1 Kontext [pro] creates a much more visually appealing and professional-looking infographic with crisp lines and legible text, though it bizarrely literalizes the name Saturn by giving the rocket rings. LongCat-Image struggles significantly with text rendering and fails to follow the step-by-step instructions of the prompt, resulting in a generic space poster rather than a specific Apollo 11 mission guide.

Next steps

Explore each model