Head to head
Esc

Models · slot A

to navigate to pick

LongCat-Image Meituan Wan 2.7 Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

LongCat-Image

12.9 arena score

#61 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.7

20.5 arena score

#39 of 62 in Text-to-Image

Vote tally

Where the votes landed

LongCat-Image

0%

win rate

Ties

0%

Wan 2.7

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent realization of glass transparency and realistic reflections.
  • + High visual quality with a modern, clean aesthetic.
  • + Perfect adherence to lighting and object placement.
  • The glass cube has an open top rather than being a solid cube structure.
  • The plant is almost entirely behind the book/cube rather than clearly visible through the glass.

Wan 2.7

  • + Successfully shows the plant through the glass as requested in the prompt.
  • + Highly detailed textures on the wooden table and red book spine.
  • + Accurately depicts a fully enclosed glass cube structure.
  • There is a strange ghosting artifact of a second blue sphere on the left side of the cube.
  • The perspective of the cube's interior base seems slightly misaligned with the exterior.

Verdict: LongCat-Image produces a cleaner, more aesthetically pleasing image with superior lighting and clarity. While Wan 2.7 follows the specific detail of seeing the plant through the glass better, it suffers from a significant visual artifact (a duplicate sphere reflection that doesn't match the physics) and a less polished look.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent depiction of rain and wet pavement reflections.
  • + Captures a strong cinematic evening mood with vibrant colors.
  • + Accurate shallow depth of field as requested.
  • The bicycle structure is physically impossible, appearing to have three wheels or a split frame.
  • The composition feels a bit too staged compared to the 'candid' request.
  • Lack of anatomical detail in the hands.

Wan 2.7

  • + Near-photorealistic textures and more natural lighting.
  • + Better adherence to the 'candid' and 'no stylization' requirements.
  • + Successfully captures the specific requested 'imperfect framing' and 50mm feel.
  • Missed the requested motion blur from passing cars.
  • The bicycle geometry is slightly distorted where the handlebars meet the frame.

Verdict: Wan 2.7 is the winner for its superior realism and adherence to the 'no stylization' and 'candid' aspects of the prompt, looking much more like a real documentarian photograph. LongCat-Image creates a more striking cinematic image but fails significantly on the technical details of the bicycle and the naturalism requested.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent depiction of high-gloss reflective armor
  • + Vibrant colors and dramatic lighting
  • + Accurate representation of the requested beads and braids
  • The facial feature (hole/wound) on the cheek looks somewhat artificial and anatomically odd
  • The character looks a bit too 'clean' for a battle-worn description

Wan 2.7

  • + Superb skin texture with realistic scarring and dirt
  • + Better 'battle-worn' aesthetic with weathered, scratched armor
  • + Highly realistic leather texture and buckle detailing
  • The braiding on top of the head appears slightly messy or less defined
  • Colors are more muted compared to the prompt's 'warm torchlight' focus

Verdict: Wan 2.7 is the superior choice for this prompt as it captures the essence of a 'battle-worn' paladin with much more convincing skin textures and weathered armor. While LongCat-Image provides more vibrant lighting and cleaner engravings, its character looks more like a model in a costume rather than a gritty warrior.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Includes all specific categories requested (pizza, mains, appetizers)
  • + Strong use of vibrant accent colors as requested
  • Text is mostly illegible gibberish
  • The layout is cluttered and lacks professional hierarchy
  • Photos are poorly cropped and some look unappetizing

Wan 2.7

  • + Excellent typography with readable prices and headings
  • + High-quality, appetizing food photography
  • + Perfect grid alignment and professional graphic design
  • Categories like 'Pizza' and 'Mains' are headers, but the grid shows a mix of items including desserts/drinks underneath them

Verdict: LongCat-Image technically attempts more of the prompt's specific menu categories but fails significantly on legibility and aesthetic quality, resulting in a chaotic layout. Wan 2.7 produces a highly professional, realistic, and visually appealing menu with mostly readable text and a clean minimalist aesthetic, making it the superior design choice.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent typography style with high-quality glowing and fiery effects.
  • + Beautiful photorealistic texture on the meat patty and buns.
  • + Clear and coherent advertising layout.
  • Failed the core prompt instruction of an 'exploded' burger with components suspended in mid-air.
  • The burger appears mostly assembled rather than dynamic.

Wan 2.7

  • + Perfect adherence to the 'exploded' and 'suspended in mid-air' instruction.
  • + Highly dynamic composition with sauce splashes and floating vegetables.
  • + Followed all text instructions including specific placement and fiery effects.
  • The 'MAGIC BURGER' text has slight inconsistencies in the flame shapes atop the letters.
  • Some elements, like the pickles and lettuce, have a slightly more illustrative look compared to the photorealistic bun.

Verdict: While LongCat-Image produces a very high-quality static ad with great textures, it fails to follow the primary conceptual prompt of an exploded burger. Wan 2.7 successfully captures the dynamic, mid-air motion requested and manages to integrate all required text elements effectively.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Captures a realistic chalk texture with dusty smudges and varying stroke pressure.
  • + The handwriting looks genuinely human-written rather than a digital font.
  • Extreme spelling errors making the text largely illegible.
  • Failed to include the third specific menu item requested in the prompt.

Wan 2.7

  • + Perfect text rendering with zero spelling errors for all requested items.
  • + Excellent layout and composition that follows all prompt instructions.
  • + Successfully interpreted the truncated 'Brown But...' request as 'Brown Butter Chocolate Chip Cookies'.
  • Text appearance is a bit too uniform, leaning towards a 'chalk font' rather than organic handwriting.
  • The lighting on the board is somewhat flat compared to the background environment.

Verdict: Wan 2.7 clearly wins this comparison by providing perfectly legible, accurate text that follows every instruction in the prompt, including completing the truncated menu item. LongCat-Image achieves a more realistic chalk texture and natural 'handwritten' messiness, but it fails significantly on spelling and fulfilling the list of required items.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent depiction of surface dust and cinematic planetary lighting.
  • + High level of detail in the lunar modules and background celestial bodies.
  • Failed the negative constraint: the astronaut is riding the horse instead of the horse riding the astronaut.
  • Anatomy issues with the horse's front left leg which appears detached or floating.

Wan 2.7

  • + Includes beautiful spiral galaxies and a clean, high-resolution aesthetic.
  • + Good rendering of the astronaut suit texture and harness.
  • Failed the negative constraint: the astronaut is riding the horse instead of the horse riding the astronaut.
  • The composition is a bit more generic compared to the first image.

Verdict: Both models failed the specific instruction to have the 'horse on top' of the astronaut, instead providing the typical 'astronaut on a horse' cliché. LongCat-Image is visually more interesting due to the surface details and cinematic lighting, whereas Wan 2.7 provides a cleaner but more standard space render.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent fur texture and lighting on the capybara.
  • + Includes two passengers which adds to the busy New York vibe.
  • + The lighting on the taxi rooftop sign and city bokeh is very realistic.
  • The passenger is sitting directly behind the capybara but the perspective feels slightly claustrophobic.
  • The capybara only has one paw clearly on the steering wheel while the other is tucked.

Wan 2.7

  • + Perfectly follows the instruction of having both front paws on the steering wheel.
  • + The passenger is clearly in the back seat with a realistic 'bored' expression as requested.
  • + The perspective captures the interior and exterior context very effectively.
  • The capybara's fur looks slightly more illustrated or 'clean' compared to the photorealism of Model A.
  • Minor anatomical oddity where the capybara's arm meets the jacket sleeve.

Verdict: Wan 2.7 is the winner as it accurately followed more specific prompt instructions, such as placing both paws on the wheel and capturing the passenger's bored expression perfectly. While LongCat-Image has slightly more realistic fur rendering, it failed to place both paws on the wheel and the passenger's placement feels less balanced in the composition.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent high-contrast cinematic lighting on the jack-o-lantern
  • + Creative integration of thorns as a physical border
  • + Large, legible gothic title text
  • Several typos in the event details section, including 'The Armiees' instead of 'The Arches'
  • The layout is a bit cluttered with overlapping elements

Wan 2.7

  • + Perfect text accuracy for all components including date, time, and location
  • + Highly symmetrical and professional layout with clean borders
  • + Illustrative vintage style matches the 'parchment' aesthetic better
  • The lighting is flatter compared to the cinematic shadows in the first model
  • Added 'Est. 1847' and 'Dress code' which were not in the prompt

Verdict: Wan 2.7 is the clear winner because it correctly renders all requested text, including the specific location 'The Arches', whereas LongCat-Image suffers from several misspellings and garbled text in the bottom section. While LongCat-Image has more atmospheric cinematic lighting, Wan 2.7 provides a more balanced and functional invitation design.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent 3D miniature toy-like texture and material quality
  • + Text rendering is clean and correctly positioned at the top-center
  • + Strong adherence to the 'gentle lighting' and 'PBR materials' keywords
  • The wooden base has some perspective inconsistencies on the left side
  • Lower variety of sushi types compared to model B

Wan 2.7

  • + Perfect 45° isometric perspective and diorama composition
  • + Includes a wider variety of sushi and accessories like soy sauce and chopsticks
  • + Very clean typography with a professional graphic design feel
  • The text is slightly off-center to the left
  • The diorama base is a bit large compared to the requested 'small' base

Verdict: Both models followed the prompt exceptionally well, particularly regarding text rendering and isometric style. LongCat-Image yields a more charming 'toy' aesthetic with superior material textures, while Wan 2.7 provides a more comprehensive sushi spread and a cleaner diorama platform. Wan 2.7 is the likely winner for its superior adherence to the isometric perspective and composition.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Excellent use of god rays and vibrant golden hour lighting.
  • + Very soft and detailed fur textures on the puppy.
  • + Captures the 'big expressive eyes' and joyful vibe perfectly.
  • Serious anatomical error where the kitten and bunny are merged into one creature with cat features and rabbit ears.
  • The butterflies appear to be floating 2D assets rather than part of the scene.

Wan 2.7

  • + Successfully includes all four distinct animals requested in the prompt.
  • + Superior interaction and movement, with the animals actually 'tumbling' and chasing.
  • + Beautiful environment with realistic dew sparkles and a wider variety of wildflowers.
  • The fox kit has eyes that look slightly less 'expressive' or more mature than a kit.
  • The kitten's mouth area has slight artifacts.

Verdict: LongCat-Image suffers from a major anatomical failure by merging the cat and bunny into a single hybrid animal, whereas Wan 2.7 correctly identifies and renders all four specific animals requested. Wan 2.7 also captures the sense of playfulness and the environment more effectively, despite LongCat-Image having slightly higher contrast in its lighting.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Strong hand-drawn vintage aesthetic
  • + Correct spelling of 'Caffè Florian'
  • + Excellent capture of the requested steam and texture elements
  • Text is repetitive with 'Caffè' appearing twice
  • Design is a bit cluttered for a minimalist logo

Wan 2.7

  • + Clean vector emblem style with balanced composition
  • + Elegant typography on the 'Est. 1720' banner
  • + Fits the minimalist and professional logo brief well
  • Spelling error in the main brand name ('Florion' instead of 'Florian')
  • Steam trailing off the cloche is slightly disconnected

Verdict: LongCat-Image captures the vintage atmosphere and texture much better and manages to spell the name correctly, although the layout is somewhat busy. Wan 2.7 produces a cleaner, more professional vector layout, but fails on the spelling of the core subject name. LongCat-Image is the preferred choice for accuracy to the specific brand prompt.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

LongCat-Image
Wan 2.7

AI Judge Analysis

LongCat-Image

  • + Successfully captures the requested color palette
  • + Solid vector illustration of the lunar module
  • Fails to include all 6 requested steps
  • Text is largely gibberish/unreadable
  • Layout is cluttered and lacks a clear logical flow

Wan 2.7

  • + Strictly follows the 6-step sequence outlined in the prompt
  • + Excellent text legibility and professional infographic layout
  • + Iconography is consistent and fits the flat-vector style perfectly
  • One spelling error ('Descript' instead of 'Descent')
  • Central vertical dashed line is slightly off-center between Earth and Translunar steps

Verdict: Wan 2.7 significantly outperforms LongCat-Image by adhering to the specific 6-step instructional sequence and providing high-quality, legible text. While LongCat-Image captures the aesthetic and colors well, it fails the basic structure of the infographic challenge, whereas Wan 2.7 delivers a functional and professional-looking poster.

Next steps

Explore each model