Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 [schnell] Black Forest Labs Wan 2.7 Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 [schnell]

18.7 arena score

#48 of 62 in Text-to-Image

Skill signature · Text-to-Image

Wan 2.7

20.5 arena score

#39 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 [schnell]

0%

win rate

Ties

0%

Wan 2.7

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Clean, modern aesthetic with high visual clarity
  • + Realistic glass reflections and soft lighting
  • + Excellent central focus and vibrant colors
  • Added an extra blue sphere on top of the book that was not in the prompt
  • The sphere looks like it is floating rather than sitting in the cube

Wan 2.7

  • + Perfect adherence to the object count and placement
  • + Highly realistic textures on the wooden table and red book
  • + Accurate depiction of a plant's position behind the glass
  • The glass cube has some internal edge artifacts and structural inconsistencies

Verdict: Wan 2.7 is the clear winner as it followed every instruction in the prompt perfectly, whereas FLUX.1 [schnell] hallucinated an additional sphere on top of the book. Wan 2.7 also exhibited superior texture realism on the wood and book, despite some minor geometric flaws in the glass rendering.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Excellent shallow depth of field with high-quality bokeh.
  • + Strong color contrast with a vibrant red bicycle.
  • + Good adherence to the 'cinematic' aspect of the prompt.
  • Lack of motion blur on the passing cars as requested.
  • The man's skin and hair look slightly too smooth/digital for a 'no stylization' request.
  • The bicycle geometry is slightly illogical near the handlebars.

Wan 2.7

  • + Perfectly captures the 'imperfect framing' and 'candid' street photography aesthetic.
  • + Realistic skin texture and clothing details, appearing much less processed.
  • + Includes visible raindrops and more realistic pavement reflections.
  • Failed to include motion blur on the background vehicles.
  • The depth of field is deeper than the 'shallow' request implies.
  • The man's hands are slightly distorted around the bicycle brake lever.

Verdict: Wan 2.7 is the winner because it captures the authentic 'candid street photo' aesthetic much more effectively than FLUX.1 [schnell], which looks more like a polished film still. Wan 2.7 better handles the requested natural skin texture and imperfect framing, making it feel like a real amateur photograph despite missing the motion blur.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Intense, high-contrast lighting that emphasizes skin texture
  • + Strong composition for a 'close portrait' specifically targeting facial details
  • + Clean, sharp rendering of eyes and facial hair
  • Missed several key prompt elements like 'hair braided with small beads' and 'ornate engraved plate armor'
  • Armor is barely visible and lacks the requested engraving detail

Wan 2.7

  • + Excellent adherence to all prompt details including braided hair with beads, scars, and ornate armor
  • + Superior material rendering on the weathered plate armor and leather straps
  • + Better environmental storytelling with visible torchlight and bokeh sparks
  • Slightly wider shot than a 'close portrait' usually implies
  • Facial skin texture is slightly less defined compared to the metallic textures

Verdict: Wan 2.7 followed every aspect of the prompt, including specific details like the beads in the braids and the intricate engravings on the armor, which FLUX.1 [schnell] largely ignored. While FLUX.1 [schnell] provided a more intimate headshot with striking eyes, Wan 2.7 captured the 'battle-worn paladin' aesthetic much more effectively with its superior material rendering and source preservation.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Strong minimalist aesthetic with plenty of white space.
  • + Features clear, bold sans-serif headers.
  • + Clean four-quadrant photo grid layout.
  • The placeholder text is gibberish and very small/distorted.
  • The photo selection is repetitive with multiple similar-looking pizzas.
  • The section naming 'ORFEFUS' is nonsensical.

Wan 2.7

  • + High level of detail with readable menu item names and prices.
  • + Excellent adherence to all prompt elements including vibrant accents and category tabs.
  • + Sophisticated composition with food-related props in the background.
  • Text in smaller descriptions becomes slightly blurry/garbled.
  • Slightly less 'minimalist' than Model A due to the high density of information.

Verdict: Wan 2.7 produced a significantly more professional and realistic menu that could actually be used, featuring readable prices, menu names, and a cohesive vibrant theme. FLUX.1 [schnell] captures the minimalist layout well but fails on content quality, using nonsensical text and repetitive imagery.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + The burger patty and melting cheese have a high level of photorealistic texture.
  • + The lighting on the bottom of the burger bun is realistic relative to the fire background.
  • + The exploding particles give a good sense of debris and motion.
  • The text is poorly rendered with a typo ('AGIC BURGER') and redundant price tags.
  • The burger is largely assembled rather than being an 'exploded' view of components.
  • The suspended chunks look more like pieces of bread or nuggets than typical burger ingredients.

Wan 2.7

  • + Perfectly adheres to the 'exploded' burger concept with clearly separated components.
  • + Text rendering is excellent, featuring the requested fiery glow, starburst, and accurate spelling.
  • + Dynamic composition with splashes of sauce and smoke provides a strong sense of motion.
  • The cucumber/pickle slice looks slightly less integrated into the physics of the scene.
  • The overall image has a slightly more illustrative/rendered feel compared to the raw realism of Model A's meat texture.

Verdict: Wan 2.7 is the clear winner as it perfectly follows all prompt instructions, including complex layout requirements and perfect text rendering. While FLUX.1 [schnell] captures a high degree of photorealism in the food itself, it fails significantly on the text/typography and does not truly 'explode' the burger as requested.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + The handwriting style looks very authentic and hand-drawn.
  • + Includes realistic chalk texture and natural variations in letter size.
  • Numerous spelling errors and gibberish throughout the menu items.
  • Failed to follow the prompt's specific text instructions for several food items.

Wan 2.7

  • + Perfect adherence to the requested text and spelling for every item.
  • + Captures the 'cozy café' atmosphere well with background lighting and plants.
  • + Excellent chalk smears and board texture for realism.
  • The text looks slightly too uniform, leaning towards a digital font style rather than raw handwriting.

Verdict: Wan 2.7 is the clear winner as it correctly rendered every specific word and date requested in the prompt with 100% accuracy. FLUX.1 [schnell] produced significantly more realistic chalk 'handwriting', but failed completely on spelling and prompt adherence for the specific menu items.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Followed the difficult logical instruction of putting the horse on top of the astronaut.
  • + Excellent cinematic lighting and surreal composition.
  • + Highly detailed rendering of the space suit and horse textures.
  • The horse has anatomical issues, appearing to have two heads or a conjoined body.
  • The astronaut's posture is somewhat ambiguous due to the perspective.

Wan 2.7

  • + Clean, high-resolution visual quality with sharp details.
  • + Beautiful background featuring galaxies and a clear planetary curve.
  • Completely failed the negative constraint/instruction for the horse to be on top.
  • Cliche interpretation that ignores the specific spatial request of the prompt.

Verdict: FLUX.1 [schnell] successfully followed the complex instruction to reverse the typical rider/mount relationship, resulting in a truly surreal image as requested. In contrast, Wan 2.7 produced a generic 'astronaut riding a horse' image, failing the primary logical challenge of the prompt despite having good technical clarity. FLUX.1 [schnell] is the winner for its superior prompt adherence and creativity.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Excellent photographic lighting and depth of field
  • + Realistic fur texture on the capybara
  • + High-quality text rendering on the hat
  • The capybara's paws are not placed correctly on the steering wheel
  • The passenger is sitting in the middle rather than clearly in the back seat

Wan 2.7

  • + Accurately places both paws on the steering wheel as requested
  • + Better spatial composition showing the taxi exterior and interior simultaneously
  • + Correct placement of the passenger in the rear seat area
  • The capybara's fur looks somewhat needle-like and less natural
  • The businesswoman's face and hands are slightly distorted

Verdict: While FLUX.1 [schnell] produces a more aesthetically pleasing image with superior lighting and texture, Wan 2.7 followed the specific physical instructions of the prompt much better, including the paw placement and the seating arrangement. Wan 2.7 captures the 'bored' expression requested for the passenger more effectively, despite having lower overall image fidelity than FLUX.1 [schnell].

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Features a moody atmosphere with good cinematic lighting on the jack-o-lantern.
  • + Strong graphic design style that fits a modern digital invitation.
  • Significant text errors including typos and repeated lines of nonsense text.
  • Fails to include the parchment texture requested in the prompt.

Wan 2.7

  • + Perfect text rendering of all requested details without typos.
  • + Excellent adherence to the 'vintage gothic' and 'parchment' aesthetic with detailed borders.
  • + Rich composition featuring all elements like thorns, webs, and twisted trees clearly.
  • The illustration style is a bit more like a cartoon/comic than a cinematic poster.
  • Slightly less 'dark' than requested as the parchment is quite light.

Verdict: Wan 2.7 is the clear winner as it followed every instruction, including specific text strings, complex border requirements, and the vintage parchment style. FLUX.1 [schnell] struggled significantly with the text, producing several typos and nonsensical lines of data at the bottom of the invitation.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Clean minimalist composition
  • + Accurate Japanese flag representation
  • + Rendered text 'JAPAN' is legible
  • Failed to include the word 'SUSHI' in text
  • Texture on the salmon looks slightly unnatural or sauce-heavy

Wan 2.7

  • + Followed all text instructions including 'JAPAN' and 'SUSHI'
  • + Excellent 3D miniature styling with high-quality PBR-like textures
  • + Great variety of sushi types that fit the 'cartoon scene' request
  • The flag icon is slightly off-center compared to the text

Verdict: Wan 2.7 followed all parts of the prompt, including the specific text requirements and a diverse miniature scene. FLUX.1 [schnell] produced a very clean image but missed the secondary text 'SUSHI' and had a much more limited interpretation of the scene.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Warm, cinematic lighting with a soft artistic bokeh
  • + Whimsical, storybook-like aesthetic
  • Failed to include the bunny
  • Anatomical issues with the center kitten appearing to have extra ears or fused features with a second kitten
  • The fox has an unnaturally small body

Wan 2.7

  • + Successfully included all four requested animals (dog, cat, bunny, fox)
  • + Excellent lighting effects including god rays and dew sparkles as requested
  • + Better rendering of animal anatomy and distinct fur textures
  • The kitten has an slightly odd, tiny tongue artifact
  • The composition is a bit crowded with the animals arranged in a single line

Verdict: Wan 2.7 is the clear winner as it followed the prompt much more accurately by including all four distinct animals, whereas FLUX.1 [schnell] missed the bunny and produced a confused anatomical mashup in the center. Wan 2.7 also better captured technical details like god rays and dew sparkles while maintaining a higher degree of realism.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Clean vector aesthetic suitable for a minimalist logo
  • + Good color palette adherence with brown and cream tones
  • Significant text spelling errors (Framilan instead of Florian, 7720 instead of 1720)
  • Missing the steam element requested in the prompt

Wan 2.7

  • + Accurately depicts the cloche dome with steam and the ribbon banner
  • + Correct date 'Est. 1720' and close spelling of 'Florion'
  • + Excellent vintage texture and balanced composition
  • One-letter spelling error in the name ('Florion' instead of 'Florian')
  • Slightly less 'minimalist' than Model A due to the detailed border

Verdict: Wan 2.7 is the clear winner as it followed almost every specific detail of the prompt, including the steam and the correct historical date. While FLUX.1 [schnell] has a nice minimalist vector feel, its catastrophic failure to spell the name correctly and the incorrect year (7720) make it unusable as a logo.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 [schnell]
Wan 2.7

AI Judge Analysis

FLUX.1 [schnell]

  • + Successfully uses the requested navy and red palette.
  • + Clean vector circles provide a cohesive diagrammatic feel.
  • Text is largely illegible gibberish.
  • The diagram logic is confusing and doesn't clearly map to the six requested steps.
  • Artifacts on the rocket icon undermine the 'crisp lines' requirement.

Wan 2.7

  • + Highly legible and accurate text rendering for the title and steps.
  • + Perfectly follows all six prompt steps with appropriate iconography for each.
  • + Excellent formatting that looks like a professional infographic.
  • Minor spelling error in 'DESCRIPT' (should be Descent) and 'Tranquiliry'.
  • Slight lack of vertical space makes the bottom astronaut section feel slightly crowded.

Verdict: Wan 2.7 significantly outperformed FLUX.1 [schnell] by following the specific logical structure of the prompt. While FLUX.1 [schnell] produced a vague diagram with unreadable text, Wan 2.7 created a fully functional educational infographic with clear headings, supporting data, and correct NASA-inspired styling.

Next steps

Explore each model