Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [max] Black Forest Labs Qwen Image Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [max]

23.7 arena score

#23 of 62 in Text-to-Image

Skill signature · Text-to-Image

Qwen Image

20.8 arena score

#38 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [max]

0%

win rate

Ties

0%

Qwen Image

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent handling of caustics and light refraction through the glass.
  • + High level of texture detail on the wooden table and the book spine.
  • + Realistic lighting that perfectly matches the 'from the left' instruction.
  • The plant is positioned more to the side than directly behind, though it is visible through the glass.

Qwen Image

  • + Clean, simple composition with clear shapes.
  • + Good prompt adherence regarding the positioning of objects.
  • + Vibrant colors on the red book and blue sphere.
  • Internal reflections within the cube are physically inconsistent, appearing as extra vertical bars.
  • Lighting is somewhat flat and lacks the dramatic window shadows present in the other model.

Verdict: FLUX.1 Kontext [max] produced a superior image with highly realistic lighting, detailed textures, and convincing physics for the glass cube. Qwen Image followed the spatial instructions well but struggled with the glass's internal reflections, creating distracting metallic-looking bars inside the cube.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Exceptional skin texture and realistic age details on the hands and face.
  • + Beautifully rendered rain environment with realistic wet pavement reflections.
  • + Accurate bike mechanical details like the chain and cassette.
  • The man's heritage appears more ambiguous/Mediterranean rather than specifically Japanese.
  • Lack of noticeable motion blur from the background traffic.

Qwen Image

  • + Stronger adherence to the request for an 'elderly Japanese man'.
  • + Good implementation of the requested motion blur on the passing white car.
  • + Accurate 'imperfect framing' that feels like a candid snapshot.
  • The bicycle frame geometry is physically impossible and distorted.
  • Lower overall resolution and softer details compared to the first image.
  • The rain effect is less convincing, looking more like static filter lines.

Verdict: FLUX.1 Kontext [max] produces a significantly higher quality image with stunning textures and realistic lighting, though it misses the specific ethnic marking requested in the prompt. Qwen Image adheres better to the subject's ethnicity and includes the requested motion blur, but fails significantly on the structural integrity of the bicycle and overall image clarity.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Extremely realistic skin texture with lifelike pores and wrinkles
  • + Masterful use of lighting and bokeh to create a cinematic close-up
  • + Highly intricate engraving detail on the plate armor
  • Missed the request for beads in the hair braids
  • The eyes, while striking, appear somewhat supernatural which wasn't specifically requested

Qwen Image

  • + Excellent adherence to the 'beads' and 'leather straps' requirements
  • + Clearly depicts the battle-worn nature with blood and scars
  • + Includes the actual torch to explain the lighting source
  • Lighting feels flatter and less integrated than the competition
  • Skin texture and hair rendering appear more like a video game render than a photograph

Verdict: FLUX.1 Kontext [max] produces a much higher quality, photorealistic image with superior lighting and texture, though it missed the specific detail of beads in the braids. Qwen Image followed every prompt instruction including the beads and straps but lacks the professional cinematic depth and realistic skin rendering of the first model.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent professional alignment and white space usage.
  • + Product photos are high quality and consistent in style.
  • + The typography feels more like a real minimalist menu design.
  • The text is largely gibberish and has scaling issues near the bottom.
  • Failed to clearly label the 'Appetizers' and 'Mains' sections as requested.
  • Most images appear to be pizza, lacking variety in food types.

Qwen Image

  • + Successfully followed the instruction for specific section headers (Appetizers, Pizza/mains).
  • + Excellent bold sans-serif font usage and vibrant color accents.
  • + The grid layout is highly creative and fits the modern casual dining brief well.
  • Text rendering contains several spelling errors (e.g., 'Pizzaurant').
  • The grid layout leaves a bit too much space on the right compared to the left.
  • The food photos look slightly more 'stock' and less natural than Model A.

Verdict: Qwen Image is the preferred model because it followed the functional requirements of the prompt more accurately, specifically including the requested section headers for appetizers and mains. While FLUX.1 Kontext [max] has a slightly more realistic photographic quality, Qwen Image captured the 'modern vibrant' aesthetic and the structural request much better.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealistic texture on the meat and bun
  • + Professional typography with a clean, centered layout
  • + Captures a gritty, cinematic atmosphere with realistic embers
  • The burger is mostly assembled rather than 'exploded' as requested
  • The price tag is not in a starburst as specified
  • Strange bun fragments floating on the sides look unnatural

Qwen Image

  • + Successfully follows the starburst price tag instruction
  • + The burger has a much better sense of motion and 'exploded' dynamics
  • + Text effect perfectly matches the 'fiery, glowing' prompt
  • Less photorealistic than Model A, appearing slightly more like a 3D render
  • The composition feels a bit cramped at the top
  • The lettuce and some mid-air fragments look slightly plastic

Verdict: Qwen Image followed the specific details of the prompt much better, including the starburst for the price and the 'exploded' nature of the burger. While FLUX.1 Kontext [max] has superior photorealistic textures and cleaner font rendering, it failed to incorporate the starburst and kept the burger mostly intact, making it a less accurate ad based on the prompt instructions.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent text rendering with zero spelling errors across the entire board.
  • + Highly realistic chalk texture with smudge details on the blackboard surface.
  • + Followed the date prompt exactly as 'APRIL 30, 2026'.
  • The handwriting style is more of a casual print than the 'elegant cursive' requested for the title.

Qwen Image

  • + Successfully captured a more stylized, handwritten aesthetic for the text.
  • + Good use of layout and spacing for a café menu.
  • Failed the date prompt by adding an extra zero, rendering it as '20026'.
  • The text 'Risotto' is misspelled as 'Risoto'.
  • The text looks more like a digital brush stroke than authentic grainy chalk.

Verdict: FLUX.1 Kontext [max] is the clear winner due to its impeccable text accuracy and superior chalk realism, including natural smudges on the board. While Qwen Image attempted a more stylistic handwriting approach, it failed on basic spelling and accurately rendering the requested date.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent cinematic lighting and textures on the horse's coat and spacesuit.
  • + Dynamic and fluid composition that feels truly surreal and ethereal.
  • + Superior background detail with realistic star fields and planet rendering.
  • Failed the negative constraint; the astronaut is riding the horse instead of the horse riding the astronaut.

Qwen Image

  • + Clean imagery with high contrast and sharp focus.
  • + Good anatomical representation of the horse and spacesuit.
  • Failed the negative constraint; the astronaut is riding the horse.
  • The composition feels a bit static compared to the dynamic floating pose in the other image.

Verdict: Both FLUX.1 Kontext [max] and Qwen Image failed the specific structural logic requested in the prompt (horse on top of astronaut), instead producing the common trope of an astronaut riding a horse. FLUX.1 Kontext [max] is the preferred image as it has significantly better cinematic lighting, fine detail, and a more believable space environment.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealism in fur texture and lighting
  • + Captures a very calm and professional expression on the capybara
  • + Accurate rendering of the black 'TAXI' text on the yellow cap
  • The passenger is holding a phone to her ear rather than looking at it as requested
  • Only one paw is clearly on the steering wheel

Qwen Image

  • + Follows the passenger's action better with her looking at the phone screen
  • + Shows both paws on the steering wheel as requested
  • + Includes a realistic taxi roof sign and ID badge on the jacket
  • The paws look more like primate hands than capybara paws
  • The text on the taxi roof sign is garbled ('YOXI')

Verdict: FLUX.1 Kontext [max] produces a more photorealistic image with superior textures and lighting, but it misses the specific detail of the passenger looking at her phone. Qwen Image adheres more closely to the technical composition of the prompt, including both paws and the passenger's specific action, though the 'hands' on the steering wheel are anatomically incorrect for a capybara.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with a cohesive gothic font choice
  • + Flawless spelling and layout of all requested text
  • + Highly atmospheric lighting and intricate border details
  • Repeats the location 'The Arches, NYC' twice at the bottom
  • The scroll banner is slightly less prominent than requested

Qwen Image

  • + Strong composition with twisted trees framing the pumpkin
  • + Captures the parchment texture effectively
  • + Included all required elements including the bats and thorny border
  • Significant spelling errors in the title text ('Halle Party')
  • Font choice for event details is a generic sans-serif that clashes with the theme
  • Text rendering on the top arc is illegible

Verdict: FLUX.1 Kontext [max] is the clear winner as it produced a professional-grade invitation with perfect spelling and beautiful gothic typography. While Qwen Image followed the layout instructions well, it failed on the critical task of rendering the title correctly and used an immersion-breaking modern font for the event details.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent text rendering with a stylized 3D bubbly font.
  • + High-quality textures on the fish and rice that give a premium '3D render' feel.
  • + Clean and balanced composition following the 45° isometric request.
  • Completely missed the requested flag icon.
  • The 'JAPAN' text is slightly off-center.

Qwen Image

  • + Accurately included the flag icon both near the text and as a 3D object.
  • + Excellent adherence to all prompt elements including the layered diorama base.
  • + Bold, clean typography that is perfectly centered.
  • The 3D model of the nigiri sushi is slightly less detailed/realistic than Model A.
  • Minor artifact on the small flag icon next to the word 'SUSHI'.

Verdict: FLUX.1 Kontext [max] has superior texture work on the food items, providing a higher sense of visual quality, but it failed to include the flag icon. Qwen Image followed the prompt more comprehensively, including multiple interpretations of the flag and maintaining a consistent miniature aesthetic, though the 3D text in FLUX.1 felt more integrated into the cartoon style. Qwen Image is the likely winner for better prompt adherence and symmetrical composition.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent soft lighting with a strong blooming effect that matches the sunrise theme.
  • + High consistency in the 'joyful' expression and facial features across all animals.
  • + Beautiful bokeh and integrated foreground flowers for a professional photographic feel.
  • The animals are largely static rather than 'playfully chasing' or 'tumbling' as requested.
  • The anatomy of the rabbit's paws is slightly indistinct.

Qwen Image

  • + Better captures the action of 'chasing' with dynamic poses for the kitten and fox.
  • + Stronger 'dew sparkles' effect on the grass as specified in the prompt.
  • + The butterflies are better defined with clearer markings.
  • The puppy’s expression is somber or neutral, which misses the 'joyful' vibe requested.
  • Perspective issues with animal sizing, making the fox and kitten appear quite tiny compared to the golden retriever.

Verdict: Both models captured all four requested animals and the specific lighting conditions well. Qwen Image followed the action part of the prompt better by showing the animals in motion, while FLUX.1 Kontext [max] delivered a more aesthetically pleasing, high-quality image with a better 'joyful' expression that matches the requested theme. FLUX.1 is preferred for its superior artistic composition and warmer, more coherent mood.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Perfect text rendering for both the name and the established date.
  • + Excellent adherence to the 'vintage minimalist' and 'vector emblem' style.
  • + Correct inclusion of the banner and the specific accent on 'Caffè'.
  • The steam effect is a bit small and simplistic.

Qwen Image

  • + Warm color palette and background texture are well-executed.
  • + Good use of whitespace within the cloche illustration.
  • Severe text rendering errors, with letters overlapping and misspelling 'Florian'.
  • Vector style is a bit chunky and less sophisticated than requested.
  • The banner lacks the vintage curve requested and looks modern in comparison.

Verdict: FLUX.1 Kontext [max] delivered a professional, high-quality logo that followed every part of the prompt, particularly the challenging typography. Qwen Image failed significantly on the text, resulting in overlapping characters and a jumbled layout. FLUX.1 Kontext [max] is the clear winner for its clarity, accuracy, and aesthetic balance.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [max]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent artistic vector style with a consistent, muted palette.
  • + Sophisticated composition that uses the layout of the poster to simulate a space journey.
  • + Very high level of detail on the lunar lander and rocket illustrations.
  • Several spelling errors in the labels (e.g., 'Lunear Modulle', 'Laarth').
  • Included Saturn and extra moons which were not part of the prompt.

Qwen Image

  • + Strong adherence to the requested sequential numbering of steps.
  • + Clean, professional flat-vector icons that match the modern infographic aesthetic.
  • + Accurately captured the mission crew names with only minor spelling issues.
  • Confused labels for the stages, mapping 'Sarth Orbit' to the launch and 'Translunar' to Earth orbit.
  • Literal inclusion of instructions like '(Stop at landing)' in the header text.

Verdict: FLUX.1 Kontext [max] produced a more visually stunning and artistically cohesive piece, though it struggled with factual consistency by adding extra planets and significant typos. Qwen Image followed the infographic structure much more closely and captured the specific numbering requested, but it suffered from a lower level of visual complexity and included literal prompt instructions in the final text. FLUX.1 Kontext [max] is the preferred choice for its superior illustration quality and better interpretation of the 'NASA-inspired' aesthetic.

Next steps

Explore each model