Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [pro] Black Forest Labs Qwen Image Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [pro]

20.3 arena score

#41 of 62 in Text-to-Image

Skill signature · Text-to-Image

Qwen Image

20.3 arena score

#40 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [pro]

0%

win rate

Ties

0%

Qwen Image

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent photographic quality and realistic material textures
  • + Perfect execution of the 'partially visible through glass' requirement for the plant
  • + Accurate lighting direction with soft shadows consistent with the window source
  • The blue sphere appears slightly fuzzy/felt-like rather than smooth
  • The blue sphere appears to be floating inside the cube rather than resting on the bottom surface

Qwen Image

  • + Natural interaction between the blue sphere and the bottom of the glass cube
  • + Clean, glossy textures on both the sphere and the glass
  • The glass refractive logic is slightly flat compared to the plant behind it
  • Lower overall image resolution and sharpness compared to Flux

Verdict: FLUX.1 Kontext [pro] creates a more sophisticated and realistic scene with superior handling of the glass transparency and the complex visual of the plant seen through the cube. While Qwen Image handles the sphere's physical placement better, it lacks the overall photographic depth and texture quality of FLUX.1 Kontext [pro].

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent skin texture and facial detail
  • + Authentic shallow depth of field and rain effects
  • + Realistic 'imperfect' street photography framing
  • The man appears to be riding or leaning on the bike rather than actively 'repairing' it
  • Minimal motion blur on passing cars

Qwen Image

  • + Better narrative adherence to the 'repairing' part of the prompt
  • + Captures the reflection on the wet pavement effectively
  • + Includes motion blur on the background car as requested
  • Anatomical issues with the man's hands and fingers
  • Bicycle geometry is inconsistent (pedal placement and kickstand)
  • Lower overall resolution and slightly plastic skin texture

Verdict: FLUX.1 Kontext [pro] creates a far more convincing and high-quality image with superior skin textures and lighting, though it misses the specific action of 'repairing'. Qwen Image follows the narrative prompt better by showing the man working on the bike with motion blur in the background, but it suffers from significant anatomical distortions and a lower quality aesthetic.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Extremely lifelike skin texture and realistic, subtle scarring.
  • + Excellent lighting integration with warm highlights feeling organic to the scene.
  • + Superior rendering of materials including the rough woven scarf and engraved metal.
  • Missed the request for 'beads' in the hair.
  • Light bokeh is very subtle, bordering on absent.

Qwen Image

  • + Successfully included beads in the hair braids as requested.
  • + Very clear representation of the 'battle-worn' theme with distinct scars and dynamic sparks.
  • The sparks have a synthetic, 'clip-art' appearance rather than realistic bokeh.
  • Skin texture and facial hair look somewhat plasticky and less lifelike than the competitor.

Verdict: FLUX.1 Kontext [pro] produces a much more realistic and cinematic portrait with sophisticated textures and lighting, though it fails to include the beads mentioned in the prompt. Qwen Image adheres better to the specific literal prompts (beads and sparks) but suffers from lower visual quality, with digital-looking artifacts and less realistic facial features. FLUX.1 Kontext [pro] is the preferred choice for its professional photographic aesthetic and superior material rendering.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typographic hierarchy with clear section headers.
  • + Legible prices and text that resembles a real menu structure.
  • + High-quality, realistic food photography integrated into the layout.
  • Slightly inconsistent content (placing a pizza under the 'Mains' heading while listing a salad under the 'Pizza' heading).

Qwen Image

  • + Successfully creates a colorful grid of food photos as requested.
  • + Modern, high-contrast aesthetic suitable for casual dining.
  • The text is largely illegible gibberish.
  • Combines sections (Pizza/mains) rather than keeping them distinct as requested.
  • Layout is a bit cluttered compared to Model A.

Verdict: FLUX.1 Kontext [pro] produced a professional, highly usable menu design with clean typography and realistic imagery that follows the section requirements perfectly. Qwen Image followed the grid request well but failed significantly on text legibility and logical sectioning.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography style with glowing fire effects
  • + Highly photorealistic texture on the burger bun and meat
  • + Clean and professional layout despite the repetition
  • Repeated the price tag twice on the image
  • The burger is barely 'exploded', mostly just missing the top bun attachment
  • Missing the requested starburst for the price

Qwen Image

  • + Successfully captured the 'exploded' dynamic with flying ingredients
  • + Correctly included the price in a starburst as requested
  • + Better dynamic composition and sense of motion
  • Slightly less photorealistic lighting on the flying vegetables
  • Minor halo artifacts around the main title text
  • The lettuce and tomato look a bit plastic in texture compared to the other model

Verdict: Qwen Image followed the layout instructions more accurately, providing the requested 'exploded' effect with flying ingredients and the starburst element. While FLUX.1 Kontext [pro] has superior photorealism and better font rendering, it failed the composition logic by repeating the price and not actually exploding the burger as requested.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent chalk texture with realistic grit and stroke transparency
  • + Perfect adherence to the requested date: April 30, 2026
  • + Handwriting looks authentically non-uniform with character variety
  • The extra text at the bottom contains spelling errors like 'ous' and 'tree'
  • The title is in print-style caps rather than the requested elegant cursive

Qwen Image

  • + The handwriting style is more consistent with a stylized chalkboard menu's slant
  • + Displays a more complete café environment in the background
  • Significant error in the date rendering it as '20026' instead of '2026'
  • The text lacks realistic chalk texture, appearing more like a digital brush or paint pen
  • Incorrect misspelling of 'Risoto'

Verdict: FLUX.1 Kontext [pro] is the better image because it successfully renders the fine texture of chalk and adheres to the correct date, despite some minor spelling errors in the unrequested footer text. Qwen Image fails on the primary year requirement and has a smoother, less realistic text finish that looks more like a digital font than handwritten chalk.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Perfectly follows the specific instruction of having the horse on top of the astronaut.
  • + High degree of surrealism with a third tiny astronaut riding the horse on top.
  • + Exceptional cinematic lighting and texture on the space suit and horse hide.
  • The astronaut's feet have strangely been replaced with horse hooves.
  • The anatomy of the horse's back legs is a bit confused where they meet the astronaut.

Qwen Image

  • + Clean, high-resolution rendering of the astronaut and horse.
  • + Great lighting and background composition with the planet and moon.
  • Completely failed the negative constraint to have the horse on top.
  • Lacks the 'surreal' quality requested, opting for a standard trope interpretation.

Verdict: FLUX.1 Kontext [pro] successfully interpreted the difficult and specific spatial requirement of placing the horse on top of the astronaut, creating a truly surreal image. Qwen Image completely ignored the 'horse on top' instruction and produced a standard, cliché astronaut riding a horse, which fails the prompt's core challenge.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent fur texture and lighting on the capybara.
  • + The passenger has a bored, disinterested expression as requested.
  • + Realistic bokeh and street lighting through the window.
  • The passenger is holding a phone to her ear or touching her face rather than looking at it.
  • The capybara's paws are not clearly defined on the steering wheel.

Qwen Image

  • + Accurately depicts the passenger looking at her phone.
  • + Both paws are distinctly placed on the steering wheel.
  • + The composition feels more like a wide cinematic shot showing both the capybara and passenger clearly.
  • The capybara's paws look more like primate hands than capybara paws.
  • The taxi sign on top of the car contains a typo ('YOXI').
  • The capybara's face looks slightly more artificial/clean compared to the gritty realism of Model A.

Verdict: While FLUX.1 Kontext [pro] offers superior textures and a more authentic New York 'night' feel with realistic lighting, it misses the specific action of the passenger looking at her phone. Qwen Image follows the prompt instructions more precisely regarding the passenger's action and the placement of the paws, though it suffers from 'monkey-like' hands for the animal and some minor text errors. Overall, FLUX.1 Kontext [pro] is more visually compelling as a photorealistic image.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography rendering with almost perfect spelling of the prompt's required text.
  • + Complex and atmospheric border design featuring fine spiderweb and thorn-like details.
  • + Cinematic lighting on the jack-o-lantern and moon contributes to a high-quality vintage feel.
  • Includes a line of hallucinated text ('Your: Vorkleat: Iight & Spans') that was not in the prompt.

Qwen Image

  • + Strong adherence to the 'parchment' and 'thorn' border elements requested in the prompt.
  • + Atmospheric lighting on the trees and background creates a good sense of depth.
  • + Correctly follows the square format with a distinct paper-edge effect.
  • Main title text is garbled ('Halla Party Invitation') and contains illegible gibberish at the top.
  • The banner text is split awkwardly between the banner and the background.
  • Typography lacks the consistent professional polish found in the other model.

Verdict: FLUX.1 Kontext [pro] is the clear winner due to its superior text rendering and overall visual coherence; while it added one line of hallucinated text, the primary titles and details are legible and aesthetically pleasing. Qwen Image captured the parchment texture well but failed significantly on the typography, resulting in misspelled words and less elegant graphic design.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent 3D material rendering with realistic subsurface scattering on the fish
  • + Very clean and readable typography that fits the cartoon aesthetic
  • + Perfectly followed the 45-degree isometric camera angle requirement
  • The flag icon is stylized to the point of being a generic rectangle rather than a clear national flag

Qwen Image

  • + Includes a more varied sushi set which adds visual interest
  • + Distinct and accurate Japanese flag rendering both in text and as a 3D asset
  • + Great use of depth of field and soft lighting to create a miniature feel
  • The 'JAPAN' text is slightly cut off at the top of the frame
  • The isometric angle is a bit shallower than the requested 45 degrees

Verdict: Both models performed exceptionally well on this task, capturing the 'isometric miniature' aesthetic perfectly. FLUX.1 Kontext [pro] creates a cleaner single-object focus with superior texture work on the salmon, while Qwen Image provides a more complete diorama scene with additional sushi types and better integration of the flag. FLUX.1 Kontext [pro] is the slight winner for overall composition and text placement within the frame.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent soft lighting and atmospheric god rays
  • + Extremely detailed fur texture across all animals
  • + Whimsical, cohesive color palette that matches the '8K masterpiece' request
  • The 'bunny' looks like a hybrid kitten-rabbit creature
  • The animals are mostly static and posing rather than 'tumbling' or 'chasing'

Qwen Image

  • + Successfully depicts four distinct animal species (puppy, kitten, bunny, fox)
  • + Captures the 'playful chasing' and 'tumbling' action better than the competitor
  • + Clearer morning dew sparkles on the grass
  • The fox kit has three front paws visible, indicating an anatomical error
  • The lighting on the butterfly on the right is inconsistent with the background light source
  • The kitten has a slightly distorted front paw

Verdict: Both models struggle with animal anatomy when groups are involved; FLUX.1 Kontext [pro] produces a more aesthetically pleasing 'masterpiece' with superior lighting and fur detail, but it fails to create a realistic baby bunny, instead making a third cat-like animal. Qwen Image adheres better to the species requested and the action of the prompt, but suffers from a significant anatomical glitch (five legs on the fox) that breaks the photorealism.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Excellent typography with perfect spelling and correct character accents.
  • + Superior texture application that gives a genuine vintage printed paper feel.
  • + Clean, professional vector-style layout that follows minimalist design principles.
  • The steam element is a bit thick compared to the delicate lines of the rest of the logo.

Qwen Image

  • + Nice interpretation of the steam element.
  • + Good use of warm brown tones.
  • Severe typography errors including overlapping text and messy cursive integration.
  • The logo composition feels cluttered and lacks the requested minimalist aesthetic.
  • The 'Est. 1720' font is too modern and sans-serif for a vintage logo.

Verdict: FLUX.1 Kontext [pro] followed all instructions perfectly, delivering a clean, professional logo with crisp typography and authentic vintage texture. Qwen Image failed significantly on the text rendering, creating an illegible jumble of characters for the brand name.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [pro]
Qwen Image

AI Judge Analysis

FLUX.1 Kontext [pro]

  • + Strong aesthetic appeal with a professional vector art style.
  • + Effective use of the requested NASA-inspired color palette.
  • + Includes a complex trajectory map that visually connects the steps.
  • Nonsensical scientific details, such as Saturn rings around the Moon and rocket.
  • Text labels are garbled and do not match the requested steps accurately.
  • Failed to include the specific Lunar Module icons requested for the final steps.

Qwen Image

  • + Successfully included icons for the Lunar Module descent and landing.
  • + Clean, modern flat-vector layout that is easy to read.
  • + Follows the sequential numbering of the prompt more closely.
  • Included the prompt instructions '(Stop at landing)' as functional text in the image.
  • Frequent spelling errors in key terms like 'Sarth' and 'Arnnstrong'.
  • Confused the labels, for example, calling step 3 both 'Lunar' and 'Earth Orbit'.

Verdict: FLUX.1 Kontext [pro] creates a much more visually sophisticated and artistic poster, but it fails significantly on technical accuracy by placing Saturn-like rings on the Moon and rocket. Qwen Image follows the specific icon instructions for the Lunar Module much better and provides a clearer infographic structure, though it suffers from poor text rendering and inclusion of prompt meta-text. Overall, Qwen Image is a more useful infographic for the specific prompt, despite the typos.

Next steps

Explore each model