Fast distilled version of Black Forest Labs' FLUX.2 [dev] optimized for speed and cost efficiency.
Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.
FLUX.2 [dev] Flash
#7 of 62 in Text-to-Image
Qwen Image 2.0
#34 of 62 in Text-to-Image
Where the votes landed
FLUX.2 [dev] Flash
0.0%
win rate
Ties
0.0%
Qwen Image 2.0
100.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent adherence to lighting and material properties.
- + Realistic glass refraction and plant visibility through the cube.
- + Highly realistic textures on the wooden table and book.
- − The sphere is slightly off-center, though this is a minor aesthetic choice.
Qwen Image 2.0
- + Accurate placement of all requested elements.
- + Bright, clear photograph-style composition.
- + Clean representation of the cube structure.
- − The sphere appears to be floating mid-air inside the cube without physical support.
- − Reflections on the side panels of the cube look more like mirrors than transparent glass.
Verdict: FLUX.2 [dev] Flash produced a significantly more realistic image with convincing physics and light interaction, particularly how the plant is refracted through the thick glass. Qwen Image 2.0 followed the prompt accurately but suffered from a 'floating' object effect and less realistic material properties for the glass and lighting.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent depiction of shallow depth of field and bokeh from background traffic.
- + Highly realistic skin textures and fine details on the elderly man's hands.
- + Captures the 'cinematic' lighting and wet pavement reflections perfectly.
- − The bike anatomy is slightly confused near the handlebars, with multiple brake cables leading nowhere.
- − The framing is a bit more centered than the 'imperfect framing' request might suggest.
Qwen Image 2.0
- + Successfully captured a more 'candid' and imperfect framing with the subject partially cut off.
- + Strong adherence to the 'reflections on wet pavement' and 'shallow depth of field' requirements.
- + The interaction between the hands and the bike pedal appears very natural.
- − Missing the 'motion blur from passing cars' requested in the prompt.
- − The background car feels somewhat static compared to the atmosphere requested.
- − Slightly less 'cinematic' in terms of color grading compared to the other model.
Verdict: FLUX.2 [dev] Flash delivered a superior cinematic atmosphere with beautiful motion blur and lighting that perfectly matched the prompt's mood. While Qwen Image 2.0 did a better job with the 'imperfect framing' and realistic candid posing, it failed to incorporate the requested motion blur, making FLUX.2 [dev] Flash the overall winner for technical prompt adherence and visual quality.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent adherence to the 'faint scars and dirt' prompt with realistic facial textures
- + Beautifully detailed engraving on the plate armor
- + Highly realistic eyes and hair bead details
- − The scars appear slightly fresh and bloody rather than just battle-worn
- − The bokeh sparks are a bit uniform across the image
Qwen Image 2.0
- + Strong 'battle-worn' character design with aged skin and clear scarring
- + Includes a sword which adds to the paladin theme
- + Vibrant warm lighting and bokeh
- − The hand and sword pommel have noticeable anatomical and structural distortions
- − The armor engraving is somewhat less refined compared to Model A
- − Resolution looks slightly lower with some blurring on the shoulder pauldrons
Verdict: FLUX.2 [dev] Flash produced a significantly cleaner and more detailed image with superior anatomical accuracy, particularly regarding the eyes and skin texture. While Qwen Image 2.0 captures an excellent 'grizzled' look for the warrior, it suffers from typical AI artifacts in the hand and weapon areas, making FLUX.2 the more polished overall generation.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent adherence to the 'sections' requirement with distinct text blocks.
- + Includes vibrant color accents on the borders that add to the professional aesthetic.
- + High resolution for the food photography within the grid.
- − The text is largely nonsensical and garbled.
- − The categorization is confusing as 'Appetizers' and 'Mains' labels are used for pizza images.
Qwen Image 2.0
- + Strict adherence to the grid layout for food photos.
- + Food images are diverse and clearly match the Appetizer/Pizza/Mains categories.
- + Clean, modern minimalist aesthetic with rounded corners on images.
- − Text is somewhat illegible/pseudonymized.
- − Missing the 'list' style sectioning, functioning more like a digital menu board than a traditional print menu.
Verdict: FLUX.2 [dev] Flash captures the specific request for a professional table menu layout with distinct sections for text and imagery, though it struggles with meaningful text. Qwen Image 2.0 provides a much better grid of diverse food photos that correctly correspond to the categories, making it more visually coherent as a concept for a digital display.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent fiery typography that perfectly captures the glow effect requested.
- + Superior photorealistic textures, particularly on the meat patty and fresh lettuce.
- + Very clean starburst design for the price point.
- − The burger components are a bit static despite being suspended.
- − The lettuce is slightly repetitive in its layers.
Qwen Image 2.0
- + Strong sense of motion with steam and flying food particles.
- + Bold, aggressive 'Magic Burger' text with detailed flame effects.
- + Dynamic lighting that makes the burger elements pop against the dark background.
- − The price starburst looks a bit cluttered and less integrated than Model A.
- − The 'Limited Time Only' text is somewhat small and less impactful.
Verdict: Both models followed the prompt exceptionally well, producing high-quality commercial-style graphics. FLUX.2 [dev] Flash is preferred for its cleaner composition and superior photorealism in the food textures, whereas Qwen Image 2.0 captures a slightly better sense of dynamic motion through the use of steam and embers.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent text rendering with no spelling errors.
- + Authentic chalk texture and realistic handwriting variation.
- + Strong adherence to the layout requested in the prompt.
- − The pricing for the cookies is repeated twice ($9).
Qwen Image 2.0
- + Natural lighting and depth of field in the cafe background.
- + Good chalk smudge effects that enhance realism.
- + Correct spelling of menu items.
- − The spacing after the dashes in prices is inconsistent and awkward.
- − The crop is a bit tight on the right side of the chalkboard.
Verdict: Both models performed exceptionally well on a difficult text-rendering task. FLUX.2 [dev] Flash is the winner because its handwriting feels more organically integrated into the chalkboard surface and it followed the specific formatting of the prompt more cleanly, whereas Qwen Image 2.0 had slightly awkward spacing around its dashes and symbols.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + High level of kinematic detail in the spacesuit and horse fur
- + Excellent background composition with celestial bodies
- + Consistent lighting across the subjects
- − Anatomical error with a third leg appearing behind the main horse body
- − Failed the logical reversal requested in the prompt
Qwen Image 2.0
- + Creative scaly texture on the horse's neck and shoulders
- + Vibrant, magical aesthetic with floating water droplets
- + Clear, high-contrast visual style
- − Failed the logical reversal requested in the prompt
- − Horse's back legs appear slightly distorted in their joints
Verdict: Both models failed the specific logic test in the prompt requesting the 'horse on top' of the astronaut, instead providing the standard astronaut-on-horse configuration. FLUX.2 [dev] Flash has superior cinematic realism and background detail, but suffers from a significant anatomical glitch (an extra leg), whereas Qwen Image 2.0 provides a cleaner, albeit more fantastical, execution.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent photorealism with cinematic lighting and realistic textures.
- + Perfect adherence to the 'both front paws on the steering wheel' instruction.
- + The businesswoman's expression and hand details are very well rendered.
- − The passenger is sitting in the front seat instead of the back seat as requested.
Qwen Image 2.0
- + Great color vibrancy and sharp details on the capybara's fur.
- + Accurately represents the bored expression of the businesswoman.
- + Dynamic angle that captures the essence of a New York night.
- − The passenger is sitting in the front seat, failing the 'back seat' part of the prompt.
- − The positioning of the capybara's paws on the wheel is less convincing and anatomically awkward compared to Model A.
Verdict: Both models failed to place the passenger in the back seat, placing her in the front next to the capybara instead. However, FLUX.2 [dev] Flash is the winner due to its superior photorealistic quality, better hand/paw rendering, and more natural integration of the subject into the environment.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent typography style that feels premium and legible
- + High detail in the thorny border and cobwebs
- + Very polished, cinematic lighting on the jack-o-lantern
- − Includes a line of gibberish text 'Hurk: 0369c' in the event details
Qwen Image 2.0
- + Accurately renders all requested text without adding hallucinations
- + Stronger parchment texture that matches the 'vintage' part of the prompt
- + Good composition with a clear background depth
- − The transition between the parchment paper and the central image is a bit harsh and lacks blending
- − The bat silhouettes are very basic compared to the rest of the illustration
Verdict: Both models followed the complex prompt effectively, including all specific text elements and atmospheric details. FLUX.2 [dev] Flash has the superior visual polish and more impressive font choices, but it fails on text accuracy by adding a line of gibberish. Qwen Image 2.0 captures the vintage parchment aesthetic more effectively and rendered the text perfectly, making it the more reliable choice for an invitation.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent typography alignment and integration of the flag icon.
- + Strong adherence to the isometric perspective and diorama style.
- + Highly refined 3D cartoon surfaces with believable soft textures.
- − The sushi roll construction is slightly nonsensical with fish draped over a maki-like interior.
Qwen Image 2.0
- + More realistic food textures on the fish and eel.
- + Accurate sushi types including nigiri and maki.
- + Good use of the dioroma base with rounded edges.
- − The flag icon is oversized and poorly placed compared to the text.
- − Shadows on the blue background are a bit inconsistent with the lighting of the base.
Verdict: FLUX.2 [dev] Flash delivered a more cohesive 'miniature 3D cartoon' aesthetic with superior graphic design for the text and flag elements. While Qwen Image 2.0 has more realistic food variety, its layout feels less like a polished 3D diorama and more like a standard photo composite.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent depiction of dew sparkles and golden hour lighting.
- + Very soft and detailed fur texture across all animals.
- + Charming, expressive character designs.
- − The composition is a bit static and posed rather than playful/tumbling.
- − Duplicate animals (two foxes and two bunnies) instead of the four specific ones requested.
Qwen Image 2.0
- + Successfully captures the requested 'tumbling' and 'playful' interaction.
- + Stronger adherence to the specific count of animals (one golden retriever, one kitten, one bunny, one fox).
- + Beautiful god rays and vibrant meadow environment.
- − The fox's face/mouth area appears slightly anatomicaly distorted in the tumble.
- − The butterfly on the right has a slightly artificial, flat appearance.
Verdict: Both models produced high-quality, endearing images, but Qwen Image 2.0 followed the prompt more accurately by including exactly one of each animal and capturing the requested action of them 'tumbling together'. While FLUX.2 [dev] Flash has slightly more refined fur textures and lighting, it failed the specific inventory of animals by duplicating the fox and bunny.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent typography rendering including the grave accent on Caffè.
- + Strong vector illustration style with a clear professional emblem layout.
- + Applied a subtle, pleasing vintage paper texture to the background.
- − The 'Est. 1720' text is placed below the banner rather than inside it.
Qwen Image 2.0
- + Successfully integrated the 'Est. 1720' text into the banner.
- + Accurate brown and cream color scheme.
- + Clean resolution and sharp outlines.
- − The steam is rendered inside the cloche dome rather than rising from it.
- − The composition feels slightly cluttered with text overlapping the cloche base.
- − The typography is a bit generic compared to the first model.
Verdict: FLUX.2 [dev] Flash produced a much more convincing logo design with a professional vector aesthetic and superior typography. While Qwen Image 2.0 followed the banner instruction more literally by placing the date inside it, the overall design and the odd placement of steam inside the dome makes it less successful as a logo.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent typography rendering for the main title and astronaut names.
- + High level of detail in the lunar module and landing site illustrations.
- + Successfully includes creative additions like astronaut portraits that fit the NASA aesthetic.
- − Confusing layout with overlapping paths and duplicate 'Lunar Orbit' and 'Descent' labels.
- − Some garbled text in the first step (Launch) and the word 'MOON'.
- − The rocket depicted is a generic booster design rather than the iconic Saturn V requested.
Qwen Image 2.0
- + Perfect logical flow following the requested sequence from top to bottom.
- + Very clean, minimalist vector aesthetic that adheres strictly to the 'flat-vector' instruction.
- + Near-perfect spelling of labels and logical placement of icons.
- − Minor spelling error in 'Translunjar'.
- − The icons are much simpler than the detailed illustrations in the other model.
- − Text formatting on 'Lunar Orbit' is slightly cramped within the graphical elements.
Verdict: Qwen Image 2.0 is the winner because it follows the logical structure of an infographic much better than FLUX.2 [dev] Flash, which suffers from a chaotic layout and repetitive labels. While FLUX.2 [dev] Flash has higher individual asset quality, Qwen Image 2.0 accurately represents the six-step sequence requested in a clean, professional manner.
Explore each model
Alibaba's Qwen Image 2.0 model with enhanced text rendering, supporting both Chinese and English prompts with up to 6 images per request