Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 [dev] Black Forest Labs FLUX.1 Kontext [dev] Black Forest Labs

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

FLUX.1 [dev]

24.6 arena score

#16 of 62 in Text-to-Image

Skill signature · Text-to-Image

FLUX.1 Kontext [dev]

16.5 arena score

#58 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 [dev]

0%

win rate

Ties

0%

FLUX.1 Kontext [dev]

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent handling of glass refraction and reflections
  • + Highly realistic textures on the book cover and wooden table
  • + Sophisticated lighting that feels natural and soft
  • Physical logic error: the sphere appears to be floating inside the cube
  • The glass cube has internal vertical planes that make it look more like a glass frame or shelving unit

FLUX.1 Kontext [dev]

  • + Stronger adherence to the layout: the plant is clearly visible through the glass
  • + Better physical grounding: the sphere rests on the floor of the cube
  • + Clear, sharp focus on all required elements
  • Lighting is slightly flatter compared to Model A
  • The bottom of the cube appears to be a mirror rather than clear glass resting on the table

Verdict: Both models followed the prompt perfectly regarding the presence and color of objects. FLUX.1 Kontext [dev] is the winner because it correctly placed the plant behind the glass as requested and grounded the sphere physically, whereas FLUX.1 [dev] produced a floating sphere and a more abstract glass structure.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent shallow depth of field and bokeh effect
  • + Natural, highly detailed skin texture on the subject
  • + Subtle and realistic rain/wet sidewalk textures
  • The subject is standing next to the bike rather than actively 'repairing' it
  • Minimal motion blur on the passing vehicle

FLUX.1 Kontext [dev]

  • + Better fulfillment of the 'motion blur' aspect in the background cars
  • + Good reflections on the wet asphalt pavement
  • + Stronger sense of 'candid' street photography framing
  • The bicycle appears to be passing through the man's leg, a major anatomical error
  • The overall image quality and textures look more digital and less like a 50mm lens photo
  • The pose is sitting/standing over the bike rather than repairing it

Verdict: FLUX.1 [dev] is the superior image due to its significant lead in anatomical correctness and photographic realism, capturing natural skin textures and light beautifully. While FLUX.1 Kontext [dev] followed the motion blur prompt more closely, it suffered from a major clipping artifact where the bicycle frame merges through the subject's leg, and the overall render quality felt less 'natural' than requested.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent skin texture with realistic pores and freckles.
  • + Intricate metal texture on the gorget reflecting warm light.
  • + Clean, lifelike eye rendering.
  • Failed to include 'braided hair with small beads', showing simple long braids instead.
  • Armor lacks the requested 'ornate engraving'.
  • Subject appears too youthful and clean for the 'battle-worn' description.

FLUX.1 Kontext [dev]

  • + Successfully captured 'ornate engraved plate armor' with high detail.
  • + Strong application of 'warm torchlight' and 'bokeh sparks' in the background.
  • + Face clearly shows 'faint scars' and a 'battle-worn' expression.
  • Missed the 'hair braided with small beads' requirement; hair is mostly loose.
  • Leather straps and cloth underlayer are less defined than the metalwork.

Verdict: FLUX.1 Kontext [dev] followed the prompt more comprehensively, particularly regarding the ornate engraving and the 'battle-worn' aesthetic of the character. While FLUX.1 [dev] produced a cleaner skin texture, it failed to incorporate the specific engraving and character-aging details that define the paladin persona.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent professional structure that closely mimics a real menu layout.
  • + Legible information hierarchy with clear sections for Appetizers and Mains.
  • + Clean use of white space and professional typography.
  • Missed the 'grid' request for the photos, placing them as isolated elements instead.
  • Missing the specific 'pizza' section requested in the prompt.

FLUX.1 Kontext [dev]

  • + Strong adherence to the 'grid' layout for food photography.
  • + Visually vibrant colors and bold, modern sans-serif fonts.
  • Poor logical structure for a menu, with very little actual menu content compared to photos.
  • Garbled text and nonsensical headings make it feel like a collage rather than a functional design.
  • Failed to include clear sections for appetizers, pizza, and mains.

Verdict: FLUX.1 [dev] produced a highly professional and usable menu design that captures the 'casual dining' vibe perfectly, despite missing the photo grid requirement. FLUX.1 Kontext [dev] followed the grid instruction well but failed significantly on the layout logic and textual content, resulting in a design that doesn't feel like a functional menu.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent 'exploded' layout showing individual burger components as requested.
  • + Realistic textures on the food items and dynamic lighting.
  • + Clean and legible text for the secondary messages.
  • Completely failed to include the primary title text 'MAGIC BURGER'.
  • Background lacks the 'glowing embers' and fiery intensity requested in the prompt.

FLUX.1 Kontext [dev]

  • + Included all requested text elements, including a clear starburst for the price.
  • + Vibrant and high-impact fiery background with glowing coals and embers.
  • + Strong commercial advertisement composition with bold typography.
  • Failed the 'exploded' instruction; the burger is fully assembled rather than deconstructed in mid-air.
  • Minor typo in the text ('LNHLY' instead of 'ONLY').

Verdict: Both models struggled with specific parts of the prompt. FLUX.1 [dev] followed the 'exploded' burger layout perfectly but missed the most important text element ('MAGIC BURGER'), whereas FLUX.1 Kontext [dev] captured the ad's atmosphere and text much better but ignored the physical deconstruction of the burger. FLUX.1 Kontext [dev] is the likely winner for better adhering to the overall advertisement aesthetic and capturing the fiery theme.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent text accuracy with no spelling errors in the main menu items.
  • + Consistent and elegant handwriting style that looks authentic to chalk.
  • + Clean composition with natural background blurring.
  • The handwriting style is slightly too uniform, appearing almost like a digital handwriting font despite the prompt's request for natural variations.
  • Minor garbling in the bottom fine-print sentence ('drish').

FLUX.1 Kontext [dev]

  • + Strong chalk texture and authentic variation in letter weights.
  • + Includes more realistic 'smudge' artifacts typical of real chalkboard art.
  • Significant spelling errors and repetitions (e.g., 'Mushroom Mashroom', 'with with', 'Risoktso').
  • The date at the top is illegible and garbled.
  • Layout becomes cluttered with overlapping symbols and repeated prices.

Verdict: FLUX.1 [dev] is the clear winner as it successfully rendered almost all the requested text with perfect spelling and a professional aesthetic. While FLUX.1 Kontext [dev] had a slightly more realistic 'chalk' texture, it failed significantly on prompt adherence by including numerous typos and garbled text strings.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent cinematic lighting and composition
  • + High level of detail in the spacesuit and horse's mane
  • Fails the specific prompt instruction to have the horse on top of the astronaut
  • Shows a standard rider/mount relationship instead of the requested surreal reversal

FLUX.1 Kontext [dev]

  • + Successfully follows the surreal instruction of 'horse on top'
  • + Clear facial features through the helmet
  • + Maintains high resolution and realistic textures
  • The 'riding' aspect is a bit ambiguous as the horse is floating behind/on him rather than clearly mounted
  • The horse's hind legs are anatomically awkward as they merge/disappear behind the astronaut

Verdict: While FLUX.1 [dev] produced a more aesthetically pleasing and cinematic image, it completely failed the specific logical instruction for the horse to be on top. FLUX.1 Kontext [dev] followed the surrealist prompt requirements much more accurately, even though the anatomical merging of the horse's legs is a bit messy.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent photographic depth of field and lighting
  • + Captures the bored, non-chalant expression of the passenger perfectly
  • + Includes both paws on the steering wheel as requested
  • Composition feels more like a front-seat perspective than a passenger in the back
  • The capybara's claws look slightly distorted

FLUX.1 Kontext [dev]

  • + Successfully places the passenger clearly in the back seat
  • + Sharp detail on the capybara's fur and jacket
  • + Realistic car interior lighting and street bokeh
  • Fails the prompt requirement of having both front paws on the steering wheel
  • Passenger's hand holding the phone appears anatomically awkward

Verdict: FLUX.1 [dev] delivers a more cinematic and atmospheric image with better attention to the physical requirements of the prompt, such as the placement of the capybara's paws. While FLUX.1 Kontext [dev] handles the spatial relationship between the front and back seats more accurately, its failure to place both paws on the wheel and the slight distortion in the passenger's hand makes it less successful overall.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Strong illustrative art style with excellent lighting and contrast
  • + Atmospheric composition with a glowing moon and detailed thorny border
  • + Clean layout that balances the visual elements and the text
  • Several typos in the text including 'Lalloween Ranty' and 'You Tre'
  • Repeated the time '7pm, 7pm' and misspelled 'Arches'

FLUX.1 Kontext [dev]

  • + Perfectly rendered main title text with no spelling errors
  • + Accurate date and time formatting in the footer
  • + Good inclusion of all requested elements like the scroll and thorny border
  • Significant text corruption within the scroll banner
  • Misspelled the location name 'The Arches'
  • Flat lighting compared to the more cinematic feel of the prompt

Verdict: Both models struggled with the complex multi-part text requirements, particularly the location name. FLUX.1 [dev] produced a far more visually appealing and atmospheric image with cinematic lighting, but FLUX.1 Kontext [dev] captured the main title text correctly and adhered better to the 'square' layout requirement for text placement.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent 3D miniature diorama feel with realistic depth of field
  • + High-quality PBR material textures on the salmon and rice
  • + Captures the request for a soft, refined aesthetic with elegant lighting
  • Failed to render the word 'SUSHI' correctly, adding extra characters
  • The flag icon is integrated into a larger, cluttered graphic cluster

FLUX.1 Kontext [dev]

  • + Perfect text rendering of 'JAPAN' and 'SUSHI'
  • + Bold, clean 2D/3D hybrid style that is easy to read
  • + Large, clear presentation of the requested flag icon
  • The flag icon is abstract and does not represent the Japanese flag
  • Lacks the sophisticated PBR textures and 'miniature' diorama detail of the other model
  • The sushi model is very basic, resembling a simple toy rather than refined 3D art

Verdict: FLUX.1 [dev] produced a much higher quality 3D scene with beautiful textures and lighting that perfectly met the 'miniature diorama' request, despite the spelling error in the text. FLUX.1 Kontext [dev] followed the text layout instructions perfectly but failed on the flag's appearance and provided a much simpler, less visually appealing model.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent golden sunrise lighting with visible god rays and bloom.
  • + Charming, almost Pixar-like expressive eyes and character design.
  • + Includes a wider variety of animals that better match the prompt request.
  • Failed the 'hyper-photorealistic' instruction by producing a CG/illustrative style.
  • The anatomical details on the fox/rabbit creatures are a bit muddy and merged.
  • The kitten is missing or has been conflated with the puppies.

FLUX.1 Kontext [dev]

  • + Successfully captures a more photorealistic texture and lighting style.
  • + Displays much better action and 'tumbling' movement as requested in the prompt.
  • + Superior fine detail in the fur and whiskers.
  • Failed to include the specific variety of animals (bunny and fox are missing).
  • The butterflies appear somewhat flat compared to the rest of the scene.
  • The composition is a bit repetitive with three very similar-looking feline/canine hybrids.

Verdict: Both models failed to accurately count and distinguish the four specific animals requested, but for different reasons. FLUX.1 [dev] followed the 'wholesome' and 'expressive eyes' vibes but drifted into a 3D animation style, whereas FLUX.1 Kontext [dev] achieved the requested photorealistic look with dynamic movement despite missing half the animal subjects. FLUX.1 Kontext [dev] is the likely winner for its superior realism and texture, which were key components of the prompt.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent vintage aesthetic and vector emblem style
  • + Good implementation of the banner and steam elements
  • + Captures the 'minimalist' and 'cloche' aspect well
  • Significant spelling errors in the brand name ('Café Flariláan' and 'Reseaurant')
  • Includes random numbers (11011, 1941) not requested in the prompt

FLUX.1 Kontext [dev]

  • + Perfect text rendering and spelling for 'Caffè Florian'
  • + Clean, bold, and high-readability design
  • + Correctly implements all textual elements including 'Est. 1720'
  • Typography is a bit more modern/blocky than requested 'classic' style
  • The cloche dome is slightly more abstract/stylized than a traditional cloche

Verdict: While FLUX.1 [dev] captures the vintage aesthetic and vector emblem style much more effectively, it fails significantly on text spelling. FLUX.1 Kontext [dev] delivers a much more usable logo with perfect spelling and high clarity, even if it feels slightly less 'vintage'.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 [dev]
FLUX.1 Kontext [dev]

AI Judge Analysis

FLUX.1 [dev]

  • + Excellent layout that visually represents a mission trajectory/cycle.
  • + Strong adherence to the requested NASA-inspired color palette.
  • + Crisp vector-style graphics and professional infographic aesthetics.
  • Text consists entirely of gibberish despite having a structured layout.
  • Iconography doesn't clearly match the specific technical steps requested (e.g., Saturn V).

FLUX.1 Kontext [dev]

  • + Successfully rendered some legible words like 'Apollo 11' and 'Launch'.
  • + Attempts a step-by-step numbering system as requested in the prompt.
  • Very poor layout for an infographic, looking cluttered and disorganized.
  • Text rendering is messy with overlapping characters and spelling errors like 'Apolo'.
  • Fails to follow the specific mission sequence requested in the prompt.

Verdict: FLUX.1 [dev] produced a much more professional and aesthetically pleasing infographic that accurately captured the requested 'modern vector' style and color palette, though its text is nonsensical. FLUX.1 Kontext [dev] struggled with a messy composition, spelling errors, and a failure to follow the requested chronological mission steps, making it less useful as a design piece.

Next steps

Explore each model