Head to head
Esc

Models · slot A

to navigate to pick

DALL-E 2 OpenAI FLUX.1 Kontext [pro] Black Forest Labs

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

DALL-E 2

15.8 arena score

#59 of 62 in Text-to-Image

Skill signature · Text-to-Image

FLUX.1 Kontext [pro]

20.3 arena score

#41 of 62 in Text-to-Image

Vote tally

Where the votes landed

DALL-E 2

0%

win rate

Ties

0%

FLUX.1 Kontext [pro]

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Decent lighting and reflection on the table surface.
  • Completely failed to follow the spatial instructions in the prompt.
  • Objects are merged together in a nonsensical way.
  • The plant is in a giant blue pot that occupies most of the frame.

FLUX.1 Kontext [pro]

  • + Perfect adherence to all spatial instructions and object descriptions.
  • + High visual quality with realistic textures on the book, sphere, and glass edges.
  • + Accurately represents the soft window light coming from the left as requested.
  • The blue sphere appears to be levitating slightly rather than resting on the table/cube bottom.

Verdict: FLUX.1 Kontext [pro] followed every specific instruction in the prompt, creating a clear, high-quality image that correctly positioned the sphere, book, and plant. DALL-E 2 produced a confusing composition where the objects were merged or scaled incorrectly, failing to follow the prompt's layout.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Strong bokeh effect matches the shallow depth of field request
  • + Captures an 'imperfect framing' look with a foreground obstruction
  • The main subject is completely out of focus, failing the prompt's implied focus on the man and textures
  • Low resolution and blurry details make it hard to distinguish the man or the bicycle clearly
  • Lacks the skin texture and cinematic realism specified in the prompt

FLUX.1 Kontext [pro]

  • + Excellent adherence to nearly all prompt details including age, ethnicity, and red bicycle
  • + High visual quality with realistic skin textures and clear rain effects
  • + Well-executed shallow depth of field with background bokeh from cars
  • The man is riding or holding the bike rather than actively 'repairing' it
  • The composition is quite centered, missing the 'imperfect framing' aspect of the prompt

Verdict: FLUX.1 Kontext [pro] creates a much more usable and high-quality image that respects the aesthetic requirements for skin texture and cinematic realism, although it interprets 'repairing' as simply handling the bike. DALL-E 2 captures the 'imperfect framing' well but fails significantly on technical quality, producing a completely out-of-focus image where no textures or details are visible.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Captures a tactile sense of rust and age
  • Severely lacks detail and clarity
  • Fails to render lifelike eyes or identifiable facial features
  • Incorrectly interprets the prompt as a macro shot of a miniature or statue
  • Heavy digital noise and artifacts

FLUX.1 Kontext [pro]

  • + Excellent adherence to all prompt details including braided hair, scars, and ornate armor
  • + Exceptional visual quality and photorealism
  • + Perfect lighting with warm torch reflections and bokeh sparks
  • The beads in the hair are minimal/subtle compared to the prompt request

Verdict: FLUX.1 Kontext [pro] followed the prompt nearly perfectly, delivering a high-fidelity, cinematic portrait with realistic skin textures and intricate armor engraving. In contrast, DALL-E 2 produced an abstract, low-resolution image that failed to capture any recognizable facial features or the requested 'lifelike eyes'.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Strong bold sans-serif typography
  • + High contrast visual style
  • Does not follow the grid layout for food photos
  • The food images are abstract and distorted into shapes
  • Fails to include the specific sections requested

FLUX.1 Kontext [pro]

  • + Excellent adherence to all prompt elements including sections and photos
  • + Clean, professional grid-like layout
  • + High visual clarity and realistic food photography
  • Text is mostly gibberish despite correct section headings
  • Slightly inconsistent pricing values

Verdict: FLUX.1 Kontext [pro] successfully created a functional and aesthetically pleasing menu design that perfectly followed the prompt's layout and content requirements. In contrast, DALL-E 2 produced an abstract artistic piece that ignored the structural instructions of the prompt, such as the specific menu sections and a clean grid of photos.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Successfully captures a sense of motion with particles
  • + Integrates the fiery theme into the core of the burger
  • Text is nonsensical and garbled
  • Image quality is low and lacks photorealistic detail
  • The burger ingredients are messy and difficult to identify

FLUX.1 Kontext [pro]

  • + Perfect text rendering for all requested strings
  • + Exceptional photorealistic detail in the food textures
  • + Dynamic composition with clear, high-quality visual effects
  • Includes the price twice, which was not requested
  • The burger is less 'exploded' than 'partially opened'

Verdict: FLUX.1 Kontext [pro] is the clear winner as it flawlessly renders all requested text and delivers a high-resolution, professional-grade advertisement aesthetic. In contrast, DALL-E 2 fails significantly on text legibility and overall image clarity, resulting in a blurry and unappealing image.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + The text has a clear chalk-like texture.
  • The text is completely illegible and does not follow the specific menu item prompt.
  • The layout is messy and does not resemble a professional cafe menu.
  • Major artifacts and nonsensical symbols appear throughout the image.

FLUX.1 Kontext [pro]

  • + The text follows the prompt with near-perfect spelling and accuracy.
  • + The chalk texture is highly realistic, including the subtle grain on the chalkboard surface.
  • + The composition is clean and looks like a real cafe menu.
  • Includes a minor typo in the supplementary text ('four' instead of 'for').
  • The handwriting is almost too neat, bordering on a digital font look in certain sections.

Verdict: FLUX.1 Kontext [pro] successfully rendered the complex text instructions with high legibility and realistic chalk textures. In contrast, DALL-E 2 failed to produce any readable words and completely ignored the specific item list provided in the prompt.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Captures a cinematic, painterly feel.
  • + Good use of negative space in the background.
  • Failed the core prompt instruction of having the horse on top.
  • Lower resolution with significant noise and artifacting.
  • Anatomical issues with the horse's legs and the astronaut's posture.

FLUX.1 Kontext [pro]

  • + Perfectly adhered to the difficult instruction of placing the horse on top of the astronaut.
  • + High visual clarity and professional-grade rendering.
  • + Creative and surreal touch by giving the astronaut horse hooves.
  • The small extra astronaut figure riding the horse is a bit confusing but adds to the surrealism.

Verdict: DALL-E 2 completely failed the logic test of the prompt, providing a standard astronaut-on-horse image with poor anatomical detail. FLUX.1 Kontext [pro] followed the specific 'horse on top' instruction perfectly, delivering a high-quality, surreal, and cinematic image that precisely matches the user's intent.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Attempts to place the viewer inside the vehicle.
  • Fails completely on anatomical realism for both the human and animal.
  • The creature looks like a melted statue rather than a capybara.
  • The composition is chaotic with severe artifacts and distorted facial features.

FLUX.1 Kontext [pro]

  • + Excellent photorealism with clear, sharp details on the capybara's fur.
  • + Strictly adheres to all prompt instructions including the hat, jacket, and boredom of the passenger.
  • + Sophisticated lighting and depth of field that accurately mimics a night-time city environment.
  • The view is from slightly outside the window rather than fully 'inside' the cabin.

Verdict: DALL-E 2 produced a low-quality, distorted image that fails to render a recognizable capybara or a coherent human face. FLUX.1 Kontext [pro] delivered a high-quality, professional-grade photograph that perfectly captured the requested humor and detailed elements of the prompt with realistic textures.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Captures a very authentic vintage, weathered parchment aesthetic.
  • + Creative, organic border design that feels hand-carved.
  • Text consists of illegible gibberish and misspellings.
  • Fails to include specific requested visual elements like the jack-o-lantern.
  • Low resolution and messy visual artifacts.

FLUX.1 Kontext [pro]

  • + Excellent prompt adherence with nearly all text rendered correctly.
  • + High visual quality with clean, cinematic lighting and sharp details.
  • + Well-composed layout that follows all design instructions.
  • Includes a line of gibberish text ('Your: Vorkleat: Iight & Spans') not requested in the prompt.
  • The 'parchment' texture is less pronounced than requested, leaning more toward a digital poster.

Verdict: FLUX.1 Kontext [pro] is the clear winner as it successfully rendered the specific spooky elements and the majority of the complex text requirements with high clarity. DALL-E 2 produced an atmospheric texture but failed entirely on legibility and specific object placement like the jack-o-lantern.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Strong isometric lighting and shadows.
  • + Bright, high-key color palette.
  • Failed to follow text instructions, rendering 'Sush' instead of the requested phrases.
  • Poor image coherence with strange, non-sushi artifacts on the plate.
  • Missing the flag icon and secondary text.

FLUX.1 Kontext [pro]

  • + Perfect adherence to all text requirements, including 'JAPAN', 'SUSHI', and a flag icon.
  • + Excellent execution of the miniature 3D cartoon style with refined PBR textures.
  • + Clean, balanced composition that follows the diorama base and background prompts.
  • The rice grains look more like rounded pellets than traditionally textured rice.

Verdict: FLUX.1 Kontext [pro] followed every detail of the prompt, including complex text and layout requirements, resulting in a high-quality, professional-looking graphic. DALL-E 2 struggled significantly with text legibility and image coherence, placing unrecognizable objects on the plate and failing to include the primary keywords.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Attempts to show variety in animal movement
  • + Includes the requested butterfly elements
  • Serious anatomical distortions and 'melting' limbs on the kitten
  • Low resolution with heavy artifacts and a painterly rather than photorealistic texture
  • Missing the fox kit and the requested 'god rays' lighting

FLUX.1 Kontext [pro]

  • + Exceptional photorealism with ultra-detailed fur and expressive eyes
  • + Follows all prompt components including the specific list of animals (puppy, kitten, bunny, fox)
  • + Beautiful lighting with clear god rays and a lush wildflower environment
  • The bunny and kitten faces look slightly blended in style
  • Composition is more of a posed portrait rather than an active 'chasing' scene

Verdict: FLUX.1 Kontext [pro] creates a vastly superior image that captures the hyper-photorealistic and wholesome aesthetic requested, whereas DALL-E 2 suffers from significant anatomical errors and poor texture quality. FLUX.1 Kontext [pro] also successfully included all four specific animals, while DALL-E 2 failed to render the fox kit.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Successfully used the requested warm brown and cream color palette.
  • + Includes the requested cloche graphic element.
  • Text is completely garbled and unreadable.
  • Missing the banner component and the date requested in the prompt.
  • Steam effect is poorly rendered and lacks clarity.

FLUX.1 Kontext [pro]

  • + Perfect text rendering of 'Caffè Florian' and 'Est. 1720'.
  • + Excellent adherence to all prompt elements including cloche, steam, and banner.
  • + Great use of subtle paper texture and consistent vector style.
  • Minor spelling error in the banner with an extra 'E' in 'EEST.'.

Verdict: FLUX.1 Kontext [pro] followed the prompt almost perfectly, delivering a professional vector-style logo with clear typography and the requested cloche and banner elements. In contrast, DALL-E 2 failed significantly on the typography and missing several key components of the logo, resulting in nonsensical text.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

DALL-E 2
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 2

  • + Matches the specific color palette requested
  • + Captures the technical, dense look of a poster
  • Text is nonsensical gibberish including the title
  • The layout is chaotic and fails to follow the requested 6-step logical flow
  • Vector lines are messy and inconsistent

FLUX.1 Kontext [pro]

  • + Correctly renders the main title 'APOLLO 11'
  • + Strictly adheres to the flat-vector, clean infographic style requested
  • + Includes clear numbered steps and trajectory paths
  • Includes bizarre planetary rings around the rocket and Moon
  • Some smaller text labels become confused (e.g., '2, far Collins')
  • Anatomical rocket details are more cartoon-like than NASA-accurate

Verdict: FLUX.1 Kontext [pro] successfully follows the aesthetic and structural requirements of the prompt, providing a clear infographic with legible titles and steps. DALL-E 2 fails significantly on prompt adherence, producing garbled text and a visual layout that does not represent the requested mission stages.

Next steps

Explore each model