Head to head
Esc

Models · slot A

to navigate to pick

DALL-E 3 OpenAI FLUX.1 Kontext [pro] Black Forest Labs

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

DALL-E 3

20.5 arena score

#38 of 62 in Text-to-Image

Skill signature · Text-to-Image

FLUX.1 Kontext [pro]

20.3 arena score

#41 of 62 in Text-to-Image

Vote tally

Where the votes landed

DALL-E 3

0%

win rate

Ties

0%

FLUX.1 Kontext [pro]

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + High visual interest and artistic rendering of the wooden frame
  • + Excellent texture and lighting detail on the interior book and sphere
  • Failed spatial prompt adherence: the book is inside the cube rather than on top of it
  • The sphere is sitting on the book instead of being a simple standalone element

FLUX.1 Kontext [pro]

  • + Perfect adherence to all spatial instructions in the prompt
  • + Highly realistic textures on the wood table and the red book
  • + Correct lighting direction from the left window
  • The blue sphere appears slightly fuzzy/felt-like rather than a smooth glass or plastic texture

Verdict: FLUX.1 Kontext [pro] followed every spatial instruction perfectly, correctly placing the red book on top of the cube and the sphere inside. DALL-E 3 failed the prompt adherence by placing both the book and the sphere inside a wooden-framed structure, ignoring the requested layout despite its high artistic quality.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Strong composition with a creative foreground element and clear floor reflections.
  • + Captures the 'repairing' aspect of the prompt well by showing the man interacting with the chain/wheel area.
  • + Effective use of atmospheric lighting and lanterns to create a cinematic feel.
  • The man appears overly gaunt and barefoot, which feels somewhat stylized despite the 'no stylization' instruction.
  • The motion blur on the passing car is quite minimal, appearing more as a static parked car.

FLUX.1 Kontext [pro]

  • + High degree of realism in skin texture and clothing fabrics.
  • + Excellent rain effect with visible streaks that blend naturally into the scene.
  • + Effective 'imperfect framing' that makes it feel like a genuine street photograph.
  • The man appears to be pushing or riding the bicycle rather than 'repairing' it as requested.
  • The red color of the bicycle is a bit muted compared to the prompt's focus.

Verdict: DALL-E 3 followed the 'repairing' action more accurately and created a more visually interesting composition with reflections, but it suffers from a slightly 'CGI' look. FLUX.1 Kontext [pro] achieved a much higher level of photographic realism and skin texture, though it failed to properly depict the specific action of repairing the bike.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent depiction of textured leather and metal engraving
  • + Dramatic use of warm lighting and high-contrast shadows
  • + Detailed facial textures including skin pores and lifelike eyes
  • Failed to include the specific 'hair braided with small beads' instruction
  • The scars look slightly stylized rather than realistic

FLUX.1 Kontext [pro]

  • + Accurately followed the braiding and bead instruction
  • + Features a more realistic and subtle 'battle-worn' appearance with dirt and faint scars
  • + Excellent depth of field and bokeh spark effects in the background
  • The lighting is a bit flat compared to the 'warm torchlight' request
  • The engraving on the armor is somewhat less intricate than Model A

Verdict: FLUX.1 Kontext [pro] is the superior model here as it adhered to all prompt details, specifically the hair braids with beads, which DALL-E 3 ignored. While DALL-E 3 produced a more traditionally 'epic' and high-contrast image, FLUX.1 provided a more realistic interpretation of a battle-worn character with consistent lighting and texture.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Strong variety of food imagery throughout four different layouts.
  • + Excellent use of a grid-based design as requested.
  • + Effective use of vibrant color blocks and accents.
  • Text is largely illegible gibberish or symbols.
  • The presentation as four separate tiles makes it feel more like a mood board than a single usable design.

FLUX.1 Kontext [pro]

  • + Text rendering is very clean and readable, despite some minor spelling errors.
  • + Perfectly adheres to the request for specific sections (Appetizers/Pizza/Mains).
  • + Excellent use of bold sans-serif fonts and clean minimalist spacing.
  • The food imagery layout is a bit scattered rather than a strict grid.
  • Some strange pricing numbers (e.g., 243 for a pizza).

Verdict: FLUX.1 Kontext [pro] is the clear winner for its superior text rendering and adherence to the specific menu sections requested. While DALL-E 3 provides more complex visual grids, it fails to produce legible text or a coherent single-page menu format.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent 'exploded' effect with clearly separated, dynamic layers
  • + Stunning photorealistic textures on the charred patty and melting cheese
  • + Atmospheric lighting with flying embers and a grounded impact zone
  • Multiple spelling errors in the text, including 'MAGIC BURGR' and 'Limiited'
  • Missing the requested 'starburst' for the price tag

FLUX.1 Kontext [pro]

  • + Perfect text rendering for all requested phrases including the currency symbol
  • + Effective fiery, glowing effect applied to the typography
  • + High-quality burger rendering with appetizing sauce drips
  • The burger is barely separated into layers, failing the 'exploded' aspect of the prompt
  • Repeated the price tag twice, creating a cluttered bottom composition
  • Failed to place the price in a starburst

Verdict: DALL-E 3 followed the 'exploded burger' layout much better, resulting in a more dynamic and interesting advertisement, but it struggled significantly with text spelling. FLUX.1 Kontext [pro] provided perfect typography and glowing effects but failed to separate the burger components as requested, making it look like a standard floating burger rather than an exploded one. DALL-E 3 is the better visual interpretation of the concept despite the typos.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent chalk texture and artistic flair
  • + Warm, atmospheric lighting that enhances the cafe aesthetic
  • Significant spelling errors throughout the text
  • Failed to follow the specific menu items and prices requested
  • Text becomes illegible gibberish in several sections

FLUX.1 Kontext [pro]

  • + Perfect text accuracy and spelling for the requested prompt
  • + Highly realistic chalk handwriting style that looks hand-drawn
  • + Clear, organized layout that follows all prompt instructions
  • Slightly plain background and lighting compared to the artistic board in Image A
  • Minor typo in the final footer line ('foiur')

Verdict: FLUX.1 Kontext [pro] is the clear winner as it successfully rendered all requested menu items with near-perfect spelling and a very convincing chalk texture. DALL-E 3 produced a more visually striking and atmospheric scene, but it completely failed the text-to-image challenge by generating gibberish and incorrect prices.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Beautiful cinematic lighting and cosmic backdrop
  • + Coherent composition with sharp details in the astronaut suit
  • Failed the core prompt instruction to have the horse on top
  • Common interpretation of the prompt without the requested surreal reversal

FLUX.1 Kontext [pro]

  • + Perfectly followed the difficult instruction of placing the horse on top of the astronaut
  • + Highly surreal and creative interpretation of the prompt
  • + Excellent physical details on the horse's fur and the suit textures
  • Anatomical glitch where a smaller astronaut appears to be part of the horse's saddle/neck area
  • The background stars and planets are slightly less vibrant than Model A

Verdict: DALL-E 3 produced a high-quality but generic image that completely ignored the specific spatial instruction 'horse on top'. FLUX.1 Kontext [pro] successfully executed the surreal prompt reversal, placing the horse on the astronaut's back, making it the clear winner for prompt adherence despite a minor anatomical artifact.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent fur detail and rim lighting on the capybara
  • + High-quality, atmospheric New York City environment
  • + Consistent 'taxi' theme with interior dashboard details
  • Failed to include the human passenger in the back seat
  • The capybara's paw/hand looks slightly morphed into the steering wheel

FLUX.1 Kontext [pro]

  • + Successfully included the human businesswoman in the back seat
  • + Highly realistic lighting and photographic texture
  • + Accurately depicted both front paws on the steering wheel
  • The passenger is on the phone rather than just looking at it, and her hands have anatomical issues
  • The yellow cap is slightly less 'professional' looking than the one in Model A

Verdict: FLUX.1 Kontext [pro] is the clear winner because it followed the complex instruction to include a bored human passenger in the back seat, which DALL-E 3 omitted entirely. While DALL-E 3 produced a more vibrant and stylized city background, FLUX.1 Kontext [pro] achieved a much higher level of photorealism and better adherence to the specific composition requested.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent intricate gothic framing with high visual depth
  • + Very atmospheric lighting and texture on the parchment
  • Text is largely illegible and contains gibberish
  • Fails to accurately present the requested event details

FLUX.1 Kontext [pro]

  • + Near-perfect text rendering for the title and date
  • + Accurately follows the banner and location details
  • + Clear composition with a dominant glowing jack-o-lantern
  • Includes a line of gibberish text in the middle of valid details
  • The border is less 'thorns and webs' and more of a simple web pattern

Verdict: While DALL-E 3 produces a more complex and artistically impressive gothic border, it fails significantly on the text requirements, resulting in illegible characters. FLUX.1 Kontext [pro] captures the prompt's practical needs much better, delivering clean, readable text and a clear layout that functions as an actual invitation, despite a minor line of hallucinated text.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent 3D miniature diorama feel with high attention to detail in the textures
  • + Follows the isometric layout and diorama base request perfectly
  • + Vibrant colors and high visual clarity
  • Failed to place the text 'JAPAN' and 'SUSHI' at the top-center as requested
  • Rendered the text as part of the 3D base rather than a separate graphic element

FLUX.1 Kontext [pro]

  • + Followed text placement instructions perfectly, including the flag icon and hierarchy
  • + Clean, minimalist aesthetic that matches the 'cartoon' and 'soft refined texture' prompt
  • + Accurate 45 degree isometric perspective
  • The diorama base is a bit simple compared to the detailed scene in the other version
  • The rice texture looks slightly like generic spheres rather than traditional rice

Verdict: FLUX.1 Kontext [pro] is the winner because it adhered to all layout instructions, particularly the specific placement and content of the text at the top-center. While DALL-E 3 created a visually denser and more complex 3D model, it ignored the text hierarchy and placement prompts entirely, embedding the text into the base instead.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent depiction of god rays and sunrise lighting
  • + Very expressive, large eyes as requested
  • + Detailed fur textures and sparkling dew effects
  • Has a distinct digital/CGI illustrative feel rather than photorealistic
  • The butterflies have strange furry faces, which is anatomically bizarre

FLUX.1 Kontext [pro]

  • + Much closer to the 'hyper-photorealistic' part of the prompt
  • + Captures the 'tumbling together' action more naturally
  • + Better bokeh and professional photography feel
  • The fox and cat look very similar in facial structure
  • The lighting is soft but lacks the dramatic 'god rays' seen in the other image

Verdict: DALL-E 3 creates a charming, storybook-like illustration with dramatic lighting but fails the 'photorealistic' requirement and produces strange hybrid butterfly-creatures. FLUX.1 Kontext [pro] delivers a much more realistic image that looks like actual wildlife photography, better satisfying the core aesthetics of the prompt.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Excellent vector emblem style with intricate detailing
  • + Follows the color palette perfectly with warm brown and cream tones
  • + High-quality stippling texture and clear typography for the date
  • Failed to include the specific name 'Caffè Florian', substituting it with 'Coffee House'
  • The composition is quite dense for a 'minimalist' request

FLUX.1 Kontext [pro]

  • + Perfect adherence to the requested text 'Caffè Florian'
  • + Accurately represents the minimalist and vintage aesthetic requested
  • + Clean layout that is highly suitable for a restaurant logo
  • Includes a minor spelling error in 'EEST.' instead of 'EST.'
  • The steam effect is very simple compared to the more artistic rendering in the other model

Verdict: While DALL-E 3 produced a more visually sophisticated emblem, it failed the primary prompt instruction to use the name 'Caffè Florian'. FLUX.1 Kontext [pro] followed the text requirements and the minimalist style far more accurately, despite a small typo in the abbreviation 'Est.'.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

DALL-E 3
FLUX.1 Kontext [pro]

AI Judge Analysis

DALL-E 3

  • + Successfully captures a sophisticated Retro-Futuristic NASA aesthetic
  • + Excellent use of the specified color palette
  • + Contains complex, organized infographic layout structures
  • Text is purely illegible gibberish
  • The rocket icons are stylistically closer to Space Shuttles than the Saturn V
  • Fails to clearly delineate the requested 6 specific steps in order

FLUX.1 Kontext [pro]

  • + Legible and accurate primary title text
  • + Adheres better to the flat-vector style requested
  • + Follows the sequence of steps more logically with labels
  • Includes a nonsensical Saturn-like ring around the Moon and Saturn V's tip
  • The Earth is labeled as 'Step 4' while the Moon is not, creating factual confusion
  • Composition feels less 'clean' and 'professional' with cluttered overlapping lines

Verdict: DALL-E 3 produces a much more visually stunning and 'NASA-inspired' piece of art, but it fails as an actual infographic because it ignores the specific numerical steps and uses placeholder text. FLUX.1 Kontext [pro] creates a more functional diagram and is able to render legible text, but it suffers from strange logical errors like placing rings around the Moon. DALL-E 3 is preferred for its superior aesthetic quality and consistent vector style.

Next steps

Explore each model