Black Forest Labs' ultra-high resolution image generation model, an enhanced version of FLUX1.1 [pro] optimized for premium quality output
Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.
FLUX1.1 [pro] Ultra
#51 of 62 in Text-to-Image
GPT Image 1.5
#7 of 62 in Text-to-Image
Where the votes landed
FLUX1.1 [pro] Ultra
100.0%
win rate
Ties
0.0%
GPT Image 1.5
0.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Exquisite glass rendering with realistic refraction and edge highlights.
- + Professional-grade photographic quality with sharp focus on the book texture.
- + Accurate interpretation of the soft window light from the left.
- − The sphere is levitating rather than sitting on the bottom of the cube, which feels slightly unnatural.
GPT Image 1.5
- + Natural grounding of the sphere on the base of the cube.
- + Good composition that centers the subject effectively.
- + Clear adherence to all spatial requirements of the prompt.
- − The glass cube looks more like a frame with thick plastic edges rather than a solid glass object.
- − Lower textural detail on the book and plant compared to the competitor.
Verdict: Both models followed the prompt perfectly in terms of object placement. FLUX1.1 [pro] Ultra is the clear winner due to its superior photographic realism, particularly in the rendering of the glass material and the fine textures on the book cover, whereas GPT Image 1.5 feels slightly more like a digital render and lacks the same level of environmental lighting depth.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent execution of long-exposure motion blur on passing vehicles.
- + Captures a very cinematic, high-contrast urban mood.
- + Strict adherence to the 50mm lens and reflections requirement.
- − The subject appears to be just standing with the bike rather than actively 'repairing' it.
- − The bicycle geometry is slightly simplified and unrealistic near the handlebars.
GPT Image 1.5
- + Strong narrative adherence with the subject actively repairing the bike with tools.
- + Excellent skin textures and clothing details that feel authentic and non-stylized.
- + Composition feels like a genuine candid street photo.
- − Failed to incorporate the requested motion blur from passing cars.
- − Includes a strange second bicycle frame overlapping the main red bike.
Verdict: FLUX1.1 [pro] Ultra excels in achieving the technical photographic effects requested, particularly the motion blur and cinematic lighting. However, GPT Image 1.5 provides a more realistic and gritty 'candid' feel with better character posing for the 'repairing' action, despite missing the motion blur and having some minor AI artifacts in the bike's structure.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- − The image failed to load or is a solid black square.
- − Zero prompt adherence as no visual content is present.
GPT Image 1.5
- + Excellent adherence to all prompt details, including the braided hair with beads, scars, and ornate armor.
- + Stunning textural quality on the leather straps, fabric, and weathered metal.
- + Masterful use of warm torchlight and bokeh sparks to create atmosphere.
- − Small anatomical inconsistency where a braid appears to merge into the metal of the pauldron.
Verdict: FLUX1.1 [pro] Ultra failed to generate a visible image, resulting in a black frame. GPT Image 1.5 successfully produced a high-quality, cinematic portrait that perfectly captures the battle-worn aesthetic and specific details requested in the prompt.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Elegant graphic design layout with a professional mockup perspective.
- + High-quality, creative food photography integrated into the design.
- + Good use of bold sans-serif headers as requested.
- − Nonsense text and significant spelling errors (e.g., 'Pizzzans').
- − The grid layout of images is a bit cluttered and overlapping.
GPT Image 1.5
- + Excellent text rendering with clear, readable, and logical menu items.
- + Perfect adherence to the grid layout and section requirements.
- + Very clean, minimalist aesthetic that looks ready for use.
- − The layout is a bit basic and lacks the modern flair seen in Model A.
- − Some food items look slightly generic.
Verdict: While FLUX1.1 [pro] Ultra produces a more stylish and artistic design mockup, GPT Image 1.5 is far superior for this specific task because it renders legible, meaningful text and follows the layout instructions with precision. GPT Image 1.5 successfully translates the prompt into a functional menu design, whereas FLUX1.1 fails significantly on text accuracy and clarity.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent typography rendering with a clean, modern aesthetic.
- + High resolution and very clean texture on the burger ingredients.
- + Includes a realistic charred rock and fire base that adds depth.
- − Missed the 'starburst' requirement for the price tag.
- − The burger stack looks somewhat static and repetitive rather than a single 'exploded' unit.
- − Includes hallucinated watermark-like text at the bottom.
GPT Image 1.5
- + Perfectly follows all prompt instructions including the starburst shape and glowing text effects.
- + Great sense of motion with flying ingredients and dynamic debris.
- + Strong adherence to the 'exploded' request, showing interior textures like the toasted bun.
- − Slightly lower clarity and more 'painterly' artifacts compared to Model A.
- − The 'MAGIC BURGER' title is slightly cropped at the top edges.
Verdict: While FLUX1.1 [pro] Ultra produces a cleaner and more professional-looking image, GPT Image 1.5 is the winner for strict prompt adherence. GPT Image 1.5 correctly included the starburst element and better captured the 'exploded' motion requested, whereas FLUX1.1 [pro] Ultra missed a specific UI element and produced a more traditional stacked burger.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent chalk texture and board realism.
- + Natural, believable handwriting style.
- − Significant text hallucinations and spelling errors like 'mushroomm' and 'frristratnisnios'.
- − Repetitive text at the bottom and fails to complete the prompt's specific menu items correctly.
GPT Image 1.5
- + Perfect adherence to specific text requirements and spelling.
- + Accurately rendered elegant cursive for the title as requested.
- + Clean composition with consistent handwriting style.
- − The chalk texture on the board looks a bit like a digital overlay rather than physical smudges.
- − The lighting is somewhat flat compared to the atmospheric lighting in Model A.
Verdict: GPT Image 1.5 is the clear winner because it followed the text instructions perfectly, whereas FLUX1.1 [pro] Ultra suffered from significant gibberish and spelling errors. While FLUX1.1 captured a more realistic 'messy' chalkboard aesthetic, GPT Image 1.5 successfully balanced elegant cursive for the header with clear, accurate menu items.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent high-resolution clarity and cinematic lighting across the atmosphere.
- + Good execution of the horse and astronaut details.
- + Dynamic composition with a sense of movement in zero-G.
- − Completely failed the negative constraint/spatial instruction 'horse on top'.
- − The inclusion of multiple small Earth-like planets is a bit repetitive.
GPT Image 1.5
- + High level of intricate detail on the astronaut suit and horse harness.
- + Strong cinematic atmosphere with realistic lighting and nebulae.
- + Impressive ground textures and particle effects.
- − Failed the specific spatial instruction 'horse on top'.
- − The anatomy of the horse's front legs and hooves is slightly distorted.
- − Standard interpretation of the prompt rather than a surreal one.
Verdict: Both models failed to follow the difficult spatial constraint 'horse on top, not vice versa,' instead opting for the common 'astronaut riding a horse' trope. FLUX1.1 [pro] Ultra produced a cleaner, more vibrant image with better composition, while GPT Image 1.5 offered superior fine textures but suffered from slight anatomical errors. FLUX1.1 [pro] Ultra is preferred for its overall visual polish and better handling of the cinematic request.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + High resolution and very clean visual quality
- + Accurate rendering of the business coat and phone use
- + Excellent background bokeh and city lighting
- − Composition error: the woman is placed in the front passenger seat instead of the back seat
- − The capybara's paws are not visible on the steering wheel as requested
- − The woman is looking at the capybara's steering wheel rather than her phone
GPT Image 1.5
- + Perfect adherence to composition instructions with the passenger in the back seat
- + Captures the capybara with both front paws on the steering wheel
- + Very realistic 'cinematic' lighting and texture that feels like a real film still
- − Slightly lower resolution compared to Model A
- − Text on the taxi meter in the foreground is blurry
Verdict: GPT Image 1.5 is the clear winner as it correctly followed the spatial instructions to place the business woman in the back seat and the capybara's paws on the wheel. FLUX1.1 [pro] Ultra produced a higher quality image technically, but failed the prompt by placing the passenger in the front seat and omitting the paws on the steering wheel.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent clarity and high resolution.
- + All requested text details are present and legible.
- + Clean graphic design style with professional layout.
- − Includes redundant and misspelled text under the banner ('noickt').
- − Omits the word 'Party' from the main title 'Halloween Invitation'.
- − The aesthetic is somewhat generic and lacks the requested vintage 'dark parchment' texture.
GPT Image 1.5
- + Perfect adherence to the 'dark parchment' and vintage gothic aesthetic.
- + Flawless text rendering for all requested fields including 'Halloween Party Invitation'.
- + Superior composition with a more atmospheric and cinematic night sky.
- − The thorns in the border are slightly repetitive in pattern.
- − Less 'clean' look if the user required a vector-style graphic rather than a painted illustration.
Verdict: GPT Image 1.5 is the clear winner as it perfectly captures the 'vintage gothic' mood, dark parchment texture, and all requested text without errors. FLUX1.1 [pro] Ultra has better technical sharpness but fails on the prompt by including redundant/misspelled text and omitting a word from the primary title.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent soft 3D cartoon aesthetic that matches the 'miniature' request.
- + Clean, professional typography integrated into the background.
- + Soft, diffused lighting and refined PBR textures.
- − Includes some strange artifacts like the red bug-like object and floating steam.
- − The garnish looks like a potted plant rather than culinary decoration.
GPT Image 1.5
- + Stronger adherence to the 'realistic PBR' instruction with believable wood and ceramic textures.
- + Excellent diorama base design with grass and layered textures.
- + Very clear and bold text placement that follows the hierarchy requested.
- − The lighting is a bit harsh compared to the 'gentle' request.
- − Less of a 'cartoon' feel than Model A.
Verdict: Both models followed the prompt instructions very well, including the specific text and icon requirements. FLUX1.1 [pro] Ultra captures a more whimsical, soft cartoon aesthetic, but GPT Image 1.5 succeeds better in providing a high-detail diorama with realistic PBR materials and a more cohesive set of sushi elements. GPT Image 1.5 is the winner for its superior texture work and more logical composition.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent adherence to the '8K masterpiece' aesthetic with clean, polished rendering.
- + Captures the 'golden sunrise' and 'god rays' with high artistic flair.
- + Consistent character design with very expressive, large eyes across all four animals.
- − Looks more like a 3D digital illustration than 'hyper-photorealistic'.
- − The animals are mostly sitting still rather than 'tumbling together' as requested.
GPT Image 1.5
- + Successfully captures a more realistic photographic style with natural lighting.
- + Better adherence to the 'tumbling together' action, showing the kitten on its back and playful poses.
- + Includes visible 'dew sparkles' on the grass as requested in the prompt.
- − The fox kit has a slightly distorted mouth/teeth area.
- − The butterfly's placement looks a bit pasted on/artificial compared to the rest of the scene.
Verdict: GPT Image 1.5 is the winner as it better follows the prompt's request for a hyper-photorealistic style and active movement like 'tumbling together.' While FLUX1.1 [pro] Ultra created a very beautiful and clean image, it leaned too far into a stylized, CG-animation aesthetic rather than realism.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent vector emblem layout with clean circular framing
- + High-quality typography with period-accurate styling
- + Good use of warm brown and cream tones against a light background
- − Misspelled the primary brand name as 'Caffé Framilian'
- − The 'Est. 1720' is not on the banner as requested
- − The cloche is somewhat flattened in perspective
GPT Image 1.5
- + Perfect text rendering for 'Caffè Florian'
- + Correctly placed 'Est. 1720' on a vintage banner
- + Excellent shading and texture on the cloche dome
- − Ignored the 'light background' requirement by using a black background
- − The steam effect is slightly chunky compared to the fine lines of the text
- − Less of a cohesive 'emblem' style compared to Model A
Verdict: GPT Image 1.5 is the winner because it followed the specific text instructions and banner placement perfectly, whereas FLUX1.1 [pro] Ultra failed on the brand name spelling. While FLUX1.1 [pro] Ultra created a more professional circular emblem composition, the disregard for the 'light background' prompt by GPT Image 1.5 is a secondary issue compared to the core brand accuracy.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Elegant vector art style with a modern professional feel
- + Excellent use of vertical space and a cohesive lunar landscape
- + Adheres strictly to the requested NASA-inspired color palette
- − Text is largely nonsensical gibberish
- − The infographic flow is confusing and logically inconsistent
- − Fails to clearly represent the specific 6-step sequence requested
GPT Image 1.5
- + Matches the 6-step sequence exactly as requested
- + All labels are perfectly legible and correctly spelled
- + Iconography is very consistent and clear for each stage
- − The 'Descent' and 'Landing' panels use the same background causing visual repetition
- − Slightly less 'modern' and 'clean' in its vector execution compared to the competitor
- − The red used is quite bright rather than the requested 'muted red'
Verdict: GPT Image 1.5 is the clear winner for its superior prompt adherence and functional infographic design, providing all six requested steps with perfectly legible text. FLUX1.1 [pro] Ultra created a more artistic and visually striking single-image composition, but it failed to follow the sequential instructions and the text is entirely unreadable.
Explore each model
OpenAI's state-of-the-art image generation model with better instruction following and adherence to prompts