Black Forest Labs' ultra-high resolution image generation model, an enhanced version of FLUX1.1 [pro] optimized for premium quality output
Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.
FLUX1.1 [pro] Ultra
#46 of 62 in Text-to-Image
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Where the votes landed
FLUX1.1 [pro] Ultra
0%
win rate
Ties
0%
GPT Image 1 Mini
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent photorealistic textures on the book cover and wooden table.
- + Highly accurate glass refractions showing the plant behind the cube.
- + Crisp lighting that clearly originates from the left as requested.
- − The sphere is floating in the center rather than resting on the bottom, which may feel slightly unnatural.
GPT Image 1 Mini
- + Natural grounding of the blue sphere on the bottom of the cube.
- + Simple and clean composition that follows all prompt instructions.
- − The plant is completely blurred out and doesn't appear 'through' the glass as clearly as the prompt implies.
- − Lower overall texture detail compared to its competitor.
Verdict: FLUX1.1 [pro] Ultra is the clear winner due to its superior rendering of glass and light. While both models followed the spatial instructions, FLUX1.1 [pro] Ultra provided much more detailed textures and more realistic optical refractions of the plant behind the glass cube, whereas GPT Image 1 Mini had a softer, more artificial appearance.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent execution of long-exposure motion blur on passing cars
- + High clarity and realistic lighting on wet pavement
- + Good full-body composition including reflections
- − The man appears to be posing with or moving the bike rather than 'repairing' it
- − Skin texture looks slightly smoothed compared to the 'natural texture' request
GPT Image 1 Mini
- + Outstanding natural skin texture and facial detail
- + Captures the 'repairing' action more accurately with a crouched pose
- + Authentic shallow depth of field and color grading for a 50mm lens feel
- − Missed the 'motion blur from passing cars' instruction as the background cars are static
- − Less emphasis on the wet pavement reflections compared to Image A
Verdict: FLUX1.1 [pro] Ultra followed the complex technical prompt more accurately, successfully incorporating the difficult motion blur and reflection requirements. However, GPT Image 1 Mini produced a much more convincing human subject with superior skin textures and a more authentic 'candid' repair pose, making it feel more like a real photograph despite missing the background motion blur.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- − The image failed to render, resulting in a solid black frame.
- − Total failure to adhere to the prompt.
GPT Image 1 Mini
- + Excellent adherence to all prompt elements including braided hair, scars, and ornate armor.
- + Impressive textural detail on the engraved plate and leather straps.
- + Masterful use of lighting and bokeh to create a cinematic atmosphere.
- − The beads in the hair are subtly integrated and could be more prominent.
Verdict: FLUX1.1 [pro] Ultra failed to generate a visible image, resulting in a completely black output. GPT Image 1 Mini provided a high-quality, atmospheric portrait that followed every detail of the prompt with exceptional clarity and artistic composition.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent professional flyer layout with realistic perspective.
- + Includes complex details such as price points, item descriptions, and branding icons.
- + Sophisticated typography variation that feels like a real restaurant menu.
- − Text contains frequent misspellings like 'Pizzzans' and 'Appet Love'.
- − Layout is slightly cluttered compared to a strict minimalist aesthetic.
GPT Image 1 Mini
- + Very clean and organized 3x2 grid that perfectly matches the prompt.
- + Perfectly legible and bold sans-serif header text.
- + Superior minimalist composition that functions effectively as a simple template.
- − Empty space for the menu items looks like an unfinished template.
- − Less 'professional' in terms of real-world menu depth compared to Model A.
Verdict: FLUX1.1 [pro] Ultra produces a high-fidelity, professional-looking mockup that captures the 'casual dining' atmosphere well, though it suffers from typical AI spelling errors. GPT Image 1 Mini provides a much cleaner, more literal interpretation of the minimalist grid request, functioning better as a design template even though it lacks the detailed content of the former.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Extremely high photorealistic detail on food textures
- + Crisp, modern commercial-grade lighting and resolution
- + Accurate text rendering for all requested strings
- − Missed the starburst element for the price
- − Includes hallucinated watermark-like text at the bottom
GPT Image 1 Mini
- + Perfect adherence to all layout instructions including the starburst
- + Superior fiery, glowing effect on all text elements
- + Captures a distinct 'magical' atmosphere consistent with the prompt
- − Visual resolution and food texture are significantly lower than the competitor
- − Composition feels slightly more static than the prompt's 'dynamic' request
Verdict: FLUX1.1 [pro] Ultra produces a stunning, high-resolution commercial image with incredible food detail, but it fails to include the requested starburst. GPT Image 1 Mini adheres more strictly to the prompt's structural requirements (starburst and specific text effects) but delivers a much lower-quality image that lacks the professional photorealism of FLUX1.1 [pro] Ultra.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Features a more authentic café-style background with depth.
- + Captures the specific imperfections and 'natural variation' requested in the prompt.
- − Significant spelling errors and word repetitions (e.g., 'mushroomm', 'fresh fresh', 'riistratnisnios').
- − Failed to render the title in the requested 'elegant cursive' style.
- − Text is messy and lacks proper spacing for a menu.
GPT Image 1 Mini
- + Excellent text spelling and layout, accurately following all menu items and prices.
- + Superior chalk texture that looks realistically grainy and feathered.
- + Good composition with a clear, readable hierarchy.
- − The font is a bit too uniform, lacking some of the requested 'handwritten' slant and variation.
- − Failed to use 'elegant cursive' for the title as specifically requested.
- − The board feels slightly flatter and less integrated into a 3D environment.
Verdict: GPT Image 1 Mini is the clear winner because it successfully renders all the complex text with perfect spelling and professional layout, whereas FLUX1.1 [pro] Ultra produces incoherent gibberish and numerous repetitions in the body text. While both models failed the 'elegant cursive' title instruction, GPT Image 1 Mini's chalk texture and overall legibility make for a much more usable image.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + High resolution and bright, cinematic lighting against the Earth.
- + Clean and polished textures on the spacesuit and horse's fur.
- + Dynamic composition with a sense of movement.
- − Completely failed the negative constraint/spatial instruction of placing the horse on top of the astronaut.
- − Included strange artifacts like the orange object on top of the astronaut's backpack.
GPT Image 1 Mini
- + Good atmospheric depth and starry background texture.
- + The horse's anatomy is reasonably well-rendered for a surreal concept.
- − Failed the specific prompt instruction to have the horse on top of the astronaut.
- − The image is somewhat dark and lacks the requested high-detail cinematic sharpness seen in the competitor.
Verdict: Both FLUX1.1 [pro] Ultra and GPT Image 1 Mini failed the specific spatial logic requested in the prompt ('horse on top, not vice versa'), instead providing the standard astronaut-on-horse interpretation. FLUX1.1 [pro] Ultra is the preferred choice for its superior image quality, vibrant colors, and sharp details, whereas GPT Image 1 Mini feels more muted and traditional.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + High resolution with vibrant colors and sharp details
- + Excellent rendering of capybara fur and whiskers
- − Failed the spatial logical request, placing the passenger in the front seat holding a steering wheel
- − Incorrectly depicts the businesswoman as a second driver or passenger with a wheel
GPT Image 1 Mini
- + Correctly follows the spatial layout, placing the passenger in the back and the capybara in the driver's seat
- + Accurately captures the 'bored' expression of the businesswoman
- + Maintains a moody, cinematic lighting that fits the nighttime city theme
- − Image is slightly grainier/noisier compared to Model A
- − The capybara's paw on the wheel looks slightly human-like in structure
Verdict: While FLUX1.1 [pro] Ultra produces a cleaner and more detailed image, it fails completely on the core logic of the prompt by placing the businesswoman in the front seat with her own steering wheel. GPT Image 1 Mini correctly interprets the scene structure, showing the capybara driving with the passenger in the back seat as requested, making it much more successful despite slightly lower technical clarity.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent adherence to the 'parchment scroll' and 'thorns' aesthetic
- + Very crisp typography with high legibility
- + Clean composition with distinct background and foreground elements
- − Includes hallucinatory/repetitive text below the scroll
- − The title text misses the word 'Party' as requested in the prompt
GPT Image 1 Mini
- + Perfect text accuracy including the presence of the word 'Party'
- + Superior artistic mood with a gritty, vintage texture
- + Centered composition feels more cohesive and cinematic
- − The border of webs and thorns is very subtle and hard to see
- − Smaller text on the banner and location details is slightly less legible than it could be
Verdict: GPT Image 1 Mini is the winner because it followed the text prompt more accurately, specifically including the word 'Party' and avoiding the redundant/garbled text found in FLUX1.1 [pro] Ultra. While GPT Image 1 Mini is moodier and more cohesive, FLUX1.1 [pro] Ultra has better clarity but failed the detail check by repeating a garbled version of the banner text.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent 3D rendering with sophisticated lighting and PBR material effects.
- + Beautifully detailed diorama with a comprehensive miniature scene.
- + Correct text placement and inclusion of all requested elements.
- − Includes a strange, non-requested ladybug-style artifact on the base.
- − The 'JAPAN' text is slightly warped along a curve rather than being strictly flat/bold.
GPT Image 1 Mini
- + Highly accurate adherence to the 'minimal' part of the prompt.
- + Perfectly clean and bold typography that matches the request exactly.
- + Consistent matte 'cartoon' texture across the entire model.
- − Simple composition lacks the visual interest of a 'miniature 3D scene'.
- − The diorama base is very basic compared to the requested 'miniature 3D scene' descriptor.
Verdict: FLUX1.1 [pro] Ultra produced a much more visually impressive 3D scene with realistic lighting and materials, capturing the 'miniature' aesthetic perfectly. While GPT Image 1 Mini followed the typography instructions with slightly more precision, FLUX1.1 [pro] Ultra is the preferred choice for its superior artistic execution and depth.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent soft lighting and atmospheric god rays
- + Cohesive artistic style across all four animals
- + Highly detailed fur texture and expressive eyes
- − Static composition that lacks the 'chasing and tumbling' movement requested
- − The kitten looks more like a stylized plush toy than a real tabby
GPT Image 1 Mini
- + Perfectly captures the 'chasing and tumbling' action in the prompt
- + More realistic animal anatomy and fur patterns
- + Accurate depiction of a tabby kitten as requested
- − The fox's front right leg has anatomical issues/blending with the body
- − Background bokeh is slightly less creamy than Image A
Verdict: GPT Image 1 Mini is the superior choice because it captures the dynamic action of the animals chasing and tumbling, whereas FLUX1.1 [pro] Ultra created a static, posed portrait. While FLUX1.1 [pro] Ultra has more beautiful lighting, GPT Image 1 Mini adhered better to the specific animal types and the overall energetic vibe of the prompt.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent vector emblem style with professional line work
- + Beautiful color palette of warm browns and creams on a light background
- + Sophisticated vintage typography and composition
- − Failed to spell the primary brand name correctly, rendering it as 'Framilian' instead of 'Florian'
GPT Image 1 Mini
- + Perfect text accuracy for 'Caffè Florian' and 'Est. 1720'
- + Clean, minimalist interpretation of the cloche and steam
- + Good use of texture for a vintage effect
- − Ignored the request for a 'light background', opting for solid black
- − Misses the 'banner' requirement for the 'Est. 1720' text, using a flat graphic instead
Verdict: GPT Image 1 Mini is the preferred choice because it successfully renders the brand name 'Caffè Florian' correctly, which is a critical failure in FLUX1.1 [pro] Ultra. While FLUX1.1 [pro] Ultra produced a more aesthetically complex and professional-looking vector emblem, the misspelling of the subject name makes it unusable for the specific request.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX1.1 [pro] Ultra
- + Excellent visual depth and complex vector artistic style.
- + Captures the specified color palette perfectly.
- + Includes a large amount of supporting graphical elements that look professional.
- − The numerical sequence of the steps is illogical and confusing.
- − Text is largely gibberish beyond the main headers.
- − The layout is cluttered and difficult to read as a functional infographic.
GPT Image 1 Mini
- + Perfect adherence to the 6 requested steps in the correct order.
- + Clean, legible text for all major headers.
- + Strictly follows the 'crisp lines' and 'flat-vector' style requested in the prompt.
- − The composition is a bit basic and lacks the 'poster' feel of the other model.
- − Iconography for 'Translunar' is a bit abstract compared to others.
Verdict: While FLUX1.1 [pro] Ultra produces a more visually stunning piece of art, GPT Image 1 Mini is the clear winner for prompt adherence and utility. GPT Image 1 Mini correctly followed the 6-step sequence and provided legible text, whereas FLUX1.1 [pro] Ultra failed to order the steps correctly and filled the page with nonsensical text.
Explore each model
OpenAI's cost-effective image generation model for when image quality isn't the top priority