Black Forest Labs' 12-billion parameter multimodal flow transformer for in-context image generation and editing with character consistency, typography handling, and commercial-ready quality
Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.
FLUX.1 Kontext [pro]
#41 of 62 in Text-to-Image
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Where the votes landed
FLUX.1 Kontext [pro]
0.0%
win rate
Ties
0.0%
GPT Image 1 Mini
100.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent photographic quality with realistic textures on the sphere and table.
- + Accurate spatial relationships, showing the plant through the glass and window light from the left.
- + Clear, sharp details on the book edges and glass frame.
- − The blue sphere appears slightly felt-like rather than a solid smooth material.
GPT Image 1 Mini
- + Good adherence to all prompt elements including the specific layout.
- + Solid lighting and shadows that match the environment.
- − The glass refractive logic is slightly flawed near the bottom edges.
- − The image has a slightly softer, more digital look compared to the crispness of Model A.
Verdict: Both models followed the complex spatial instructions perfectly. FLUX.1 Kontext [pro] is the winner due to its superior photographic realism, particularly in the rendering of light and the fine wood grain of the table. GPT Image 1 Mini is also excellent but appears slightly more like a CGI render than a real photograph.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent depiction of falling rain and wet textures
- + Matches the 'candid' and 'imperfect framing' request with a centered, slightly low-angle perspective
- − The subject appears to be riding or standing with the bike rather than 'repairing' it
- − Hand anatomy on the right handlebar is slightly distorted
GPT Image 1 Mini
- + Successfully depicts the 'repairing' action with a squatting pose and focused hands
- + Highly realistic skin texture and facial lighting
- + Stronger background bokeh and motion blur on cars looking more natural
- − The rain is less visible compared to the other model, feeling more like post-rain dampness
- − The subject's hands are overlapping in a slightly confusing way with the bike spokes
Verdict: GPT Image 1 Mini followed the specific action of 'repairing' much better than FLUX.1 Kontext [pro], which showed the man simply holding the handlebars. While FLUX.1 Kontext [pro] captured the atmosphere of falling rain more vividly, GPT Image 1 Mini took a superior 'candid' approach with better realistic textures and a more convincing 50mm shallow depth of field.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent photorealistic skin texture and lifelike eyes
- + Beautifully detailed engraving on the polished armor plate
- + Superior rendering of cloth and leather strap textures
- − Missed the request for beads in the hair braids
- − The hair and face look a bit too clean for a battle-worn warrior
GPT Image 1 Mini
- + Stronger adherence to the battle-worn theme with significant dirt and scars
- + Excellent atmospheric lighting and bokeh sparks throughout the background
- + Captures the braided hair style with a more rugged look
- − Eyes look slightly less realistic than in the second model
- − The metallic sheen on the armor is somewhat muted compared to the prompt's focus on reflection
Verdict: Both models performed exceptionally well on this complex prompt. FLUX.1 Kontext [pro] produced a much more realistic, high-fidelity human portrait with incredible detail on skin and metal, though it missed the specific detail of beads in the hair. GPT Image 1 Mini captured the 'battle-worn' essence and atmospheric lighting much more effectively, although the overall image has a slightly more painterly feel than FLUX.1's photorealism.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent typography rendering with almost perfectly legible text and numbers.
- + Sophisticated layout that integrates food photography and text in a professional editorial style.
- + Accurate color usage for section headers to create visual distinction.
- − The 'Mains' section lists pizza items, creating a programmatic mismatch in the menu content.
- − The 'Pizza' section features a salad photo instead of a pizza.
GPT Image 1 Mini
- + Strict adherence to the 'grid' requirement for food photos.
- + Clean, minimalist layout that looks like a usable template.
- + High-quality, vibrant food photography that corresponds well to the section categories.
- − Complete lack of menu item text, providing only placeholders for categories.
- − The layout feels a bit generic compared to a professional graphic design.
Verdict: FLUX.1 Kontext [pro] creates a much more realistic and professional-looking menu with impressive text rendering, though the logic of the images versus the section headers is slightly flawed. GPT Image 1 Mini follows the grid layout request more literally and provides better categorization of food photos, but fails to include any actual menu item text beyond the headers. FLUX.1 Kontext [pro] is preferred for its superior graphic design quality and successful rendering of complex text elements.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent photorealistic texture on the bun and melting cheese
- + The glowing fiery text 'MAGIC BURGER' is highly polished and legible
- + Vibrant use of embers and fire adds to the dynamic sense of motion
- − Failed to include the starburst element for the price
- − Repeated the price twice, creating a cluttered bottom of the image
- − The burger layers are slightly compressed rather than 'exploded'
GPT Image 1 Mini
- + Successfully followed all instructions including the starburst shape and secondary text placement
- + Excellent 'exploded' layout with clear separation between every ingredient
- + The fiery glow effect is applied consistently across all text elements
- − The lighting on the burger feels a bit flat compared to the backgrounds intensity
- − The tomato slice looks slightly less realistic/3D than other elements
Verdict: While FLUX.1 Kontext [pro] achieves a higher level of photorealistic detail and more professional typography, GPT Image 1 Mini is the overall winner for prompt adherence. GPT Image 1 Mini correctly included the starburst and secondary messages without redundancy, while also capturing the 'exploded' nature of the burger much more effectively than FLUX.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent chalk texture and natural handwriting style.
- + Accurately rendered all requested menu items and prices.
- + High-resolution chalkboard details like dust and smudge patterns.
- − Typos in the footer text ('ous gluten tree').
- − The title is in print blocks rather than the requested elegant cursive.
GPT Image 1 Mini
- + Perfectly accurate spelling across all text, including fine print.
- + Strong chalk texture that feels authentic to the medium.
- + Better layout and spacing between the menu items.
- − Failed to use 'elegant cursive' for the title as specified.
- − The handwriting looks slightly more uniform than Model A, appearing almost font-like.
Verdict: Both models performed exceptionally well on text rendering and texture. GPT Image 1 Mini is the winner because it maintained 100% spelling accuracy, whereas FLUX.1 Kontext [pro] had significant errors in the small text at the bottom. GPT Image 1 Mini also presented a more professional layout for a menu.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Successfully followed the difficult spatial reasoning prompt of putting the horse on top of the astronaut.
- + Included surreal details like the astronaut having hooves for feet.
- + High level of detail on the space suit and horse textures.
- − The composition is a bit crowded with the additional small astronaut riding the horse.
- − The rendering of the horse's back hooves/legs against the astronaut is slightly confusing.
GPT Image 1 Mini
- + Very clean, cinematic lighting and atmosphere.
- + Good anatomical drawing of the horse.
- − Completely failed the negative constraint/spatial instruction to put the horse on top.
- − Standard cliché interpretation of an astronaut on a horse.
Verdict: FLUX.1 Kontext [pro] is the clear winner as it successfully interpreted the unusual and specific prompt instruction to place the 'horse on top, not vice versa,' while also adding surreal touches like hooves on the astronaut. GPT Image 1 Mini ignored the spatial instruction entirely, providing a standard astronaut riding a horse which failed the core challenge of the prompt.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent photographic lighting and fur texture detail
- + Captures a realistic New York taxi exterior aesthetic
- + The capybara has a very expressive professional appearance
- − The passenger is holding her hands to her ears/face rather than looking at a phone
- − The capybara's paw placement on the wheel is slightly awkward with the perspective
GPT Image 1 Mini
- + Perfect adherence to the passenger's action of looking at a phone
- + Classic taxi driver hat design with checkered pattern
- + Stronger interior perspective that frames both characters clearly
- − The capybara's hand/paw looks a bit more like a human hand covered in fur than a natural paw
- − Slightly dimmer lighting compared to Model A
Verdict: GPT Image 1 Mini followed the prompt instructions more accurately by including the specific detail of the passenger looking at her phone, whereas the passenger in FLUX.1 Kontext [pro] appears to be talking on a phone or covering her ears. However, FLUX.1 Kontext [pro] produced a more vibrant and photographically sharp image with superior texture quality.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent typography with a beautiful gothic font style.
- + Superior lighting and contrast with a glowing jack-o-lantern and bright moon.
- + Very detailed border featuring intricate cobwebs as requested.
- − Includes a line of gibberish text ('Your: Vorkleat: Iight & Spans') not in the prompt.
- − Used a comma instead of a period in the date ('30.10,2026').
GPT Image 1 Mini
- + Accurately followed all text instructions without adding extra hallucinated words.
- + Atmospheric vintage texture that feels like dark parchment.
- + Good adherence to the border requirement with thorns and webs visible.
- − The jack-o-lantern and background lack the 'cinematic lighting' punch seen in the competitor.
- − The gothic font for the title is quite basic compared to the requested 'elegant gothic' style.
Verdict: FLUX.1 Kontext [pro] produces a much more visually striking image with superior contrast and high-quality gothic lettering, but it suffers from significant text hallucinations. GPT Image 1 Mini is more reliable for the specific copy requested, accurately placing the text without errors, though the final aesthetic is slightly flatter. FLUX.1 Kontext [pro] is the winner for its professional graphic design quality, despite the minor text artifacts.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent preservation of the source facial structure and skin texture.
- + High-fidelity hair texture that matches the existing beard quality.
- + Perfect retention of the lighting environment and background details.
- − The hairline on the forehead is slightly harsh and sharp.
GPT Image 1 Mini
- + Natural, voluminous hair style with realistic flyaways.
- + Good integration of the hair with the forehead.
- − Significantly alters the person's facial features, making him look noticeably younger.
- − Changes the face shape and ear structure compared to the source image.
- − Slightly softens the skin texture, losing the weathered look of the original.
Verdict: FLUX.1 Kontext [pro] is the superior model for this editing task because it successfully added the hair while keeping the identity of the person perfectly intact. In contrast, GPT Image 1 Mini changed the man's facial features and ear shape, failing to preserve the source identity despite the high quality of the hair itself.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent typography with a playful 3D bubble effect
- + High-quality PBR textures on the salmon and tray
- + Clean, centered composition with a professional diorama feel
- − The rice grains look more like glossy pearls than individual rice grains
- − The flag icon is stylized as a waving banner rather than a standard flag icon
GPT Image 1 Mini
- + Accurate depiction of multiple types of sushi (salmon, tuna, shrimp)
- + Clean and professional graphic design for the text and flag
- + Rice texture is more realistic and faithful to actual sushi
- − Added chopsticks which were not requested in the prompt
- − The diorama base is slightly cropped at the bottom
- − The lighting is a bit flatter compared to the soft shadows in the other model
Verdict: Both models followed the prompt instructions very well, with FLUX.1 Kontext [pro] providing superior 3D text styling and better lighting depth. While GPT Image 1 Mini captured the sushi rice texture and variety better, FLUX.1 Kontext [pro]'s overall composition and 3D miniature aesthetic feel more cohesive as a single scene.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent preservation of the source image's pose, composition, and clothing.
- + Effective comic-book style illustration.
- − Completely failed to incorporate the requested thematic elements like hockey, dogs, or a news anchor setting.
GPT Image 1 Mini
- + Successfully incorporated all requested elements including the news desk, dog, and hockey stick.
- + Strong caricature style with exaggerated features and a colored pencil texture.
- + Good balance of layout and clear storytelling within the image.
- − Loss of the original 'selfie' pose from the source image.
Verdict: FLUX.1 Kontext [pro] successfully converted the subject into a stylized illustration while maintaining the original composition, but it completely ignored the thematic instructions regarding the profession and hobbies. GPT Image 1 Mini followed the prompt perfectly, creating a detailed caricature that includes the dog, hockey equipment, and news anchor environment, making it the clear winner for this specific editing task.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent depiction of warm golden light and god rays
- + Extremely soft and detailed fur rendering
- + Highly expressive and cute facial features
- − The 'bunny' looks like a white cat with slightly longer ears
- − The animals are mostly static and sitting rather than tumbling and chasing
GPT Image 1 Mini
- + Perfectly captures the dynamic action of chasing and tumbling requested
- + Accurately represents all four species, including a distinct baby bunny and tabby kitten
- + Great sense of movement and energy in the composition
- − The fox kit has slightly strange, dark-colored paws
- − Lighting is less dreamy and 'masterpiece' quality compared to Image A
Verdict: While FLUX.1 Kontext [pro] creates a more visually stunning and polished image with superior fur texture, it fails to accurately render the bunny and ignores the dynamic action in the prompt. GPT Image 1 Mini follows the prompt much more accurately, successfully depicting all four specific animals in the middle of playful movement, making it the better interpretation of the text.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Perfectly captures the Studio Ghibli cel-shaded aesthetic with clean line art.
- + Maintains excellent preservation of the source image's layout and clothing details.
- + The facial expressions are highly expressive and culturally accurate to the Ghibli style.
- − The transition from the foreground character to the background is a bit sharp.
GPT Image 1 Mini
- + Features a beautiful soft, hand-drawn texture that feels like colored pencils or pastels.
- + Applying a warm, nostalgic color palette that matches the prompt's request for 'gentle lighting'.
- − The faces look more like generic Western illustrations than Studio Ghibli specifically.
- − Lacks the characteristic clean linework and distinct cel-shading associated with the requested studio.
Verdict: FLUX.1 Kontext [pro] successfully captures the specific recognizable style of Studio Ghibli, from the character design to the cel-shading. While GPT Image 1 Mini creates a lovely atmospheric illustration, it feels more like a general pastel drawing rather than a transformation into the requested anime style. FLUX.1 Kontext [pro] is the winner for its closer adherence to both the source image's structure and the specific artistic target.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent source preservation, maintaining the woman's face and the dog's features accurately.
- + Subtle and realistic hair movement that conforms to the wind direction.
- + Adds flying leaves as requested without cluttering the scene.
GPT Image 1 Mini
- + Successfully adds wind-blown hair and many flying leaves.
- + Maintains the overall composition and color palette of the original.
- − Significant loss of detail and character likeness in both the woman and the dog.
- − The leaves are somewhat oversized and repetitive, looking like a simple overlay.
- − Anatomical glitches appear on the woman's face and the dog's eyes compared to the source.
Verdict: FLUX.1 Kontext [pro] is the clear winner as it successfully applies the requested motion edits while perfectly preserving the identity of the woman and the dog from the source image. In contrast, GPT Image 1 Mini fails to preserve original details, significantly altering the facial features of the subjects and producing a less realistic result.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Excellent typography with proper accent on 'Café'
- + Accurate adherence to the light background and subtle texture request
- + Clean vector-style composition with a classic banner
- − Includes a minor typo 'EEST' instead of 'EST' in the banner
GPT Image 1 Mini
- + Beautifully stylized cloche dome with detailed shading
- + Perfect text accuracy including the established date
- + Elegant gold-on-black color scheme
- − Failed the request for a light background
- − The typography has slight alignment inconsistencies in the letter 'I'
Verdict: FLUX.1 Kontext [pro] followed the stylistic instructions more closely, particularly the request for a light background and subtle texture, creating a very authentic vintage logo feel. GPT Image 1 Mini produced a visually striking image with perfect text spelling, but it completely ignored the instruction for a light background, resulting in a dark theme instead.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.1 Kontext [pro]
- + Captures a cinematic and atmospheric feel with the navy night sky.
- + Features high-resolution textures on the lunar surface and celestial bodies.
- − Fails to follow the numbered step-by-step instructions or icon requirements.
- − Contains nonsensical elements like Saturn's rings around the Moon and rocket.
GPT Image 1 Mini
- + Strictly adheres to all numbered steps and requested iconography.
- + Maintains a consistent and clean flat-vector style with a professional layout.
- − Typography has slight alignment issues.
- − The translunar trajectory line is somewhat abstract and loopy.
Verdict: GPT Image 1 Mini followed the instructional prompt perfectly, delivering a logical infographic sequence with all requested steps and icons. FLUX.1 Kontext [pro] ignored the structural requirements of the prompt, creating a visually confusing scene with factual inaccuracies regarding planetary rings.
Explore each model
OpenAI's cost-effective image generation model for when image quality isn't the top priority