Black Forest Labs' state-of-the-art image generation model with maximum quality and speed, supporting text-to-image and multi-reference image editing with up to 4MP output
Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.
FLUX.2 [pro]
#8 of 62 in Text-to-Image
GPT Image 1.5
#7 of 62 in Text-to-Image
Where the votes landed
FLUX.2 [pro]
42.9%
win rate
Ties
9.5%
GPT Image 1.5
47.6%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.2 [pro]
- + Perfect adherence to the positional and spatial requirements of the prompt.
- + Superior photorealistic quality with natural depth of field and soft lighting.
- + Accurate glass physics, including subtle refractions of the table and background.
- − The sphere is slightly matte rather than gloss (minor preference).
GPT Image 1.5
- + Excellent color vibrance and texture on the red book.
- + Good clarity and sharp focus on the central objects.
- − The bottom of the cube appears to be a mirror rather than clear glass as requested.
- − The sphere reflection inside the base is physically inconsistent with the glass material.
- − Objects appear slightly more 'rendered' and less photographic than Model A.
Verdict: FLUX.2 [pro] followed the prompt more accurately, particularly in the rendering of the glass cube's transparency and the relationship between the objects. While GPT Image 1.5 produced a vibrant image, it turned the bottom of the cube into a mirror, whereas FLUX.2 [pro] maintained a realistic, consistent glass appearance throughout the scene.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent preservation of the car's exterior design and details.
- + Seamlessly integrates the man into the driver's seat from the source photo.
- + Captures a convincing motion blur on the wheels and road.
- − The man's hair is slightly simplified compared to the source image.
- − The lighting on the man is a bit flat compared to the bright coastal surroundings.
GPT Image 1.5
- + Strong adherence to the California coastline setting with palm trees and cliffs.
- + Maintains the man's scarf and clothing textures from the source image.
- + Good high-angle composition.
- − Significant distortion of the car's interior and dashboard architecture.
- − The man's hand on the steering wheel has anatomical issues (merged fingers and awkward grip).
- − Loss of the iconic Rolls-Royce Spirit of Ecstasy hood ornament and front-end presence.
Verdict: FLUX.2 [pro] is the clear winner as it successfully combines both source images into a single, coherent scene while maintaining the integrity of the car's design. GPT Image 1.5 struggles with anatomical details in the hands and heavily alters the car's interior, making it look like a generic convertible rather than the specific model provided.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent photographic quality with realistic skin texture and lighting.
- + The bicycle details (chain, spokes, derailleur) are highly coherent and physically plausible.
- + Subtle and realistic inclusion of rain droplets on the bike and clothing.
- − The motion blur on the background car is very subtle, making it look almost static.
- − The framing feels a bit too clean and professional for the 'imperfect framing' request.
GPT Image 1.5
- + Strong adherence to the 'imperfect framing' and 'candid' aspects of the prompt.
- + The motion blur on the passing vehicle is more pronounced and dynamic.
- + Great atmosphere with visible rain and a tool tray adding to the storytelling.
- − Significant anatomical issues with the man's hands, which appear mangled and merged with the tire.
- − The bicycle's rear structure is physically impossible, with spokes disappearing and a strange wheel hub configuration.
- − Overall image has a slightly grainier, more processed look compared to the requested 50mm lens clarity.
Verdict: FLUX.2 [pro] produces a much higher quality image with superior technical accuracy in terms of human anatomy and mechanical bicycle parts. While GPT Image 1.5 captures the 'candid' and 'motion blur' energy of the prompt more effectively, its failure to render hands and the bike wheel correctly makes it less realistic overall.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent depiction of colored beads in the braids
- + The engraving on the armor is deep and crisp
- + Strong lifelike quality in the skin texture and eyes
- − The scars look a bit like surface paint rather than physical tissue damage
- − The lighting on the leather straps is slightly uniform
GPT Image 1.5
- + Superior battle-worn aesthetic with realistic skin moisture and grit
- + Armor engraving and cross detail are highly intricate and feel aged
- + Excellent depth of field with better integration of bokeh sparks
- − The beads in the hair are less prominent and colorful compared to Model A
- − Slightly less 'paladin' feel compared to the heavy plate look of Sample A
Verdict: Both models followed the prompt exceptionally well, but GPT Image 1.5 wins due to its more realistic 'battle-worn' texture, featuring subtle skin moisture and integrated grit that feels authentic. While FLUX.2 [pro] did a better job with the specific request for beads in the hair, GPT Image 1.5 offered a more cinematic composition with superior lighting and material interaction.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent photographic quality and lighting in the food images.
- + Sophisticated use of typography and icons.
- + Clean, professional presentation as a physical menu mockup.
- − Significant text errors including 'MINS' instead of 'MAINS' and repetitive menu items under different headings.
- − Food photos do not match the labels (e.g., steak labeled as pizza).
- − The pricing is confusingly formatted with commas and unrealistic values.
GPT Image 1.5
- + Excellent adherence to all prompt elements including categories and grid layout.
- + Perfect text rendering for both titles and item descriptions.
- + High logical consistency between food images and their corresponding labels.
- − Composition feels a bit more like a web UI than a physical printed menu.
- − Slightly less 'premium' feel in the food photography compared to the competitor.
Verdict: While FLUX.2 [pro] produces very high-quality aesthetic visuals and a professional mockup feel, it fails significantly on the actual content logic, mislabeling food items and misspelling the 'Mains' category. GPT Image 1.5 is the superior choice because it provides a perfectly functional, readable menu with accurate text, appropriate pricing, and high logical consistency between the grid photos and the menu sections.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.2 [pro]
- + Extremely clean and legible text rendering with a believable glow effect.
- + Superior photographic texture on the burger bun and melting cheese.
- + Elegant composition with a clear sense of motion through swirling light trails.
- − The background is slightly more static compared to the explosive chaos of model b.
- − The '€6.99' starburst is a bit flat and graphic rather than integrated into the physical scene's lighting.
GPT Image 1.5
- + Highly dynamic 'exploded' effect with debris and high-intensity fiery background.
- + The starburst element features better integration with the fiery theme of the overall image.
- + Very appetizing texture on the charred meat patty and toasted bun.
- − The 'MAGIC BURGER' text is warped and partially cut off at the top edge.
- − The composition feels a bit cluttered, making the individual ingredients harder to distinguish.
- − Less realistic lettuce rendering compared to Model A.
Verdict: FLUX.2 [pro] is the winner due to its superior text rendering, clean composition, and photorealistic textures that make for a more professional-looking advertisement. While GPT Image 1.5 captures the 'exploded' and 'fiery' aspects with more raw intensity, it fails on basic graphic design principles like text placement and legibility.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent chalk texture with realistic dust and smudging on the slate.
- + Natural variation in letter sizing and line spacing consistent with real handwriting.
- + Perfect adherence to the requested text and pricing.
- − The 'slate' board looks a bit like a piece of floor tile rather than a standard café wall chalkboard.
GPT Image 1.5
- + Classic chalkboard aesthetic with authentic-looking eraser marks and chalk dust.
- + Very clean and readable cursive handwriting that looks professional yet hand-drawn.
- + Successfully completed the truncated prompt for the final menu item.
- − Text lacks the slight 'grain' or 'grit' of real chalk compared to Model A.
- − The underlining of the title is a bit too perfectly straight for a hand-drawn line.
Verdict: Both models followed the prompt successfully, including the specific date and pricing. FLUX.2 [pro] is more realistic in its rendering of the chalk texture and the physicality of the board, whereas GPT Image 1.5 feels slightly more like a digital font overlaying a chalkboard background, though it is still very high quality.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent character consistency including face, glasses, and outfit details.
- + Accurately recreates the lighting and yellow background of the source.
- + High visual quality with realistic textures on the scarf and skin.
- − Fails the complex leg-crossing pose from Image 1, opting for a generic crouch.
- − The feet are merged/unnatural at the contact point with the box.
GPT Image 1.5
- + Successfully replicates the specific crossed-leg pose from the reference image.
- + Maintains very good character and clothing consistency from Image 2.
- + Captures the lean and balance of the source pose more accurately than the competitor.
- − Anatomy of the feet is slightly distorted with extra toes or blurred details.
- − A small artifact or hand fragment is visible at the very top right edge of the frame.
Verdict: While both models did an excellent job preserving the identity and clothing of the character from Image 2, GPT Image 1.5 is the clear winner because it actually followed the instruction to replicate the 'exact dynamic pose' from Image 1. FLUX.2 [pro] ignored the complex leg positioning in favor of a standard squat, whereas GPT Image 1.5 successfully managed the difficult task of crossing the legs while standing on the red ottoman.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent adherence to the 'horse on top' spatial instruction
- + High-quality, cinematic rendering with clean details
- + Surreal and creative interpretation of the prompt
- − Anatomical glitch where the horse appears to be merging with or emerging from the astronaut
- − The scale between the astronaut and horse is slightly inconsistent
GPT Image 1.5
- + Strong visual quality and detailed textures in the space suit and lunar surface
- + Dynamic action pose with dust and motion fragments
- − Failed the primary spatial instruction of 'horse on top'
- − Produced a standard 'astronaut riding a horse' cliché despite the specific prompt text
- − Crowded composition with many competing elements
Verdict: FLUX.2 [pro] followed the complex spatial instruction to put the horse on top of the astronaut, whereas GPT Image 1.5 defaulted to the common trope of an astronaut riding a horse. While FLUX.2 [pro] has some anatomical clipping issues, it is the clear winner for successfully interpreting a difficult and surreal prompt.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent side profile that captures the scale of the animal in the seat.
- + High photographic realism with convincing lighting from the dashboard and street.
- + The capybara's leather jacket looks very detailed and tactile.
- − The hands on the steering wheel look more like human hands with textured skin rather than capybara paws.
- − The 'professional expression' is less evident in a profile view compared to a front-on view.
GPT Image 1.5
- + Perfectly captures the 'both front paws on the steering wheel' instruction with anatomically appropriate paws.
- + The taxi driver cap is more detailed and recognizable with the taxi badge.
- + Excellent framing that emphasizes the interaction between the driver and the passenger in the back.
- − The lighting in the cabin is a bit flat compared to the atmospheric shadows in Model A.
- − Slightly less 'photorealistic' in the textures of the background bokeh.
Verdict: Both models followed the complex prompt exceptionally well. FLUX.2 [pro] has better cinematic lighting and a more realistic environment, but GPT Image 1.5 wins on specific detail adherence, particularly with the capybara's paws on the wheel and the clear 'bored' expression of the passenger. GPT Image 1.5's composition feels more like a passenger's POV, which fits the scene better.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent text legibility and alignment
- + High cinematic quality with smooth lighting transitions
- + Clean and professional graphic design layout
- − The parchment texture is subtle and feels more like a modern digital poster than vintage paper
- − The border is a bit minimalist compared to the prompt's request for thorns and webs
GPT Image 1.5
- + Strong 'vintage gothic' aesthetic with a highly detailed parchment texture
- + Excellent border execution with prominent thorns and sprawling webs
- + Good atmospheric depth with the added graveyard and castle silhouettes
- − Text rendering is slightly inconsistent, with a stray line after 'frights'
- − The lighting is a bit flatter and more sepia-toned compared to the cinematic prompt
- − Alignment of the text 'Location' is slightly off-center
Verdict: Both models followed the prompt exceptionally well, particularly regarding the complex text requirements. FLUX.2 [pro] produced a cleaner, more readable invitation that looks ready for print, while GPT Image 1.5 captured the 'vintage gothic' and 'parchment' aesthetic much more effectively through its textured and detailed border. FLUX.2 [pro] is the likely winner for its superior polish and perfect text rendering.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
FLUX.2 [pro]
- + Successfully added thick, voluminous hair
- + Excellent blending of hair into the sideburns
- + Maintains high resolution and realistic hair texture
- − Significantly altered the subject's facial structure, especially around the eyes and brow
- − The hair color doesn't perfectly match the existing beard color
GPT Image 1.5
- + Perfect preservation of the original facial features and bone structure
- + Hair texture and color match the beard and overall aesthetic of the source image much better
- + Realistic hairline and blending with the forehead
- − The hair volume is slightly more modest compared to Model A's interpretation of 'full and thick'
Verdict: While FLUX.2 [pro] provided a very thick head of hair, it failed the preservation requirement by significantly altering the person's face. GPT Image 1.5 successfully added a natural, realistic head of hair that perfectly matches the existing beard while keeping the identity of the person in the source image identical.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent adherence to the 'cartoon scene' and 'minimal' style requested
- + Clean, high-quality rendering of bold text and the flag icon
- + Accurate 45-degree isometric perspective on a simple diorama base
- − Texture of the rice is a bit bubbly and less realistic compared to the PBR request
- − The plate looks slightly simplified even for a cartoon style
GPT Image 1.5
- + Very high detail in materials and textures, especially the wood and ceramic
- + Balanced and vibrant colors with great lighting
- + Effective use of the isometric diorama concept with additional elements
- − Ignored the request for 'minimal' garnish and plate by adding a teapot, cups, and soy sauce bottle
- − The text 'SUSHI' is relatively small and less bold than requested
- − Technically more of a 3D render than a 'cartoon scene'
Verdict: FLUX.2 [pro] followed the prompt more closely by keeping the scene minimal and the text prominent, whereas GPT Image 1.5 added many extra elements not requested in the prompt. While GPT Image 1.5 has superior texture work with its wood and ceramic materials, FLUX.2 [pro] captured the specific 'cartoon' aesthetic and specific layout requirements better.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
FLUX.2 [pro]
- + Perfectly captures the caricature style with simplified cartoon features.
- + Excellent integration of all themes into a cohesive TV studio set.
- + Cleverly includes ice rinks on the desk and background screens.
- − The facial likeness to the source image is somewhat reduced by the generic cartoon style.
- − The character's hands are a bit blocky and awkwardly proportioned.
GPT Image 1.5
- + Stronger facial likeness to the source person within the caricature style.
- + High level of detail in the hair and background elements.
- + Humorous and creative inclusion of a dog wearing a hockey helmet.
- − The 'NEWS' microphone head is oddly shaped and slightly detached from the handle.
- − The hand holding the papers has an extra finger/anatomical confusion.
Verdict: Both models followed the instructions well, but GPT Image 1.5 is the winner for its superior ability to maintain the facial likeness of the source subject while applying the caricature effect. While FLUX.2 [pro] created a very clean cartoon scene, it felt more like a generic character, whereas GPT Image 1.5 captured the specific smile and features of the woman in the original photo alongside a funnier 'hockey dog.'
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent anatomical details on the kitten and fox paws.
- + Superior lighting with realistic dew sparkles and gentle god rays.
- + Clean, artistic composition that feels less cluttered.
- − Failed to include the baby bunny requested in the prompt.
- − The kitten's stripes look slightly artificial/painted on.
GPT Image 1.5
- + Included all four requested animals: puppy, kitten, bunny, and fox.
- + Captured the 'tumbling together' action more effectively than the other model.
- + Vibrant golden hour lighting with strong god rays.
- − Noticeable anatomical errors, such as the kitten having three hind legs/paws in the air.
- − The fox's face/mouth area is a bit messy and less defined.
- − Higher level of visual noise and over-sharpening compared to Model A.
Verdict: FLUX.2 [pro] produced a much cleaner and more aesthetically pleasing image with better lighting and textures, but it completely missed the bunny. GPT Image 1.5 adhered better to the prompt by including all characters and the requested tumbling action, but the image contains significant anatomical glitches and lacks the refined polish of the first model. FLUX.2 [pro] is preferred for its technical quality despite the missing element, as GPT Image 1.5's errors are quite distracting.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
FLUX.2 [pro]
- + Perfectly captures the Studio Ghibli art style with clean linework and watercolor textures.
- + Maintains excellent structural preservation of the original meme characters' poses and clothing.
- + The background adds Ghibli-esque details like Japanese signage and a hand-painted town aesthetic.
- − The color palette is slightly desaturated compared to the 'lush' look sometimes associated with Ghibli.
GPT Image 1.5
- + Excellent use of warm, dreamy lighting and soft pastel colors as requested.
- + Captures the emotional essence of the prompt through a shimmering, nostalgic atmosphere.
- − The art style leans more toward generic modern 'shoujo' anime rather than the specific Ghibli aesthetic.
- − Significant loss of detail and clarity due to excessive glow and soft-focus effects.
Verdict: FLUX.2 [pro] is the clear winner as it masterfully replicates the specific Studio Ghibli art style while retaining the unmistakable composition and character details of the original meme. GPT Image 1.5 provides a beautiful, dreamy atmosphere, but it fails to capture the distinct Ghibli linework and character design, resulting in a more generic anime look.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
FLUX.2 [pro]
- + Successfully added wind effect to hair while keeping the face recognizable.
- + Leaves are integrated into the foreground, including some overlapping the dog.
- + Good preservation of the original bridge and background elements.
- − The hair edit creates some jagged artifacts against the sky.
- − A leaf is awkwardly sticking out of the dog's mouth.
GPT Image 1.5
- + Excellent addition of many dynamic, semi-blurred leaves for a sense of movement.
- + Very natural hair-in-the-wind effect with fine strands.
- + High level of consistency with the original background and lighting.
- − Tiny alteration to the smile/facial features compared to the original.
Verdict: Both models followed the instructions well, adding blowing hair and flying leaves. FLUX.2 [pro] placed many leaves in the immediate foreground, but some feel static or poorly placed (like the leaf in the dog's mouth). GPT Image 1.5 created a much more 'energetic and lively' feel by using different sizes and slight motion blurs on the leaves, along with a more realistic wind effect on the hair.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.2 [pro]
- + Excellent adherence to the 'light background' and 'vector emblem' prompt requirements.
- + Perfect text rendering for both 'Caffè Florian' and 'Est. 1720'.
- + Clean, balanced composition with professional minimalist aesthetic.
- − The steam effect is a bit abstract compared to a natural plume.
GPT Image 1.5
- + Strong vintage feel with high-quality shading on the cloche.
- + Accurately captured all text elements and the banner style.
- − Ignored the 'light background' requirement by using a black background.
- − The composition feels slightly bottom-heavy with the large banner.
Verdict: FLUX.2 [pro] followed the prompt more accurately, specifically adhering to the request for a light background and a minimalist vector style. While GPT Image 1.5 produced a detailed and visually appealing illustrative logo, it failed the background color requirement and feels less like a clean vector emblem than its competitor.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.2 [pro]
- + Clean, minimalist vector aesthetic that perfectly matches the 'modern infographic' request.
- + Higher text legibility for the main headings and names.
- + Sophisticated composition that uses vertical space well to tell a story.
- − The Saturn V icon for the 'Launch' step is poorly rendered and inaccurate.
- − Filler body text is gibberish.
GPT Image 1.5
- + Much better technical accuracy for icons, specifically the Saturn V and Lunar Module.
- + Stronger adherence to the NASA-inspired color palette with bold use of muted red.
- + Clearer distinction between all six requested phases of the mission.
- − The layout is quite cramped, with elements overlapping the borders and the top text cut off.
- − Inconsistent line weights and styles across different icons.
Verdict: FLUX.2 [pro] produces a more 'finished' looking piece of graphic design with elegant typography and a cohesive modern style, though its technical icons are weak. GPT Image 1.5 creates a much more accurate representation of the Apollo hardware and follows the step-by-step prompt more closely, but the final image suffers from poor framing and a cluttered composition. GPT Image 1.5 is the preferred choice for a functional infographic despite the layout issues, as its icons actually represent the subject matter correctly.
Explore each model
OpenAI's state-of-the-art image generation model with better instruction following and adherence to prompts