Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [flex] Black Forest Labs GPT Image 1.5 OpenAI

Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.

FLUX.2 [flex]

24.8 arena score

#14 of 62 in Text-to-Image

Skill signature · Text-to-Image

GPT Image 1.5

27.1 arena score

#5 of 62 in Text-to-Image

Top 3 in Image Editing
Vote tally

Where the votes landed

FLUX.2 [flex]

41.2%

win rate

Ties

0.0%

GPT Image 1.5

58.8%

win rate

41.2% 0.0% ties 58.8%
Shared challenges 19

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [flex]
GPT Image 1.5
33% wins 0% ties 67% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Clean, minimalist composition with a high-quality photographic feel.
  • + Excellent rendering of the glass cube's beveled edges and physical interactions with the book.
  • + Subtle and realistic soft lighting from the left window.
  • The plant is quite blurred, making its visibility through the glass less distinct than requested.
  • The blue sphere has a matte texture that lacks the reflective properties typically found in such a scene.

GPT Image 1.5

  • + Strong adherence to the requirement of seeing the plant through the glass.
  • + The blue sphere features more realistic reflections, including the window light.
  • + Good texture on the red book cover and the wooden table grain.
  • The light reflection on the table looks slightly repetitive or artificial.
  • The bottom of the glass cube has a mirrored base which wasn't requested and slightly confuses the geometry.

Verdict: Both models followed the complex spatial instructions perfectly. FLUX.2 [flex] produced a more aesthetically pleasing and physically consistent image with superior lighting, while GPT Image 1.5 did a better job of making the plant visible through the glass as requested in the prompt.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [flex]
GPT Image 1.5
86% wins 0% ties 14% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Successfully integrates the exact car and man into a California coastline setting.
  • + Preserves the full silhouette and design details of the Rolls Royce Phantom Drophead Coupé.
  • + Maintains excellent visual coherence with realistic motion blur on the wheels and road.
  • The man is slightly small in the composition compared to the car's scale.
  • The character's face loses some detail due to the distance of the shot.

GPT Image 1.5

  • + Excellent character preservation, capturing the man's face, hair, and clothing accurately.
  • + Stronger focus on the 'driving' aspect with a clear view of the interior and hands on the wheel.
  • + The coastline background is highly scenic and matches the requested location well.
  • Crops out the majority of the car, which was a primary subject of the source images.
  • The perspective of the car interior feels slightly claustrophobic compared to the open-top source.

Verdict: Both models successfully combined the man and the car into the requested setting. FLUX.2 [flex] is the better overall edit because it preserves the entire car and shows the action from a cinematic distance, whereas GPT Image 1.5 crops the image so tightly that the identity of the car is mostly lost.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [flex]
GPT Image 1.5
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to 'motion blur from passing cars' with long-exposure style headlights.
  • + Very realistic skin texture and lighting on the subject.
  • + Strong cinematic composition with a shallow depth of field reflecting a 50mm lens.
  • The bicycle frame geometry is slightly physically impossible near the pedals.
  • The man's hands are a bit muddled in the interaction with the chain.

GPT Image 1.5

  • + Captures a more authentic 'candid' street photography feel with tools and accessories.
  • + Excellent texture on the wet pavement and clothing.
  • + The bicycle design is more coherent and realistic for a commuter bike.
  • Missing the 'motion blur from passing cars' requested in the prompt.
  • The car in the background is static rather than blurred by movement.

Verdict: FLUX.2 [flex] followed the technical instructions of the prompt more closely, specifically the motion blur and the bokeh typical of a 50mm lens. However, GPT Image 1.5 captured a more believable scene with better bicycle details and a natural 'candid' atmosphere, though it failed the specific motion blur requirement.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Symmetrical and clear armor engraving and leather strap details.
  • + Excellent representation of the bead requirement in the braided hair.
  • + Balanced composition with effective use of bokeh sparks.
  • The scars look like surface-level digital paint rather than integrated skin texture.
  • The lighting on the face is a bit flat compared to the dramatic armor highlights.

GPT Image 1.5

  • + Exceptional skin texture with realistic pores, dirt, and integrated scarring.
  • + Stunning interplay of warm torchlight and cool shadows on the skin and metal.
  • + Highly lifelike eyes with complex reflections.
  • Some minor geometric inconsistencies in the armor engravings on the left side.
  • The straps/buckles are less defined than in the competing image.

Verdict: While FLUX.2 [flex] provides a cleaner more symmetrical armor design with better adherence to the hair bead request, GPT Image 1.5 offers a significantly more lifelike and cinematic portrait. GPT Image 1.5 excels in hyper-realistic texture on the face and skin, and its use of light and shadow creates a much more convincing 'battle-worn' atmosphere.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [flex]
GPT Image 1.5
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Strict adherence to the requested grid layout for photos.
  • + Balanced use of colors corresponding to the specific sections.
  • + Professional mock-up presentation with realistic shadows.
  • Text is largely gibberish and unreadable.
  • The food photos in the grid don't always match the section they are under (e.g., pizza and steaks under the 'Appetizers' header).

GPT Image 1.5

  • + Perfectly legible and logical text with realistic menu items and descriptions.
  • + Strong photographic quality with food that looks appetizing and relevant to the sections.
  • + Good use of bold sans-serif fonts as requested.
  • Layout is more of a horizontal split than a distinct grid as requested.
  • The 'Appetizers' section has pizza photos adjacent to it, which is slightly confusing.

Verdict: While FLUX.2 [flex] adhered better to the 'grid' layout requested in the prompt, GPT Image 1.5 produced a far more usable and professional result. GPT Image 1.5's ability to render perfectly legible, contextually accurate text makes it the superior choice for a graphic design task like a menu.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Extremely clean and legible typography with a realistic flame effect.
  • + The 'exploded' effect feels more controlled and allows for a clearer view of individual high-quality textures.
  • + Excellent lighting on the food items that matches the fiery background.
  • The composition feels slightly more like a static stack than a dynamic explosion.

GPT Image 1.5

  • + Stronger sense of motion and 'exploded' energy with flying debris and angled components.
  • + Very high level of detail on the grilled texture of the patty and the toasted bun.
  • + Great integration of the starburst effect into the fiery scene.
  • The text is a bit cramped at the top and bottom of the frame.
  • The overall image is very busy, which slightly reduces the focus on the product.

Verdict: FLUX.2 [flex] produces a much cleaner and professional-looking advertisement with superior text rendering and a clear visual hierarchy. While GPT Image 1.5 captures the 'dynamic explosion' aspect with more energy, the legibility and polished aesthetics of FLUX.2 [flex] make it a better fit for a commercial ad prompt.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent chalk texture with realistic smudges and dust on the board.
  • + Perfect spelling and completion of the text prompt including the footer.
  • + Strong photographic composition with a realistic wooden frame and depth of field.
  • The font style feels slightly more uniform and less 'handwritten' than requested.
  • The lighting reflection is a bit intense at the top center of the board.

GPT Image 1.5

  • + Very realistic chalk handwriting style with authentic variations in letter height and slant.
  • + Excellent adherence to the 'slanted' and 'natural variation' part of the prompt.
  • + Perfect transcription of the requested text including the added menu item details.
  • The overall image resolution and sharpness are lower than the competitor.
  • The framing is very tight, losing the 'cozy café' environmental context.

Verdict: Both models followed the complex text instructions perfectly, which is impressive. FLUX.2 [flex] produced a much higher quality image with better lighting and environmental details, while GPT Image 1.5 captured the 'natural handwriting' and 'chalk slant' request more authentically.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent character replication including the sunglasses, scarf patterns, and facial hair.
  • + Good environment matching with the yellow lighting and red ottoman.
  • + Accurately places the character in a crouched, dynamic position on the ottoman.
  • The pose is significantly altered from the source, missing the specific 'crossed-leg' balance seen in Image 1.
  • The raised arm is bent and pointing rather than the straight, elegant extension in the reference.

GPT Image 1.5

  • + Successfully replicates the specific crossed-leg pose and body lean from the reference image.
  • + Maintains the character's clothing details including the scarf tassels and sunglasses.
  • + High visual quality and seamless integration of the character into the yellow studio environment.
  • The lower hand pose is slightly generic compared to the delicate finger position in Image 1.
  • The top edge of the image is cut off slightly compared to the reference frame.

Verdict: GPT Image 1.5 is the clear winner as it successfully combined the character from Image 2 with the 'exact' complex pose of Image 1, particularly the difficult leg crossing. FLUX.2 [flex] captured the character details well but defaulted to a more generic crouching pose, failing to follow the core instruction of replicating the precise body position.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the specific prompt instruction of the horse being on top.
  • + Beautiful cinematic lighting and color palette with a clear surreal aesthetic.
  • + High anatomical quality for both the horse and the astronaut's suit.
  • The connection point between the horse's front legs and the astronaut's hands is a bit physically ambiguous.
  • Small artifacts present in the asteroid field.

GPT Image 1.5

  • + High level of surface detail on the lunar ground and space lander.
  • + Clear, sharp focus across the entire composition.
  • Completely failed the negative constraint to have the horse on top.
  • Commonplace, literal interpretation rather than the requested surreal subversion.
  • The dust effect looks a bit disconnected from the vacuum of space context.

Verdict: FLUX.2 [flex] successfully followed the difficult logical constraint of placing the horse on top of the astronaut, resulting in a truly surreal and creative image according to the prompt. GPT Image 1.5 produced a high-quality but generic image that directly ignored the specific instruction to reverse the typical horse-rider relationship.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent lighting and high-definition texture on the capybara's fur.
  • + Dynamic composition with a wide-angle view that captures more of the NYC city lights.
  • + Realistic rendering of the capybara's paws gripping the steering wheel.
  • The transition between the capybara's head/neck and the human-looking hands is slightly uncanny.
  • The capybara's head is turned away from the road, making the pose feel less natural.

GPT Image 1.5

  • + Perfect framing through the windshield creates a cinematic and immersive 'inside the taxi' feel.
  • + The capybara's expression is very professional and forward-facing as a driver should be.
  • + Successfully captures the specific requested 'bored' expression of the passenger.
  • The passenger's hair and facial features are slightly blurry/out of focus compared to the foreground.
  • The taxi cap logo is a bit generic compared to the detailed badge in Model A.

Verdict: Both models followed the prompt excellently, but GPT Image 1.5 offers a more convincing and cohesive composition by framing the scene through the front windshield, which creates a more realistic narrative. While FLUX.2 [flex] has superior technical resolution and sharper details on the far background, the placement of the capybara and the 'normalcy' of the scene feels more natural in GPT Image 1.5.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent typography with clean, legible gothic fonts
  • + High-quality central jack-o-lantern rendering with effective glowing light
  • + Well-defined torn parchment border effect
  • The composition feels slightly empty in the mid-ground
  • The border design is a bit repetitive

GPT Image 1.5

  • + Atmospheric and rich vintage texture that feels like real weathered parchment
  • + Intricate border with high detail on the thorns and spiderwebs
  • + Stronger thematic coherence between the background elements and the foreground
  • Text rendering is slightly less sharp than Model A
  • The 'thorns' in the border are very busy, making the edges look a bit messy

Verdict: Both models followed the prompt exceptionally well, including all requested text and visual elements. FLUX.2 [flex] produced a cleaner, more modern graphic design with superior typography, while GPT Image 1.5 captured the 'vintage' and 'moody' atmosphere more effectively through its textured, painterly style. GPT Image 1.5 is the winner for its superior artistic depth and more immersive gothic aesthetic.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [flex]
Before After
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of original identity and facial details.
  • + Natural-looking buzz-cut texture that matches the existing beard density.
  • + Flawless integration of the hairline and ears.
  • The hair is somewhat thin/short, bordering on a buzz cut rather than a 'full, thick head' as requested.

GPT Image 1.5

  • + Successfully provides a much thicker, fuller head of hair.
  • + Matches the hair texture well to the existing beard texture.
  • + Preserves the overall environment and composition.
  • Slightly alters the facial structure, making the man look somewhat younger or like a different person.
  • The hairline integration on the forehead is a bit harsh.

Verdict: FLUX.2 [flex] provides the most realistic integration, perfectly preserving the subject's identity, though it opted for a shorter buzz-cut style rather than a long, thick head of hair. GPT Image 1.5 followed the prompt's instruction for 'thick hair' more literally but at the cost of subtly changing the subject's face. FLUX.2 [flex] is the winner for its superior realism and identity preservation.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the 'cartoon' and 'soft texture' requirement
  • + Clean minimalist composition with a perfect 45-degree isometric angle
  • + Strong text clarity and accurate flag icon
  • Texture of the sushi toppings is slightly too plastic-looking for PBR materials

GPT Image 1.5

  • + Highly impressive realistic PBR materials, especially on the wood and ceramic
  • + Detailed interpretation of a diorama with a 3D base
  • + Excellent text rendering and layout
  • Ignored the 'minimal garnish' instruction by adding excessive accessories like a teapot and soy sauce
  • Less 'cartoon' in style than requested, leaning more towards hyper-realism

Verdict: Both models followed the complex prompt very well, particularly regarding text layout and isometric perspective. FLUX.2 better captured the requested 'cartoon' and 'minimalist' aesthetic, whereas GPT Image 1.5 produced a much more detailed and texture-rich scene that arguably ignored the 'minimal' constraint but looked more professional overall.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Successfully incorporates all elements including hockey, dogs, and the news desk.
  • + Maintains the exact denim outfit from the source image.
  • + Consistently applies a clean, 2D cartoon caricature style across the whole image.
  • The 'hockey' element is a bit subtle, represented primarily by the dogs' jerseys and one stick.
  • The text on the news desk ('THEIANAES') is nonsensical.

GPT Image 1.5

  • + Excellent caricature of the subject's face while maintaining a strong resemblance to the source.
  • + High level of humorous exaggeration with the hockey-playing dog and news ticker.
  • + Clearly legible and relevant text in the graphic overlay.
  • Changed the subject's outfit from the denim shirt in the source to a red dress.
  • The hand holding the microphone has anatomical issues (too many fingers/blurry).

Verdict: Both models followed the prompt well, but GPT Image 1.5 captured a much better caricature of the specific woman in the source photo while adding more creative and humorous details like the dog in the hockey helmet. FLUX.2 [flex] was better at preserving the subject's original clothing, but the overall style is more generic and the text is garbled.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [flex]
GPT Image 1.5
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent composition with clear space between animals, allowing their full bodies and movements to be seen.
  • + Superior lighting and atmosphere, featuring beautiful god rays and a distinct sunrise glow.
  • + Includes all four requested animals with distinct, accurate features for each species.
  • The kitten and puppy appear to be floating or jumping rather than 'tumbling' as specified.
  • The butterflies look slightly pasted on rather than fully integrated into the lighting of the scene.

GPT Image 1.5

  • + Captures the 'tumbling together' part of the prompt much more effectively with a cozy, grouped interaction.
  • + Exceptional fur texture and detail, particularly on the kitten’s paws and the fox's face.
  • + The warm, golden light feels very integrated into the scene's dew and sparkles.
  • The fox is missing its lower body/hind legs, looking like a disembodied head and torso.
  • The composition is very crowded, making it harder to distinguish the individual forms of the four animals.

Verdict: Both models followed the prompt well, but FLUX.2 [flex] produced a much more coherent and anatomically correct image with a better sense of scale. While GPT Image 1.5 captured the 'tumbling' aspect and soft textures beautifully, the cramped composition resulted in missing limbs for the fox, whereas FLUX.2 [flex] delivered a professional, 8K-style landscape composition with clear species differentiation.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of the original pose and character positions.
  • + Captures the iconic line-heavy Studio Ghibli 'Only Yesterday' or 'Princess Kaguya' style aesthetic.
  • + Maintains the specific plaid pattern of the man's shirt very accurately.
  • The woman on the right misses the angry/shocked expression from the original.
  • The color palette is a bit washed out compared to typical Ghibli vibrant landscapes.

GPT Image 1.5

  • + Successfully captures the warm, nostalgic, and dreamy lighting requested.
  • + Preserves the emotional context of the scene, including the girlfriend's angry expression.
  • + Beautiful watercolor-inspired texturing and soft pastel color palette.
  • The woman in the foreground has significantly longer and more voluminous hair than the source image.
  • More of a general shoujo-manga aesthetic rather than specifically Studio Ghibli.

Verdict: FLUX.2 [flex] provides a more technically accurate edit in terms of preserving the exact shirt patterns and character silhouettes, adopting a clean Ghibli line-art style. However, GPT Image 1.5 better captures the emotional 'warmth' and 'dreamy mood' requested in the prompt, successfully translating the girlfriend's angry expression into an illustrated form. FLUX.2 is preferred for structural preservation, but GPT Image 1.5 is preferred for thematic atmosphere.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [flex]
Before After
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of the source image's identity and background.
  • + The hair motion looks natural and maintains the woman's facial structure.
  • + The dog's tail is subtly modified to suggest wagging movement.
  • The flying leaves look like static green shapes pasted over the foreground rather than part of the environment.
  • The motion feel is slightly less energetic compared to model_b.

GPT Image 1.5

  • + Stronger sense of 'dynamic motion' with more abundant and varied leaf textures.
  • + Great hair dynamics that create a wind-swept appearance.
  • + Good color integration of the leaves with the autumn-toned lighting.
  • Noticeable anatomical distortion around the woman's right shoulder and neckline where the hair meets the jacket.
  • The dog's left front leg is morphed and looks unnatural compared to the source.

Verdict: Both models successfully followed the instruction to add wind and leaves. FLUX.2 [flex] is the better choice for high-fidelity editing as it preserves the proportions and details of the original woman and dog perfectly, whereas GPT Image 1.5 introduces structural errors in the woman's jacket and the dog's legs. However, GPT Image 1.5 captured the 'energetic feel' better with more varied leaf shapes and motion blur.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [flex]
GPT Image 1.5

AI Judge Analysis

FLUX.2 [flex]

  • + Perfect adherence to the light background request.
  • + Excellent minimalist vector style that feels like a clean modern-vintage logo.
  • + Very clean and readable typography for both the main title and the banner text.
  • The texture is extremely subtle, almost invisible compared to the prompt's request.

GPT Image 1.5

  • + Excellent use of vintage textures and shading on the cloche dome.
  • + Accurate recreation of all requested text elements, including the accent on 'Caffè'.
  • + Stronger 'vintage' character in the illustration style.
  • Completely ignored the request for a light background, providing a black one instead.
  • Composition feels a bit more cluttered and less 'minimalist' than Model A.

Verdict: Both models followed the complex text instructions perfectly, which is impressive. FLUX.2 [flex] is the winner because it adhered to the 'light background' instruction, whereas GPT Image 1.5 generated a black background. FLUX.2 [flex] also captured the 'minimalist vector' aesthetic much better for a modern logo use case.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [flex]
GPT Image 1.5
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the clean, modern vector aesthetic.
  • + Very high text clarity and accuracy across all labels.
  • + Elegant use of negative space and layout for a professional infographic feel.
  • Failed to include the 6th step (Landing).
  • The 'Translunar' iconography is a bit disconnected from the other elements.

GPT Image 1.5

  • + Successfully included all 6 requested steps from Launch to Landing.
  • + Strong consistency in the flat-vector illustration style.
  • + Creative inclusion of the crew silhouettes and the 'Tranquility' landing site marker.
  • Image is cropped at the top, cutting off the main title.
  • The 'Translunar' arc is somewhat cluttered compared to the other clean sections.

Verdict: FLUX.2 [flex] produced a much more professional and aesthetically pleasing 'modern' infographic with crisp typography, but it failed the prompt instruction to include 6 steps. GPT Image 1.5 followed the sequence instructions perfectly and included all steps, though significantly suffered from poor framing that cut off the top of the poster.

Next steps

Explore each model