Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [klein] 9B Black Forest Labs GPT Image 1.5 OpenAI

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

FLUX.2 [klein] 9B

20.6 arena score

#13 of 32 in Image Editing

Skill signature · Image Editing

GPT Image 1.5

27.1 arena score

#7 of 62 in Text-to-Image

Top 3 in Image Editing
Vote tally

Where the votes landed

FLUX.2 [klein] 9B

50.0%

win rate

Ties

0.0%

GPT Image 1.5

50.0%

win rate

50.0% 0.0% ties 50.0%
Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent photorealism in the wood texture and window lighting
  • + Natural-looking refractions within the thick glass
  • + High resolution with cinematic shallow depth of field
  • The glass container is more of a hollow vase/vessel than a perfect cube
  • The blue sphere appears to be floating unnaturally instead of resting on the bottom

GPT Image 1.5

  • + Perfect adherence to the geometry of a cube
  • + Accurate placement of the sphere resting on the internal floor
  • + Clear visibility of the plant through the glass panels
  • Lighting is a bit flat compared to Model A
  • The glass edges look slightly artificial/digital

Verdict: Both models followed the prompt instructions perfectly, including the spatial relationships between the book, cube, sphere, and plant. FLUX.2 [klein] 9B offers a more photographic aesthetic with superior lighting and texture, but GPT Image 1.5 achieved better geometric accuracy for the 'cube' and more realistic physics for the sphere.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the original car's design, proportions, and lighting.
  • + Successful implementation of California coastline scenery in the background.
  • + Good sense of motion with wheel blur and road lines.
  • The subject is barely visible, with only the top of his hair showing behind the headrest.
  • Does not effectively integrate the specific person from the source image.

GPT Image 1.5

  • + Excellent subject preservation, accurately featuring the man's face, hair, and scarf from the source image.
  • + Dynamic composition that creates a connection with the viewer.
  • + High level of detail in the car interior and background coastline.
  • The car's proportions are slightly distorted compared to the source, appearing slightly compressed.
  • The lighting on the man does not perfectly match the bright, high-sun environment of the background.

Verdict: FLUX.2 [klein] 9B does a better job of placing the original car into a new environment while maintaining its integrity, but it fails to include the requested man in a meaningful way. GPT Image 1.5 successfully combines both source images by placing the specific man inside the car, making it the superior editor for following the prompt instructions.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent skin texture and realistic hand details
  • + Effective use of shallow depth of field and street reflections
  • + Great subject focused interaction with the bike
  • The white car in the background lacks the requested motion blur
  • Framing feels a bit too clean and centered for a candid prompt

GPT Image 1.5

  • + Better execution of motion blur on the passing vehicle
  • + More realistic 'candid' and 'imperfect' framing
  • + Excellent wet pavement reflections
  • The man's hand/tool interaction with the wheel is anatomically muddled
  • The skin texture is slightly more smoothed compared to Model A

Verdict: Both models followed the prompt well, but they succeeded in different areas. FLUX.2 [klein] 9B produced a much higher quality subject with superior skin textures and hand details, though it failed to incorporate motion blur. GPT Image 1.5 captured the 'candid' atmosphere and technical motion blur more accurately, but at the cost of some fine detail and anatomical clarity in the hands.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Exceptional intricate engraving on the plate armor
  • + Clear and accurate rendering of the beaded braids
  • + Very strong adherence to the strap and cloth texture requirements
  • The lighting feels a bit more studio-like and less atmospheric compared to Model B
  • Symmetry in the facial features and braids feels slightly artificial

GPT Image 1.5

  • + Beautiful cinematic lighting with realistic warm reflections on the metal
  • + Excellent battle-worn aesthetic with realistic dirt and grittiness
  • + Very lifelike eyes and expressive facial features
  • The braids are a bit messy and the beads are less distinct than in Model A
  • The engraving on the armor is less sharp and detailed

Verdict: Both models followed the prompt exceptionally well, but with different artistic directions. FLUX.2 [klein] 9B excels at technical clarity, showing every detail of the engraved armor and the specific hair beads, while GPT Image 1.5 offers a more cinematic, gritty, and atmospheric composition that feels more like a lived-in character. GPT Image 1.5 wins slightly due to superior lighting and more realistic integration of the battle-worn elements.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Includes all requested text categories
  • + Clean white background with professional shadows
  • Non-existent language and garbled text throughout the body
  • Repetitive food photos featuring four similar pizzas instead of varied dishes
  • Overlapping text in the header

GPT Image 1.5

  • + Excellent text legibility and coherent menu items
  • + High alignment between dish names and corresponding high-quality photos
  • + Very professional and clean multi-column layout
  • Crop is slightly tight on the bottom edge
  • The 'grid' for photos is asymmetrical compared to the text section

Verdict: GPT Image 1.5 is the clear winner as it produces a functional menu design with perfectly legible, relevant text and varied food photography that matches the sections. FLUX.2 [klein] 9B fails significantly on text rendering, producing gibberish, and lacks variety in its food images despite a decent overall aesthetic.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [klein] 9B
GPT Image 1.5
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography rendering and clean starburst icon.
  • + High photorealistic quality with clear, appetizing textures on the bun and meat.
  • + Modern professional composition suitable for a high-end food advertisement.
  • The 'exploded' effect is very conservative, with layers barely separated.
  • The background is more static and less 'fiery' compared to the other model.

GPT Image 1.5

  • + Captures the 'exploded' request perfectly with wide separation of all ingredients.
  • + Intense, high-energy background with dynamic embers and smoke.
  • + Good text integration with a fiery glowing effect on the main title.
  • The 'LIMITED TIME ONLY' text is placed at the very bottom and slightly cut off by the frame.
  • Image has a slightly over-processed or 'deep fried' HDR look that reduces realism.

Verdict: FLUX.2 [klein] 9B produces a cleaner, more professional-looking advertisement with superior font rendering, but it plays it very safe with the 'exploded' burger concept. GPT Image 1.5 follows the core instruction of an exploded burger more literally and creates a more energetic atmosphere, though the text placement and image texture are less refined. FLUX.2 is preferred for overall design quality and realism.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent chalk texture on the board surface
  • + Highly accurate text rendering of the complex prompt
  • + Strong aesthetic with realistic cursive and layout
  • Small spelling error in 'fress' instead of 'fresh'
  • Composition feels slightly crowded with the smudges

GPT Image 1.5

  • + Clean layout with consistent handwriting
  • + Accurate spelling throughout the entire text
  • + Realistic wooden frame integration
  • Texture appears more like a digital overlay than physical chalk
  • The cursive title is less elegant compared to the competing model

Verdict: Both models followed the complex prompt remarkably well, including the specific date and menu prices. FLUX.2 [klein] 9B is the preferred choice due to its superior chalk texture and more authentic handwriting variations, despite a minor typo. GPT Image 1.5 is visually cleaner but lacks the tactile, dusty quality of a real chalkboard.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [klein] 9B
GPT Image 1.5
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Near-perfect adherence to the complex anatomical pose including head tilt and arm angle.
  • + Highly accurate recreation of character accessories like the scarf pattern and sunglasses.
  • + Excellent lighting consistency with the yellow studio environment.
  • One hand is clenched into a fist, differing from the reference pose's open fingers.
  • Minor skin tone inconsistency between the face and feet.

GPT Image 1.5

  • + Faithful reproduction of the character's facial features and specific expression.
  • + Good preservation of the red box prop from the source environment.
  • Fails the core pose requirement by keeping the torso and head upright.
  • Proportions are slightly distorted where the legs cross.

Verdict: FLUX.2 [klein] 9B is the clear winner as it successfully replicates the difficult, dynamic body position and head tilt from the pose reference. GPT Image 1.5 fails to capture the lean and specific orientation of the original image, resulting in a much more static composition.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent high-detail rendering of the space background and planets
  • + Clean composition with a clear 'astral' aesthetic
  • + Precise rendering of the astronaut's face and suit details
  • Failed the core positional prompt: 'horse on top'
  • Includes distracting hallucinated text at the bottom
  • The horse's Anatomy lacks appropriate space gear, breaking immersion

GPT Image 1.5

  • + Stronger cinematic lighting and dynamic movement
  • + Detailed lunar surface and space debris integration
  • + Realistic horse harness and textures matching the environment
  • Failed the core positional prompt: 'horse on top'
  • Composition feels slightly cluttered with too many floating elements
  • Small artifacts in the horse's mane and dust clouds

Verdict: Both models completely failed the negative constraint to have the 'horse on top' of the astronaut, instead opting for the common 'astronaut riding horse' trope. FLUX.2 [klein] 9B provides a cleaner, more vibrant space scene, while GPT Image 1.5 offers a grittier, more cinematic lunar landing style. FLUX is slightly less effective due to the inclusion of nonsensical text.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Successfully preserved the subject's face, hair, and vitiligo patterns
  • + Accurately replicated the coat, scarf, and jeans from the reference image
  • + Included the gold jewelry and watch as requested
  • Lighting on the coat is slightly too bright compared to the beach environment

GPT Image 1.5

  • + Perfectly replicated the texture and drape of the plaid scarf
  • + Maintained the background and wooden structure accurately
  • Failed to include the subject's face, which was a core instruction
  • Missing the gold jewelry shown in the base outfit

Verdict: FLUX.2 [klein] 9B followed all instructions, including the critical requirement to keep the person's face unchanged while applying the complex outfit from the second image. GPT Image 1.5 failed by cropping out the subject's head entirely and omitting jewelry. FLUX.2 is the clear winner for maintaining character identity and outfit completeness.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent photorealism with sharp textures on the capybara and the woman.
  • + Clean composition with a clear view of both subjects.
  • + The expression on the woman perfectly matches the 'bored' instruction.
  • The woman is sitting in the passenger seat rather than the back seat.
  • The text on the hat contains gibberish ('NEW TALA').

GPT Image 1.5

  • + Correctly places the woman in the back seat as requested in the prompt.
  • + The text on the taxi cap is legible and accurate.
  • + High-quality realistic lighting and atmospheric bokeh in the background.
  • The capybara's paws look a bit like bird talons or strange human hands rather than natural capybara feet.
  • The perspective through the windshield is slightly confused compared to the car interior.

Verdict: While FLUX.2 [klein] 9B offers a slightly sharper and more detailed image, it failed a key spatial instruction by placing the businesswoman in the front passenger seat. GPT Image 1.5 adhered better to the layout of the prompt by putting her in the back seat and correctly rendering the text on the driver's cap.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography with clean, highly readable font choices
  • + Crisp cinematic lighting on the jack-o-lantern and trees
  • + Perfect layout with well-balanced negative space
  • The art style is a bit more 'modern digital' than 'vintage gothic'

GPT Image 1.5

  • + Authentic vintage gothic aesthetic with a distressed parchment texture
  • + Stronger thematic atmosphere with the graveyard and castle background
  • + Ornate border design that feels tactile and illustrative
  • Text is slightly less sharp and harder to read against the busy background
  • The layout feels a bit cramped compared to Model A

Verdict: Both models followed the complex prompt instructions perfectly, including all requested text and visual elements. FLUX.2 [klein] 9B produces a cleaner, more modern commercial layout that is very professional, while GPT Image 1.5 captures the 'vintage' and 'dark parchment' aspect of the prompt much more effectively with its gritty, detailed art style.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [klein] 9B
Before After
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent texture and realistic hair volume
  • + Near-perfect preservation of the original face, glasses, and background
  • + Seamless blending at the temples where the hair meets the existing beard
  • The hair covers a bit more of the forehead than might be expected for a classic hairline

GPT Image 1.5

  • + Strong prompt adherence for a 'thick head of hair'
  • + Good color matching between the new hair and existing beard
  • Slightly alters the facial features, making the eyes and nose look different from the source
  • The hairline connection to the glasses and temples is a bit blurry
  • The overall lighting on the face feels slightly flatter compared to the original

Verdict: FLUX.2 [klein] 9B is the clear winner as it successfully adds the hair while keeping the subject's identity and the image's original quality perfectly intact. GPT Image 1.5 manages to add the hair, but it inadvertently modifies the man's facial structure and eye shape, losing the likeness of the original subject.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent 3D cartoon aesthetic with soft, pillowy textures.
  • + Perfectly follows the request for a minimalist diorama style.
  • The flag icon is incorrect, resembling the flag of Yemen rather than Japan.
  • The sushi roll on the right has some structural clipping/warping issues.

GPT Image 1.5

  • + High-fidelity PBR materials with realistic textures on the wood and ceramics.
  • + Accurate Japanese flag icon.
  • + Very clean and bold text typography.
  • Included many extra items (soy sauce, tea pot, chopsticks) despite the request for 'minimal garnish'.
  • The scene feels a bit crowded relative to the 'minimalist' prompt.

Verdict: While GPT Image 1.5 provides much higher material detail and gets the flag correct, FLUX.2 [klein] 9B better captures the 'cartoon' and 'minimalist' aesthetic requested in the prompt. GPT Image 1.5 is the preferred model overall, however, due to its superior text rendering, correct cultural iconography, and impressive lighting and texture quality.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent caricature style with exaggerated facial features
  • + Cleverly integrates dogs into the hockey arena crowd
  • + Maintains a cohesive cartoon aesthetic throughout
  • The 'hockey' element is relegated to the background rather than being central to the character
  • Hand holding the microphone is poorly rendered with an extra digit/merged finger

GPT Image 1.5

  • + Strong resemblance to the subject in the source image
  • + Excellent integration of all prompt elements (desk, mic, dogs, and a literal hockey game)
  • + Humorous and creative detail of a dog wearing a hockey helmet
  • The 'caricature' style is less exaggerated and borders more on a digital painting
  • Minor background artifacts in the hockey player's leg area

Verdict: GPT Image 1.5 is the winner as it successfully blends all the requested elements—news anchor, dogs, and hockey—into a singular, humorous composition that clearly resembles the subject. While FLUX.2 [klein] 9B provides a more traditional caricature art style, it hides the hockey theme in the background and suffers from noticeable anatomy issues in the hands.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent clarity and vibrant flower details
  • + Crisp backlighting and god rays implementation
  • + Highly expressive and clean character rendering
  • Completely missed the 'baby bunny' requirement
  • Anatomy of the fox's front legs is slightly awkward

GPT Image 1.5

  • + Successfully included all four requested animals
  • + Beautiful 'tumbling together' composition that feels more organic
  • + Excellent soft fur texture and dew sparkle effects
  • Fox's paws are rendered as dark, indistinct blobs
  • The butterfly near the kitten is missing a body/head

Verdict: GPT Image 1.5 is the winner because it successfully included all four requested animals (dog, cat, fox, and bunny), whereas FLUX.2 [klein] 9B failed to generate the bunny. While FLUX.2 offered slightly higher clarity, GPT Image 1.5 better captured the 'tumbling together' action and the specific lighting requests like dew sparkles.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent watercolor aesthetic that perfectly nails the 'hand-painted' texture request.
  • + Near-perfect preservation of the character poses, clothing patterns, and spatial layout.
  • + Captures the Studio Ghibli character design style with clean line art and expressive eyes.
  • Has largely replaced the background urban street with a generic countryside/meadow setting.
  • The woman on the right has lost her distinctive angry expression, looking neutral instead.
  • The man's hand has some minor anatomical awkwardness.

GPT Image 1.5

  • + Successfully captures the requested dreamy, warm lighting and nostalgic glow.
  • + Preserves the urban street background from the source image better than the competitor.
  • + Maintains the distinct facial expressions, particularly the jealousy of the woman on the right.
  • The color palette is overly saturated and orange, drifting away from the requested 'soft pastels'.
  • Texture looks more like a digital filter than a hand-painted illustration.
  • The blur on the character in the foreground is a bit distracting.

Verdict: FLUX.2 [klein] 9B is the overall winner because it successfully transforms the image into a high-quality illustration that feels like genuine Studio Ghibli concept art, especially with its beautiful watercolor textures. While GPT Image 1.5 does a better job of preserving the specific background and facial expressions from the 'distracted boyfriend' meme, its execution feels like a digital filter rather than a true artistic transformation and fails to meet the 'soft pastel' requirement.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [klein] 9B
Before After
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the source subject's facial features and clothing.
  • + The flying leaves are crisp and integrate well into the environment.
  • + Maintains the exact scale and background of the original image.
  • The wind effect on the hair is slightly less dramatic than Model B.
  • A few leaves appear slightly flat against the subject.

GPT Image 1.5

  • + Stronger, more natural-looking wind effect on the hair.
  • + Good variety in leaf colors and shapes to suggest movement.
  • + Preserves the overall layout and dog's appearance perfectly.
  • Subtle change to the woman's face, specifically making the jawline slightly more angular.
  • One leaf appears to be growing out of the dog's head.

Verdict: Both models followed the instructions exceptionally well, adding hair motion and falling leaves while preserving the source image. GPT Image 1.5 is the winner for its more convincing wind effect on the hair, which creates a better sense of 'dynamic motion' compared to FLUX.2 [klein] 9B, which was slightly more static despite the additions.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography and spelling accuracy for both the brand name and the specific accents.
  • + Clean, vector-style execution that aligns perfectly with modern logo standards.
  • + Follows the light background requirement with subtle paper texture.
  • The steam element is a bit simplified/stylized compared to a natural plume.
  • The banner placement cuts into the cloche unlike a traditional emblem.

GPT Image 1.5

  • + Beautiful use of vintage illustrative textures and shading on the cloche.
  • + Elegant typography with a classic artisan feel.
  • + Good use of the banner at the base of the design.
  • Failed the negative space/background requirement by providing a black background instead of a light one.
  • The steam looks slightly detached from the cloche lid.
  • Slightly less legible 'Caffè' text compared to Model A.

Verdict: FLUX.2 [klein] 9B is the clear winner as it followed every instruction, including the specific requirement for a light background and subtle texture. While GPT Image 1.5 produced a beautiful and artistic vintage logo, it completely ignored the background color instruction, making it less useful for the requested application.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [klein] 9B
GPT Image 1.5

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Clean vector aesthetic with consistent high-contrast shapes.
  • + Includes creative silhouette avatars for the astronauts.
  • + Good use of the requested color palette.
  • Severe spelling errors in almost every label (e.g., 'EARDHT', 'TRANSLURAL', 'LANDINING').
  • Logical flow of icons is confusing and non-linear.

GPT Image 1.5

  • + Excellent layout that follows a clear, logical step-by-step sequence.
  • + Perfect spelling on all technical labels and mission steps.
  • + Strict adherence to the NASA-inspired color palette and flat-vector style.
  • Iconography is a bit crowded within the individual panels.
  • The 'Translunar' icon is slightly repetitive of the orbit tiles.

Verdict: GPT Image 1.5 is the clear winner as it perfectly follows the requested 6-step structure with accurate spelling and a professional infographic layout. While FLUX.2 has high visual clarity, its labels are gibberish and the sequence of steps is difficult to follow.

Next steps

Explore each model