Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [klein] 9B Black Forest Labs GPT Image 1 OpenAI

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

FLUX.2 [klein] 9B

20.6 arena score

#13 of 32 in Image Editing

Skill signature · Image Editing

GPT Image 1

23.2 arena score

#28 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [klein] 9B

0.0%

win rate

Ties

0.0%

GPT Image 1

100.0%

win rate

0.0% 0.0% ties 100.0%
Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent photographic realism with natural lens blur and lighting.
  • + Complex accurately rendered glass refractions of the sphere and background plant.
  • + High texture detail on the wooden table and book edges.
  • The glass container looks more like a thick vase than a perfect geometric cube.
  • The blue sphere is quite small and looks somewhat like a marble suspended in glass rather than sitting in a cube.

GPT Image 1

  • + Strong adherence to the 'cube' geometry and structural request.
  • + Clean, minimalist composition with clear separation of objects.
  • + Correct placement of the plant behind the cube as requested.
  • The glass cube lacks a top surface; the red book is floating over an open space.
  • The blue sphere appears to be floating inside the cube rather than resting on the bottom surface.
  • Lighting on the table is a bit flat compared to Model A.

Verdict: FLUX.2 [klein] 9B produces a much more realistic photograph with beautiful light and refraction, though the 'cube' is slightly rounded like a vase. GPT Image 1 follows the geometric instructions more literally with a sharp cube shape, but fails on physics as the book floats over an open top and the sphere floats in mid-air. FLUX.2 is the preferred model for its superior visual quality and more convincing handling of transparency.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the car's original proportions and design details.
  • + Effective implementation of motion blur on the wheels and road.
  • The man is barely visible, with only his hair silhouette showing behind the driver's seat.
  • The car has been shifted to the left-hand drive side of the frame despite being a right-hand drive model in the source.

GPT Image 1

  • + Successfully places the man prominently in the driver's seat with recognizable features from the source image.
  • + Captures the classic 'California coastline' aesthetic with cliffs and winding roads.
  • + Keeps the driver on the correct side of the vehicle based on the car's interior layout.
  • The hood of the car appears slightly elongated compared to the source image.

Verdict: While FLUX.2 [klein] 9B does a better job at preserving the technical details of the car, it fails to meaningfully include the 'man' from the second source image, only showing the top of his head. GPT Image 1 successfully integrates the specific man from the source image into the scene while maintaining high visual quality and a more dynamic composition.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent skin texture and realistic age details on the hands and face.
  • + Includes logical environmental details like a toolbox and wet ground reflections.
  • + Strong adherence to the 50mm lens and shallow depth of field request.
  • Lack of motion blur on the background car as requested.
  • The bicycle's chain and frame geometry are slightly physically inconsistent.

GPT Image 1

  • + Successfully captures a more 'imperfect' and candid framing.
  • + Excellent color grading that feels cinematic and cohesive.
  • + Better depiction of light rain through surface droplets and atmosphere.
  • Anatomy of the hand interacting with the hub is poorly rendered.
  • Lacks the requested motion blur on the background vehicles.
  • The bicycle's rear frame and chain guard area are structurally nonsensical.

Verdict: FLUX.2 [klein] 9B provides a much sharper image with superior detail in the man's skin and the mechanical tools, though it feels a bit more staged. GPT Image 1 captures the 'candid' atmosphere and cinematic lighting more effectively but fails significantly on technical details like hand anatomy and the bicycle's structure. FLUX.2 is the overall winner for its clarity and rendering of fine details.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent detail on the engraved plate armor and leather straps.
  • + Perfect adherence to hair braiding with small beads.
  • + Strong, clear lighting and high resolution contrast.
  • The scars look a bit painted on rather than naturally integrated into the skin texture.
  • The symmetry of the armor and composition is a bit rigid.

GPT Image 1

  • + More realistic skin texture and integration of dirt and scars.
  • + Dynamic use of shadows and lighting for a more cinematic feel.
  • + The expression feels more organically 'battle-worn'.
  • The armor engraving is less distinct and more muddy compared to Model A.
  • Fewer beads visible in the braids compared to the prompt requirements.

Verdict: FLUX.2 [klein] 9B delivers a sharper, more detailed image with superior representation of the ornate armor and requested hair accessories. GPT Image 1 offers a more evocative and gritty atmosphere with superior skin modeling, but lacks the intricate texture clarity and prompt-specific details found in FLUX.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Stronger adherence to the 'grid' request with four distinct photos
  • + Professional, high-end restaurant menu typography and spacing
  • + Clever use of color accents on section dividers
  • Nonsensical text and overlapping characters in the main header
  • The food images are repetitive, mostly showing four variations of pizza instead of distinct sections

GPT Image 1

  • + Excellent food photography that clearly distinguishes categories (salad, pizza, chicken, pasta)
  • + Cleaner, more readable sans-serif typography
  • + Better color vibrancy and professional composition
  • Small spelling errors like 'descrigion' and 'descripion'
  • Missed 'Mains' as a primary section header, though included it as a line item

Verdict: GPT Image 1 is the superior choice because its food photography actually corresponds to the menu categories requested, whereas FLUX.2 [klein] 9B filled its grid with four similar-looking pizzas. GPT Image 1 also features a much cleaner layout with significantly fewer garbled text artifacts compared to the messy header and strange overlapping text in FLUX.2.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography with clean, professional-grade rendering.
  • + Highly photorealistic textures on the bun and patties.
  • + Dynamic composition with splashes of sauce and sparks that create a sense of action.
  • The burger is less 'exploded' vertically compared to Model B, appearing more like a slightly separated stack.

GPT Image 1

  • + Perfectly captures the 'exploded' concept with distinct gaps between every layer.
  • + Vibrant fiery text effects that glow intensely against the background.
  • + Strong adherence to the starburst container for the price.
  • The price text is missing a digit, rendering as '€.99'.
  • The top bun and tomato slices have a slightly less realistic, waxy texture.
  • The composition feels a bit flatter and more centered compared to the dynamic angle of Model A.

Verdict: FLUX.2 [klein] 9B produces a more professional and commercially viable advertisement with superior text rendering and better texture realism. While GPT Image 1 followed the 'exploded' instruction more literally by spacing out the ingredients, it failed on a critical detail by omitting the '6' from the price and lacked the sharp photographic clarity found in FLUX.2.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent adherence to the 'elegant cursive' requirement for the title.
  • + Realistic chalk texture with visible smudges and erased marks adding to authenticity.
  • + Exceptional text rendering accuracy, completing the 'Brown Butter' item correctly.
  • Minor spelling error in the bottom footer ('fress' instead of 'fresh').

GPT Image 1

  • + Natural variation in lettering size and spacing.
  • + Clear and legible layout with central alignment.
  • Title is in block capitals rather than the requested 'elegant cursive'.
  • Clipped number '9' for the price of the cookies without a dollar sign.
  • The chalk texture feels slightly more digital and uniform compared to Model A.

Verdict: FLUX.2 [klein] 9B is the clear winner as it followed the stylistic instruction for 'elegant cursive' in the title and maintained a highly realistic chalkboard aesthetic including smudges and authentic chalk grain. GPT Image 1 missed the cursive requirement and had minor formatting issues with the final price listing.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Perfect adherence to the complex physical pose from Image 1.
  • + High character fidelity including facial features, sunglasses, and the specific scarf pattern.
  • + Excellent integration of lighting and environment from the source image.
  • One lens of the sunglasses is slightly distorted due to the head tilt.
  • The lower hand is clenched in a fist unlike the open hand in the source.

GPT Image 1

  • + Good color match for the yellow background and red ottoman.
  • + Successfully includes the scarf and sunglasses from Image 2.
  • Anatomy is significantly broken, particularly the legs and feet which do not align with the stool.
  • Character likeness is weak, appearing more like a caricature of the person in Image 2.
  • Poor quality on fine details like the hands and the scale of the scarf.

Verdict: FLUX.2 [klein] 9B followed the instructions excellently, maintaining a very high level of character consistency while maping the character perfectly into the difficult pose from Image 1. GPT Image 1 struggled significantly with anatomy and proportions, resulting in a distorted figure that lacked the character's likeness and the precision of the pose.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent high-detail rendering of the space background with varied celestial bodies
  • + Clear, cinematic lighting that highlights the textures of the horse and suit
  • + Solid anatomy for both the horse and the astronaut
  • Failed the negative constraint entirely by showing an astronaut riding a horse
  • Includes nonsensical AI-generated text at the bottom
  • A bit cluttered with too many planetary elements in the background

GPT Image 1

  • + Stronger atmospheric and moody lighting
  • + Good texture on the horse's coat and the weathered space suit
  • + Composition is more focused and less busy than the competitor
  • Failed the negative constraint by showing an astronaut riding a horse
  • The horse's front-left hoof has an anatomical merging issue with the leg
  • The astronaut's hand/glove on the reins is slightly malformed

Verdict: Both FLUX.2 [klein] 9B and GPT Image 1 completely failed the complex logic instruction to have the 'horse on top' (a horse riding an astronaut), instead providing the common 'astronaut riding a horse' trope. FLUX.2 [klein] 9B is the slightly better image due to its superior clarity, vibrant detail, and cleaner rendering, despite the presence of hallucinated text.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent adherence to the clothing details, featuring the correct pea coat and plaid scarf texture.
  • + Maintains the original background and identity of the person with high precision.
  • + Captures the accessories well, including the watch and jewelry.
  • The person's facial features and vitiligo patterns on the face are slightly modified compared to the source.
  • Includes extra jewelry (beads) not prominently featured in the source clothing image.

GPT Image 1

  • + Very clean and professional lighting on the clothing that blends with the subject.
  • + Successfully transfers the coat, scarf, and watch from Image 2.
  • Changes the background significantly, losing many of the wooden structure and beach details from Image 1.
  • Alters the subject's face and vitiligo patterns to the point where they no longer match the source person exactly.
  • Crops the image, losing the full-body context requested.

Verdict: FLUX.2 [klein] 9B followed the prompt much more effectively by preserving the original background and the specific details of the wooden structure, while successfully adapting the outfit to the original person's pose. GPT Image 1 generated a high-quality image but failed the source preservation requirement by changing the background and altering the subject's face too much.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [klein] 9B
GPT Image 1
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent photorealism in the capybara's fur and the woman's face.
  • + Accurately places the passenger in the back seat as requested.
  • + Creative and high-quality lighting, including the phone screen glow.
  • The woman is in the front passenger seat instead of the back seat.
  • Text on the hat is gibberish ('NEW TALA').

GPT Image 1

  • + Perfectly follows the spatial instruction of placing the woman in the back seat.
  • + Legible 'TAXI' text on the cap.
  • + Stronger cinematic mood with bokeh city lights.
  • The capybara's hands/paws on the steering wheel look slightly muddy/low-detail.
  • The passenger is quite blurry due to depth of field.

Verdict: While FLUX.2 [klein] 9B offers superior photorealism and lighting, it failed to place the passenger in the back seat, seating her beside the driver instead. GPT Image 1 followed all spatial instructions perfectly and correctly rendered the 'TAXI' text on the hat, making it the better interpretation of the prompt despite slightly lower sharpness on the capybara.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent text rendering with no spelling errors.
  • + Perfect adherence to the layout including the scroll banner location.
  • + Rich visual details in the thorny border and atmospheric lighting.
  • The lighting in the sky is slightly brighter than a 'dark parchment' aesthetic might suggest.

GPT Image 1

  • + Strong 'vintage' color palette with muted, earthy tones.
  • + The jack-o-lantern has a more menace-filled expression.
  • Failed to include 'Location:' and placed the city address in the 'Time' field.
  • The text quality is slightly fuzzy and harder to read at the bottom.
  • The border is very dark and loses the thorn detail requested in the prompt.

Verdict: FLUX.2 [klein] 9B is the clear winner as it followed all textual instructions perfectly, including specific dates and locations. GPT Image 1 suffered from a layout error where it placed the location text next to the 'Time' header and omitted the location header entirely.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [klein] 9B
Before After
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent source preservation, keeping the face and background identical to the original.
  • + The hair texture and lighting integrate perfectly with the existing environment.
  • + Natural-looking curls and realistic salt-and-pepper coloration.
  • None notable; the edit is seamless.

GPT Image 1

  • + Successfully added a large volume of hair.
  • + Preserved the background and clothing accurately.
  • The hair looks like a poorly blended wig with a visible seam above the glasses.
  • Altered the facial features slightly, particularly the bridge of the nose and forehead area.
  • The hairline looks unnatural and lacks fine-grained detail where it meets the skin.

Verdict: FLUX.2 [klein] 9B provided a near-perfect edit, seamlessly integrating a realistic head of hair while preserving every detail of the original subject's face. In contrast, GPT Image 1 produced a result that looks like a digital paste-on with visible artifacts around the hairline and minor unintended changes to the facial features.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography rendering for 'JAPAN' and 'SUSHI'.
  • + Sophisticated wood texture on the diorama base.
  • + High-quality isometric perspective and realistic material rendering.
  • The flag icon is incorrect, appearing similar to the flag of Yemen instead of Japan.
  • The sushi roll in the front has a strange structural intersection with the salmon nigiri.

GPT Image 1

  • + Correctly identifies and renders the Japanese flag icon.
  • + Very clean 'soft refined texture' style that matches the 3D cartoon request.
  • + Better sushi anatomy and recognizable garnishes like ginger and wasabi.
  • The text is slightly less crisp than Model A's bold black text.
  • The diorama base is a bit plain compared to the wood grain requested.

Verdict: Both models followed the complex prompt very well, but GPT Image 1 is the overall winner because it accurately rendered the Japanese flag icon, which FLUX.2 [klein] 9B failed to do. While FLUX.2 exhibited better text clarity and a nice wood texture, GPT Image 1's superior attention to the cultural context of the prompt and cleaner 3D character-style modeling makes it more successful.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Successfully incorporates a wide variety of dogs and hockey fans in the background.
  • + Strong caricature style with the 'big head' aesthetic often associated with the genre.
  • + Good depiction of a news/sports desk setting.
  • The facial likeness to the source image is significantly weaker than Model B.
  • Hands and fingers are poorly rendered, especially on the right hand holding the pen.
  • The eyes feel slightly too manic and disconnected from the source's expression.

GPT Image 1

  • + Excellent preservation of the subject's facial features and likeness while applying the caricature effect.
  • + High source preservation by keeping the subject in her original denim shirt outfit.
  • + Clarity of storytelling by showing the dog, the news desk, and the hockey element in a cohesive watercolor style.
  • The 'TV 13' screen in the background is a bit simplistic compared to the rest of the image.
  • The hockey stick on the desk feels like a slight afterthought in terms of placement.

Verdict: GPT Image 1 is the winner because it maintains a much stronger facial likeness to the original woman while perfectly capturing all elements of the prompt (TV anchor, dogs, and hockey). FLUX.2 [klein] 9B creates a generic caricature face that loses the identity of the source image and suffers from significant issues with hand anatomy.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent fur texture and individual hair detail
  • + Vibrant, high-contrast colors in the wildflower meadow
  • + Beautifully rendered butterflies with clear patterns
  • Completely missed the 'baby bunny' requirement in the prompt
  • The kitten's pose and expression look slightly unnatural/distorted

GPT Image 1

  • + Includes all four requested animals: puppy, kitten, fox, and bunny
  • + Captures the 'playfully chasing' and 'tumbling' dynamic much better
  • + High photorealism with soft, natural lighting and god rays
  • One butterfly in the top left is missing a body/wing connection
  • Low-resolution artifacts around the fox's whiskers

Verdict: While FLUX.2 exhibited superior texture and color vibrancy, it failed to include the baby bunny from the prompt. GPT Image 1 followed the instructions more accurately by including all four animals and capturing the action-oriented nature of 'tumbling together in a meadow,' making it the overall winner despite slightly softer details.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the original composition and poses.
  • + Captures the Ghibli-style character line work and watercolor textures accurately.
  • + Successfully transforms the urban background into a dreamy pastoral Ghibli setting.
  • Loss of the specific 'jealous/angry' facial expression on the girl in blue.
  • The man's expression is a bit too cheerful compared to the source material.

GPT Image 1

  • + Successfully retains the 'jealous' facial expression of the partner.
  • + Beautiful hand-painted crayon/pastel texture and warm lighting.
  • + Preserves the urban street layout of the original image.
  • The woman in the foreground has her eyes closed, which differs from the source.
  • Slightly less 'Ghibli' in character design compared to Model A, leaning more towards general illustration.

Verdict: Both models did an excellent job translating the 'Distracted Boyfriend' meme into a hand-drawn style. FLUX.2 [klein] 9B followed the Ghibli aesthetic more closely in terms of character design and environment transformation, though it lost some of the emotional nuance in the faces. GPT Image 1 better preserved the expressions and the urban setting of the original photo while applying a beautiful, warm texture.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [klein] 9B
Before After
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the subject's face and original clothing details
  • + Clean and distinct hair movement that looks natural and energetic
  • + High quality leaf renders that blend well with the scene's lighting
  • The leash handle shape is slightly altered from the original
  • The leaves are somewhat larger than realistic for this perspective

GPT Image 1

  • + Successfully captures hair blowing in the wind and flying leaves
  • + Maintains the overall composition and color palette of the source
  • Noticeable loss of facial detail and texture compared to the original
  • Artificial-looking artifacts on the hair strands
  • Significant distortion of the leash and the person's hand area

Verdict: FLUX.2 [klein] 9B is the clear winner as it successfully adds the requested dynamic effects while preserving the high resolution and facial features of the original image. GPT Image 1's output suffers from significant quality degradation, particularly in the subject's face and the hand holding the leash, making the edit feel less integrated.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography with correct accentuation on 'Caffè'
  • + Superior vector-style rendering with clean lines
  • + Perfect adherence to the light background and subtle texture request
  • The Cloche is slightly offset from the circular border background

GPT Image 1

  • + Classic serif typography matches the 'vintage' aesthetic well
  • + Correct inclusion of all requested elements including the cloche and banner
  • Completely ignored the request for a light background
  • Overall composition is less balanced than Model A
  • Texture is a bit heavy/grainy rather than subtle

Verdict: FLUX.2 [klein] 9B followed the prompt much more accurately, especially regarding the color scheme and the light background. GPT Image 1 generated a solid logo but failed the background requirement and has less polished vector lines compared to the clean aesthetic of FLUX.2.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [klein] 9B
GPT Image 1

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Stronger visual hierarchy with a clear narrative flow across the canvas.
  • + Detailed icons including a well-rendered lunar lander and earth.
  • + Good thematic use of background colors to represent space and the lunar surface.
  • Numerous spelling errors including 'AOILLO', 'EARDHT', 'TRANSLURAL', and 'LANDINING'.
  • The logical flow of the icons is jumbled and doesn't follow the 1-6 sequence requested.

GPT Image 1

  • + Excellent adherence to the requested NASA-inspired muted color palette.
  • + Much better text rendering with correct spelling for 'TRANQUILITY' and astronaut names.
  • + Clean, consistent flat-vector iconography that perfectly matches the requested style.
  • The layout is a bit cluttered with some overlapping text and icons.
  • Includes a typo 'EARLLUNAR' in the bottom navigation bar.

Verdict: GPT Image 1 is the superior infographic as it follows the requested flat-vector style much more closely and features significantly better text accuracy than FLUX.2 [klein] 9B. While FLUX.2 [klein] 9B has more illustrative detail, its failure to spell the primary title correctly and the disorganized flow of the steps make it less effective as an information graphic.

Next steps

Explore each model