Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1 Mini OpenAI Qwen Image Alibaba

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

GPT Image 1 Mini

24.9 arena score

#13 of 62 in Text-to-Image

Skill signature · Text-to-Image

Qwen Image

21.4 arena score

#35 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1 Mini

100.0%

win rate

Ties

0.0%

Qwen Image

0.0%

win rate

100.0% 0.0% ties 0.0%
Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent photographic texture on the book cover and paper edges.
  • + Realistic lighting and soft shadows that feel natural.
  • + Clean composition with a high level of photorealism.
  • The glass cube is missing a front face, appearing more like a frame or hollow box.
  • The plant is not clearly visible through the glass as requested.

Qwen Image

  • + Successfully renders all faces of the glass cube including reflections.
  • + Accurately places the plant behind the glass with visible refraction.
  • + Better adherence to the spatial requirements of the prompt.
  • The reflections on the glass are a bit messy and inconsistent.
  • The blue sphere has a slightly artificial, CGI look compared to Model A.

Verdict: While GPT Image 1 Mini has superior textures and lighting, it fails to render a solid glass cube, leaving the front open like a display case. Qwen Image correctly interprets the physics and spatial relationships of the prompt, showing the plant through the glass and including realistic refractions, making it the better choice for prompt adherence.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent skin texture and realistic fine details in the face and hands
  • + Strong adherence to the 'shallow depth of field' and '50mm lens' aesthetic
  • + Realistic rain drops on the bicycle and clothing
  • The 'motion blur from passing cars' is minimal and looks more like static bokeh
  • The bicycle mechanics are slightly nonsensical where the hand meets the frame

Qwen Image

  • + Better capture of the 'candid street' atmosphere with a wider framing
  • + Successful implementation of reflections on the wet pavement
  • + Good sense of the environment and weather conditions
  • Fails the 'shallow depth of field' request by keeping most of the scene relatively sharp
  • The man's feet are poorly rendered and floating slightly above the ground
  • The 'motion blur' on the car looks artificial and doesn't match the shutter speed of the rest of the image

Verdict: GPT Image 1 Mini creates a much more convincing and high-quality photograph with superior skin textures and realistic lighting. Qwen Image handles the environment and reflections well, but suffers from significant AI artifacts in the man's feet and a lack of the requested shallow depth of field.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1 Mini
Qwen Image
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1 Mini

  • + Naturalistic and lifelike eyes and skin texture.
  • + Exquisite and consistent engraving details on the plate armor.
  • + Superior warm torchlight lighting that feels integrated into the scene.
  • Missed the request for beads in the hair braids.
  • Very subtle scars that are less 'battle-worn' than requested.

Qwen Image

  • + Explicitly included the beads in the braids as requested.
  • + Strong adherence to the 'battle-worn' prompt with visible scars and blood.
  • + Clear distinction of leather straps and cloth underlayers.
  • The sparks look like artificial icons rather than natural bokeh.
  • The lighting on the face feels somewhat flat and mismatched with the bright torch in the foreground.
  • Armor engravings look slightly more generic/repetitive compared to Model A.

Verdict: GPT Image 1 Mini produces a significantly more cinematic and photorealistic outcome with superior lighting and texture, though it missed the specific detail of the hair beads. Qwen Image adhered more closely to the literal prompt elements like the beads and scars, but the overall visual quality is lower due to artificial-looking sparks and less realistic lighting. GPT Image 1 Mini is the winner for its professional composition and believable, high-quality textures.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Perfect text legibility for the section headers and main title.
  • + Clean and logical layout that aligns specific dishes with their category headers.
  • + Extremely high-quality, realistic food photography in a consistent grid.
  • Lacks actual menu item text or prices, leaving large empty spaces.
  • The design leans slightly more toward a poster than a functional menu.

Qwen Image

  • + Includes realistic menu elements like prices and item descriptions.
  • + More creative use of color blocks and varied grid sizes for a 'modern' feel.
  • + Captures the vibrant accents requested in the prompt.
  • Significant text artifacts and gibberish in the subtext and logo.
  • Redundant food photos that do not necessarily match the specific menu categories (many salads).

Verdict: GPT Image 1 Mini produced a much cleaner and more professional-looking design with perfectly legible headers and high-quality photography, though it failed to populate the actual menu details. Qwen Image attempted a more complete menu layout with prices and descriptions, but the text is garbled and the overall composition feels more cluttered. GPT Image 1 Mini is the better choice as a design template.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography with a consistent fiery, glowing effect on all text elements.
  • + High-quality photorealistic texture on the bun and patty.
  • + Clean composition that follows the 'starburst' and 'exploded' constraints accurately.
  • The 'exploded' effect is quite static and lacks the dynamic energy suggested by the prompt.
  • The tomato and sauce look slightly plastic compared to the realistic bread and meat.

Qwen Image

  • + Features a very dynamic sense of motion with flying debris and ingredients.
  • + The fiery background is more intense and visually engaging.
  • + Accurately includes all requested text and the starburst element.
  • The 'LIMITED TIME ONLY' text is plain white rather than having the requested fiery, glowing effect.
  • The burger remains mostly assembled rather than having all main components suspended separately in mid-air.
  • The starburst looks like a flat graphic compared to the 3D-feeling burger.

Verdict: GPT Image 1 Mini adhered better to the text styling and the specific 'exploded view' request, creating a polished and professional-looking advertisement. Qwen Image captured the 'dynamic' and 'fiery' atmosphere much more effectively, although it failed to apply the glow effect to all text and kept the burger mostly whole. GPT Image 1 Mini is the winner for its superior text rendering and better execution of the exploded component layout.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Perfectly follows the date requested: April 30, 2026.
  • + Shows very realistic chalk texture with dusty, grainy edges.
  • + Correctly renders all text requested in the prompt without errors.
  • The header is in print-style block capitals rather than the requested elegant cursive.
  • The layout is a bit tight against the bottom frame.

Qwen Image

  • + Excellent handwriting style that looks more natural and less like a digital font.
  • + Great environmental context showing the board inside a cozy café.
  • + Creative interpretation of the request with better layout and spacing.
  • Major error in the date, rendering '20026' instead of '2026'.
  • Spelling error in 'Risotto' (spelled 'Risoto').
  • Contains nonsensical scribbled text at the bottom right.

Verdict: GPT Image 1 Mini is the winner because it provides high accuracy in both date and spelling, whereas Qwen Image includes significant errors such as '20026' and 'Risoto'. While Qwen captures a better café atmosphere and a more fluid handwriting style, GPT Image 1 Mini's adherence to the literal text requirements makes it more usable.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent cinematic lighting and atmosphere
  • + High level of detail in the lunar surface and stars
  • + Seamless integration of the subjects into the environment
  • Failed the negative constraint; the astronaut is riding the horse

Qwen Image

  • + Bright, clear colors and sharp details
  • + Good anatomical representation of both horse and astronaut
  • Failed the negative constraint; the astronaut is riding the horse
  • The harness/reins have floating artifacts near the horse's mouth

Verdict: Both models completely failed the specific negative constraint to place the horse on top of the astronaut, instead providing standard 'astronaut riding a horse' imagery. GPT Image 1 Mini is the better image due to its superior cinematic lighting and artistic composition compared to the more generic and slightly glitched output from Qwen Image.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent cinematic lighting and photorealistic textures on the fur and coat.
  • + Very narrow focus on the passengers correctly conveys the requested bored atmosphere.
  • + The capybara's expression is very professional and human-like in its calm demeanor.
  • The steering wheel is positioned strangely low relative to the capybara's body.
  • The lighting is bit too dim to see much of the taxi interior.

Qwen Image

  • + Bright, clear composition that shows more of the New York taxi environment.
  • + Good adherence to the prompt with 'both front paws' clearly on the steering wheel.
  • + Realistic inclusion of details like a seatbelt and a name tag for the driver.
  • The text on the taxi sign says 'YOXI' instead of anything recognizable.
  • The human passenger's hand holding the phone looks slightly distorted with extra or elongated fingers.
  • The lighting on the capybara is a bit flat compared to the background.

Verdict: GPT Image 1 Mini produces a much more atmospheric and photorealistic image with superior texture rendering and cinematic lighting. While Qwen Image captures more of the taxi's exterior and follows the 'both paws' instruction more literally, it suffers from anatomical issues on the human passenger and a less convincing integration of the subjects into the scene. GPT Image 1 Mini is the better choice for its artistic quality and mood.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent text rendering with no spelling errors
  • + Consistent vintage texture and cohesive color palette
  • + Professional composition that feels like a polished invitation
  • The 'webs and thorns' border is very subtle compared to the request
  • The 'parchment' effect is less pronounced than a physical paper edge

Qwen Image

  • + Great border design featuring clear webs and thorns
  • + Stronger parchment aesthetic with torn paper edges
  • + Atmospheric background depth and twisted tree detail
  • Severe spelling errors in the main title
  • The font choice for the date and location is too modern and clean for the theme
  • The text on the scroll banner is awkwardly split and slightly garbled

Verdict: GPT Image 1 Mini is the clear winner because it successfully renders all the specific text accurately, which is crucial for an invitation. While Qwen Image has superior atmospheric details and a more creative border, the significant typos in the header and the use of a generic sans-serif font for the event details break the vintage gothic theme.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography and flag icon integration.
  • + Higher quality textures and more realistic lighting.
  • + Clean, professional 3D render look.
  • The larger base feels a bit simple compared to a 'miniature diorama' concept.

Qwen Image

  • + Captures the 'miniature diorama' and 'small raised base' prompt elements better.
  • + Includes extra details like the flag on the base and garnishes.
  • + Good adheresence to the isometric perspective.
  • The text rendering is slightly less clean than Model A.
  • Lighting is a bit flat compared to the PBR request.
  • The flag icon next to the text is slightly distorted.

Verdict: GPT Image 1 Mini produced a much cleaner and more professional-looking graphic with superior text rendering and material textures. While Qwen Image better captured the 'miniature diorama' theme by building a more complex base, the overall visual clarity and execution of GPT Image 1 Mini make it the stronger choice for the specific high-clarity request.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent sense of motion and action with multiple animals mid-jump.
  • + Beautifully rendered backlighting and god rays filtering through the scene.
  • + Strong anatomical accuracy for all four distinct animals.
  • The fox's front right paw is somewhat indistinct against its body.

Qwen Image

  • + High level of detail on the butterflies and dew sparkles in the foreground.
  • + Center-focused composition creates a clear focal point.
  • + Soft, pleasing lighting that captures the 'wholesome' vibe.
  • The dog appears static and disconnected from the play occurring around it.
  • The fox's ears and head shape lean slightly towards a feline appearance.
  • The butterflies appear somewhat flat/pasted on compared to the environment.

Verdict: GPT Image 1 Mini captured the 'playfully chasing' and 'tumbling' aspect of the prompt much more effectively, as all four animals show dynamic motion. While Qwen Image has lovely foreground details and lighting, the puppy's static pose breaks the energy of the scene. GPT Image 1 Mini is the winner for its superior composition and portrayal of the requested activity.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography with correct spelling and accents
  • + Solid vector-style layout that is easy to read
  • + High contrast and sharp line work
  • Ignored the request for a light background
  • The gold/brown tones are a bit grainy

Qwen Image

  • + Followed the color palette and background request perfectly
  • + Good 'hand-drawn' vintage feel with nice texture
  • + Strong banner design for the establishment date
  • Severe spelling and layout error with the name 'Florian'
  • The steam lines are slightly asymmetrical

Verdict: GPT Image 1 Mini produced a much more professional and usable logo by handling the text perfectly, though it failed to provide the requested light background. Qwen Image captured the aesthetic and colors better, but the significant spelling error ('FLOraiAN') makes the logo unusable for its intended purpose.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1 Mini
Qwen Image

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent text rendering with no spelling errors.
  • + Consistent and clean vector illustration style across all icons.
  • + Logical flow that follows all six requested steps in order.
  • The 'Translunar' icon is a bit abstract and messy compared to the others.
  • The layout is somewhat fragmented and lacks a cohesive title at the top.

Qwen Image

  • + High visual appeal with a strong NASA-inspired dark background composition.
  • + Great inclusion of the NASA logo and mission title for a poster feel.
  • + Good use of the requested color palette.
  • Significant spelling errors including 'Tranar Orbit' and 'Aldin'.
  • Failed to include all 6 steps correctly, jumping numbers and mislabeling stages.
  • Includes a literal instruction '1 Stop at landing' as text in the image.

Verdict: GPT Image 1 Mini is the clear winner for its superior text accuracy and strict adherence to the six-step sequence requested. While Qwen Image has a more striking aesthetic and better 'poster' layout, it suffers from several typos and fails to follow the logical progression of the mission steps.

Next steps

Explore each model