OpenAI's cost-effective image generation model for when image quality isn't the top priority
Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Z-Image Turbo
#12 of 62 in Text-to-Image
Where the votes landed
GPT Image 1 Mini
37.5%
win rate
Ties
12.5%
Z-Image Turbo
50.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent photographic realism and soft lighting.
- + Accurate glass thickness and realistic refractions.
- + Book texture is highly detailed and convincing.
- − The blue sphere is large and appears to be levitating rather than sitting inside the cube naturally.
- − The plant is behind the cube but doesn't show much distortion through the glass.
Z-Image Turbo
- + The sphere size is more 'small' relative to the cube as requested.
- + The plant is visible through the glass with realistic refraction.
- + The cube has a reflective base plate which adds to the visual complexity.
- − The book appears physically detached and floating slightly above the glass cube.
- − The perspective of the cube's top surface is slightly skewed.
- − Lighting is a bit flatter compared to the atmospheric quality of the other image.
Verdict: GPT Image 1 Mini produced a visually superior image with better lighting and textures, though it failed on the scale of the sphere and its physical placement. Z-Image Turbo followed the 'small' sphere instruction better and showed the plant through the glass more effectively, but it suffered from a major structural error where the book is floating. GPT Image 1 Mini is preferred for its significantly higher aesthetic quality and cohesive scene construction.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent photographic quality with a shallow depth of field and soft cinematic lighting.
- + Realistic skin textures and weather effects, including visible rain and wet pavement reflections.
- + Accurately captures the 'repairing' action in a natural, candid-style composition.
- − The bike's mechanical details (spokes and chain area) become surreal and tangled toward the rear wheel.
Z-Image Turbo
- + Good inclusion of background traffic as requested in the prompt.
- + Clear, realistic subjects with natural lighting.
- − Does not show the subject 'repairing' the bike; he is simply holding or walking it.
- − Lacks the cinematic shallow depth of field and 'imperfect framing' requested in the prompt.
- − The rain effect and reflections are much less noticeable and lower quality compared to the other model.
Verdict: GPT Image 1 Mini feels like a high-end cinematic photograph, successfully capturing the texture of the rain, the mood of the lighting, and the specific action of repairing the bicycle. While Z-Image Turbo captures the background traffic well, it fails to depict the core action of the prompt and has a much flatter, less professional aesthetic. GPT Image 1 Mini is the clear winner for its superior atmospheric rendering and adherence to the 'cinematic but realistic' instruction.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent detailed engraving on the plate armor
- + Subtle and realistic skin texture with believable scars and dirt
- + Atmospheric warm lighting that feels integrated into the scene
- − Missed the request for small beads in the braided hair
- − The leather and cloth underlayers are mostly obscured and less detailed
Z-Image Turbo
- + Perfect adherence to the beads in the hair requirement
- + High contrast lighting with visible leather straps and chainmail/cloth layers
- + Dynamic sparks around the torch provide extra visual interest
- − The torch and sparks look somewhat digitally overlaid rather than naturally integrated
- − Facial scars look more like fresh paint/blood than healed battle scars
Verdict: Z-Image Turbo followed the prompt more precisely by including the specific detail of beads in the hair and providing visible leather/cloth layers. However, GPT Image 1 Mini achieved a much more cohesive and realistic visual quality, particularly in the subtle rendering of the skin and the intricate engravings on the armor.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfect text rendering for all requested section titles.
- + Clean, professional grid layout that follows the prompt instructions exactly.
- + High-quality, appetizing food photography with consistent lighting.
- − The menu contains no item names or prices, leaving large empty spaces.
- − Somewhat generic corporate aesthetic.
Z-Image Turbo
- + More complex layout including prices and item placeholders.
- + Vibrant orange accents create a strong visual identity.
- + Good use of the grid for food photos.
- − Significant text errors including 'PIZZA MANS' and 'SETIIION'.
- − The sections (Appetizers, Pizza, Mains) were merged or confused in the layout.
- − Background of the food photos is dark grey rather than the requested white background for the design.
Verdict: GPT Image 1 Mini adhered much better to the specific layout requirements and provided flawless text rendering, despite leaving the menu items blank. Z-Image Turbo attempted a more complete menu but suffered from significant spelling errors and failed to implement the 'Mains' category correctly, rendering it as 'PIZZA MANS'.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfect adherence to the 'exploded' layout requested in the prompt.
- + Excellent text rendering with a consistent fiery, glowing effect.
- + High-quality, photorealistic textures on the bun and patty.
- − The composition feels slightly empty in the mid-ground compared to the busy background of the competitor.
Z-Image Turbo
- + Vibrant, energetic background with visible flames and embers.
- + Strong lighting on the burger components that matches the environment.
- + Detailed texture on the patties and fresh-looking vegetables.
- − Failed to create an 'exploded' view; the burger is mostly assembled and tilted.
- − The text lacks the requested 'fiery effect' and looks more like a standard 3D font.
- − Minor artifacts in the sauce droplets at the bottom.
Verdict: GPT Image 1 Mini followed the prompt instructions much more accurately, specifically capturing the 'exploded' view where all components are suspended individually. While Z-Image Turbo created a more visually dense and dynamic background, it failed the core compositional requirement of an exploded burger and delivered standard text instead of the requested fiery glowing effect. GPT Image 1 Mini is the clear winner for its superior prompt adherence and cleaner typography.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfectly accurate spelling for all items and dates.
- + Excellent chalk texture with realistic grainy gradients within the strokes.
- + Natural-looking spacing and framing within a physical chalkboard frame.
- − The title is in print style rather than the requested 'elegant cursive chalk handwriting'.
Z-Image Turbo
- + Captures a very convincing chalk-on-blackboard aesthetic with smudges and erasing artifacts.
- + Good attempts at varied handwriting styles.
- − Includes a spelling error ('Mustroom' instead of 'Mushroom').
- − The handwriting is more of a generic casual print than the requested cursive for the title.
Verdict: GPT Image 1 Mini is the clear winner because it successfully spelled every word in the complex prompt correctly, whereas Z-Image Turbo failed on 'Mushroom'. While neither model fully delivered 'elegant cursive' for the title, GPT Image 1 Mini's superior text rendering and better alignment with the requested menu items make it the more useful image.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1 Mini
- + High visual quality and atmospheric, cinematic lighting.
- + Excellent integration of the horse and astronaut into the starry environment.
- − Fails to follow the specific spatial instruction to put the horse on top of the astronaut.
Z-Image Turbo
- + Clear, bright details on the spacesuit and horse tack.
- + Good anatomical rendering of the horse legs.
- − Fails to follow the negative constraint; the astronaut is still on top of the horse.
- − The background is less cinematic and feels like a simple backdrop.
Verdict: Both models failed the specific prompt instruction to place the horse on top of the astronaut, instead defaulting to the standard 'astronaut riding a horse' trope. GPT Image 1 Mini is superior in terms of artistry, lighting, and composition, whereas Z-Image Turbo looks more like a composite of two separate images.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent cinematic lighting that accurately captures a night scene.
- + Good focus on the capybara's expression and professional attire.
- + Better background bokeh that suggests Manhattan at night.
- − Only one paw is visible on the steering wheel, missing the 'both paws' requirement.
- − Lighting is a bit too dark on the human passenger.
Z-Image Turbo
- + Follows the 'both paws on the steering wheel' instruction perfectly.
- + Clear rendering of both subjects and the car interior.
- + The capybara's hands are rendered with surprising detail, looking like a mix of paw and hand to grip the wheel.
- − The lighting looks too bright for a night scene, appearing more like dawn or dusk.
- − The human passenger is seated strangely, looking more like a co-pilot than someone in the back seat due to the perspective.
- − Internal car geometry is slightly confused between the front and back rows.
Verdict: GPT Image 1 Mini creates a much more atmospheric and photorealistic night scene with superior lighting, but fails to show both paws on the wheel. Z-Image Turbo adheres better to the specific literal instructions regarding the paws, but the spatial arrangement and lighting make the passenger look like she is sitting next to the driver rather than in the back seat. GPT Image 1 Mini is the preferred choice for its cinematic quality and better interpretation of a New York taxi environment.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfectly rendered text with no spelling errors.
- + Consistent vintage aesthetic with a dark, moody parchment feel.
- + Excellent composition with a strong central focus on the glowing jack-o-lantern.
- − The thorns in the border are very subtle and blend into the dark background.
- − The scroll banner is more of a flat ribbon than a distinct scroll.
Z-Image Turbo
- + Creative use of layering with the torn parchment effect over the background.
- + Strong inclusion of all requested elements including webs and prominent thorns.
- + Gothic typography matches the 'elegant' part of the prompt well.
- − Contains a spelling error in the location text ('Archves' instead of 'Arches').
- − The scroll banner formatting is a bit messy, appearing as three separate pieces.
- − Lighting is a bit less cinematic and more flat compared to the other model.
Verdict: GPT Image 1 Mini is the winner because it successfully followed every text instruction without spelling errors and maintained a cohesive, polished vintage aesthetic. Z-Image Turbo had a more dynamic layout but failed on detail accuracy, specifically misspelling the location 'The Arches'.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully added a full, thick head of hair as requested.
- + Matches the hair color well to the existing beard.
- + Maintains the overall composition and lighting profile.
- − Significantly altered the facial features, making the man look like a different person.
- − The hairline looks slightly artificial where it meets the forehead.
Z-Image Turbo
- + Excellent preservation of the subject's original facial features.
- + High visual quality with realistic skin texture.
- − Failed the primary edit instruction by only adding a very thin buzz cut instead of a full, thick head of hair.
- − Removed the subject's glasses without being asked.
Verdict: GPT Image 1 Mini followed the instructions to add a full head of hair, whereas Z-Image Turbo only provided a thin buzz cut. However, GPT Image 1 Mini changed the man's face significantly, while Z-Image Turbo preserved the person's likeness and background much better despite failing the main prompt and removing the glasses.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Accurate Japanese flag icon.
- + Excellent text rendering and spacing.
- + Higher level of detail with multiple sushi types and chopsticks.
- − The 'JAPAN' text is slightly off-center to the left.
Z-Image Turbo
- + Clean 3D cartoon aesthetic.
- + Good adherence to the isometric perspective.
- − Displays the flag of China instead of Japan.
- − Text is less refined and slightly crowded.
- − Only shows a single piece of sushi which feels less complete.
Verdict: GPT Image 1 Mini is the clear winner as it correctly identifies the Japanese flag and provides a much more polished and detailed scene including multiple sushi varieties and chopsticks. Z-Image Turbo failed the geographical alignment by placing a Chinese flag next to the word 'Japan' and provided a much simpler model.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully incorporates all elements: microphone, news desk, hockey stick, and dog.
- + Excellent caricature style with exaggerated features that remain recognizable.
- + Thoughtful composition that balances professional and personal interests.
- − The fingers on the hand holding the microphone have some anatomical distortion.
- − The colored pencil texture is a complete stylistic shift from the source photo.
Z-Image Turbo
- + Successfully preserves the subject's facial identity and the original photographic style.
- + Adds a small dog in the background consistent with the original room.
- − Fails to apply the requested 'caricature' and 'exaggerated' artistic style.
- − Completely misses the 'hockey' and 'tv show anchor' elements of the prompt.
- − The edit is too subtle for the specific creative request.
Verdict: GPT Image 1 Mini followed the creative brief perfectly, creating a humorous caricature that integrated every requested element (news, dog, and hockey). While Z-Image Turbo did a better job of preserving the woman's actual face, it failed to perform the most important parts of the edit, ignoring the occupation and hobby requirements entirely.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfectly adheres to the list of animals requested.
- + Excellent lighting with visible god rays and convincing dew sparkles.
- + Highly detailed fur textures and dynamic action-oriented composition.
- − The butterfly on the left has slightly simplified wing patterns.
- − The puppy's left ear has a slightly unnatural shape at the top.
Z-Image Turbo
- + Captures the 'tumbling together' aspect of the prompt very well.
- + Cute expressions on the animals' faces.
- + Vibrant butterfly colors.
- − Anatomical issues, particularly the kitten's limb merging into the puppy's fur.
- − The fox has a white-tipped tail that looks slightly disconnected or artifacted.
- − The 'god rays' are less defined compared to the other model.
Verdict: Both models captured the essence of the prompt, but GPT Image 1 Mini is the clear winner due to its superior anatomical accuracy and lighting effects. While Z-Image Turbo created a charming group hug, it suffered from significant merging artifacts between the animals' bodies, whereas GPT Image 1 Mini maintained distinct, high-quality textures for each creature.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1 Mini
- + Strong adherence to the Studio Ghibli illustration style.
- + Excellent translation of facial expressions into an anime-inspired aesthetic.
- + Effective use of hand-painted textures and a nostalgic color palette.
- − The color saturation is a bit monochromatic (very yellow/orange).
- − The foreground woman's face is slightly sharp compared to the blurry source, losing some depth of field contrast.
Z-Image Turbo
- + Perfect source preservation of the original image's composition and subjects.
- + High resolution and clarity.
- − Completely failed to apply the requested artistic style edit.
- − The image remains a photograph rather than an illustration.
Verdict: GPT Image 1 Mini successfully transformed the photograph into a clear Studio Ghibli–inspired illustration while maintaining the recognizable character interactions and composition of the meme. Z-Image Turbo essentially ignored the stylistic instructions and provided a slightly sharpened version of the original photo. GPT Image 1 Mini is the clear winner for following the creative transformation prompt.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1 Mini
- + Strong prompt adherence with clearly blowing hair and numerous flying leaves.
- + Maintains the pose and color palette of the original subject well.
- + Includes movement in the dog's fur and the leash to enhance the 'energetic' feel.
- − The woman's facial features were altered significantly from the source image.
- − Significant background changes, such as the removal of the bridge reflection and white flowers.
Z-Image Turbo
- + Excellent source preservation, keeping the woman's face nearly identical to the original.
- + Subtle and realistic hair movement that feels integrated into the scene.
- + Maintains the original background details, including the bridge and vegetation, with more accuracy.
- − The flying leaves are sparse and less dynamic than requested.
- − Missed the opportunity to add motion to the dog or the leash.
Verdict: GPT Image 1 Mini was more successful at capturing the 'energetic and lively' instruction by adding significant wind effects to the hair and many flying leaves, though it failed to preserve the woman's facial identity. Z-Image Turbo excelled at preserving the original image's look and feel but the requested edits were too subtle to truly satisfy the demand for dynamic motion. GPT Image 1 Mini is preferred for its better commitment to the creative prompt despite the face change.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with correct inclusion of the accent mark in 'Caffè'
- + Great texture and distressed metallic effects
- − Failed to provide a light background as requested
- − The cloche illustration is a bit cluttered with uneven shading marks
Z-Image Turbo
- + Perfect adherence to the light background and minimalist vector style
- + Clean, professional layout that feels like a real logo
- + Accurate spelling and typography
- − Missed the request for a banner around 'Est. 1720'
- − The cloche icon is somewhat generic and lacks the 'retro' character of the first model
Verdict: Z-Image Turbo followed the background and style instructions more accurately, resulting in a cleaner vector-style logo that is better suited for actual use. While GPT Image 1 Mini had better texture and a more accurate banner, it completely ignored the request for a light background, making Z-Image Turbo the overall winner for prompt adherence.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text rendering with no spelling errors.
- + Includes all six requested steps in a logical narrative flow.
- + Perfect adherence to the flat-vector style and NASA color palette.
- − The 'Translunar' icon is a bit cluttered with overlapping lines.
- − The composition is slightly bottom-heavy due to cut-off elements at the base.
Z-Image Turbo
- + Features a clean, centered layout.
- + Captures the modern vector aesthetic well.
- − Significant spelling errors throughout, including 'APOLIO E 11' and 'Descenty'.
- − Failed to include several requested steps, missing the numbered list structure.
- − The rocket design is generic and does not resemble the Saturn V requested.
Verdict: GPT Image 1 Mini is the clear winner as it successfully rendered all six requested steps with perfect spelling and a cohesive narrative layout. In contrast, Z-Image Turbo struggled with basic text tasks and missed more than half of the specific content requested for the infographic.
Explore each model
Tongyi-MAI's 6-billion parameter distilled text-to-image model optimized for speed, achieving high-quality generation in 8 steps or fewer with support for bilingual text rendering