Head to head
Esc

Models · slot A

to navigate to pick

Wan 2.7 Alibaba Z-Image Turbo Alibaba

Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.

Wan 2.7

20.5 arena score

#39 of 62 in Text-to-Image

Skill signature · Text-to-Image

Z-Image Turbo

25.3 arena score

#12 of 62 in Text-to-Image

Vote tally

Where the votes landed

Wan 2.7

0.0%

win rate

Ties

0.0%

Z-Image Turbo

100.0%

win rate

0.0% 0.0% ties 100.0%
Shared challenges 17

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent realism with high-quality textures on the wood and book.
  • + Accurate physical properties with reflections and refractions through the glass cube.
  • + Follows all spatial instructions perfectly including positioning of the sphere and plant.
  • The glass cube has some internal structural lines that look slightly inconsistent.

Z-Image Turbo

  • + Clean, minimalist composition with a high-gloss sphere.
  • + Accurate adherence to the basic prompt elements.
  • The plant is too blurry and not clearly visible 'through' the glass as requested.
  • Lighting is a bit flat compared to Model A.
  • The glass cube physics look a bit simplified and less realistic.

Verdict: Wan 2.7 is the winner due to its superior realism and detail, particularly in the way the environmental light and the plant interact with the glass cube. While Z-Image Turbo followed the prompt correctly, it has a more generic aesthetic with less convincing textures and depth.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent atmosphere with realistic reflections and wet textures.
  • + Captures the requested 'imperfect framing' and 'candid' feel perfectly.
  • + Highly realistic skin and clothing textures for the main subject.
  • The hands merging with the bicycle handlebars show some anatomical merging.
  • Slightly lacks the 'motion blur' requested for the passing cars.

Z-Image Turbo

  • + Successfully incorporates motion blur in the background vehicles.
  • + Clear subject focus and bright, saturated colors on the red bicycle.
  • The character looks less like he is 'repairing' and more like he is simply standing with the bike.
  • Visual quality is lower, with a 'plastic' skin texture and less realistic lighting.
  • Lack of depth in the environment compared to Model A.

Verdict: Wan 2.7 is the clear winner as it perfectly captures the cinematic, candid street photography aesthetic requested. While Z-Image Turbo followed the motion blur instruction better, it failed to deliver the natural skin textures and atmospheric realism that Wan 2.7 achieved with its superior handling of lighting and reflections.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent detail on the engraving and worn textures of the armor.
  • + Highly lifelike facial features and realistic skin texturing.
  • + Great implementation of beads within the braids.
  • The bokeh sparks are a bit large and slightly distracting.
  • The torch in the background is somewhat clipped by the frame.

Z-Image Turbo

  • + Atmospheric lighting with a clear source for the torchlight glow.
  • + Beautifully intricate chainmail and cloth underlayer textures.
  • + Strong sense of shallow depth of field and dynamic sparks.
  • Skin texture feels slightly smoothed or airbrushed compared to Model A.
  • Facial features are less detailed than the armor and surroundings.

Verdict: Wan 2.7 provides a superior close-up portrait with much higher skin and eye realism, capturing the 'battle-worn' aspect through more convincing scars and skin texture. Z-Image Turbo has excellent costume detail and lighting atmosphere, but the character's face lacks the lifelike grit and intensity found in the Wan 2.7 generation.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent typography with high readability and correct spelling.
  • + Professional and balanced layout that mimics a real-world menu.
  • + Clear sections for appetizers, pizza, and mains as requested in the prompt.
  • The surrounding lifestyle objects (pen, oil, salt) were not explicitly requested, though they add to the presentation.
  • Some minor formatting artifacts in the small description text.

Z-Image Turbo

  • + Follows the grid layout requirement for photos effectively.
  • + Clean minimalist aesthetic with bold sans-serif fonts.
  • + Vibrant color palette in the food photography.
  • Numerous spelling errors including 'PIZZA MANS' and 'SETIIION'.
  • The text sections are jumbled and less organized than image A.
  • Smaller selection of food items compared to the overall layout size.

Verdict: Wan 2.7 is the clear winner as it produces a functional, professional menu with accurate spelling and a layout that logically follows the user's categories. Z-Image Turbo adheres well to the grid photo requirement but fails significantly on text rendering and general coherent design, resulting in several typos.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent adherence to the 'exploded' concept with all ingredients clearly separated.
  • + Strong typography that integrates perfectly with the fiery theme.
  • + High levels of detail in the food textures and dynamic splash effects.
  • The 'seeds' or nuts floating in the background feel slightly disconnected from a standard burger.

Z-Image Turbo

  • + Beautiful warm lighting and clean, readable text.
  • + Nice use of depth of field with the fiery foreground elements.
  • Failed to provide an 'exploded' view, showing a mostly assembled burger instead.
  • The burger patties look a bit more like generic meatballs or thick ground meat clumps rather than flat grill-marked patties.

Verdict: Wan 2.7 followed every aspect of the prompt, particularly the difficult 'exploded' layout and specific fiery text effects. Z-Image Turbo produced a high-quality ad image, but failed to separate the burger components as requested and had slightly less creative typography.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent layout and overall visual aesthetic.
  • + Higher image quality and better environmental context.
  • + Completes the third menu item correctly despite the prompt being cut off.
  • The font looks like a clean digital script rather than actual chalk handwriting.
  • The text has a consistent drop shadow and uniform thickness that feels artificial.
  • Missing 'elegant cursive' for the title.

Z-Image Turbo

  • + Persuasive chalk texture with realistic powdery edges.
  • + Captures the 'handwritten' request much better with natural letter variations.
  • + Adheres to the specific request for chalk texture instead of a digital overlay.
  • Contains a spelling error ('Mustroom').
  • The composition is a bit more cramped with the text nearly touching the edges.
  • Slightly less clarity in the overall image compared to model A.

Verdict: Wan 2.7 produces a more visually polished image, but fails the negative constraint by using what appears to be a digital font with artificial shadows. Z-Image Turbo much more accurately captures the requested realistic chalk handwriting style and texture, making it the better choice for this specific prompt despite a minor spelling error.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent adherence to the 'cinematic' and 'detailed' descriptors with vibrant background elements.
  • + Highly detailed texture on the space suit and horse's coat.
  • + Strong anatomical consistency with the horse's legs and the astronaut's posture.
  • The horse's hind legs are slightly elongated and awkward near the bottom edge.

Z-Image Turbo

  • + Successfully captures the surreal nature of the prompt.
  • + Detailed rendering of the saddle and riding equipment.
  • The background is very sparse and lacks the cinematic quality requested.
  • The horse's front legs exhibit anatomical warping and strange proportions.
  • Noticeable lighting inconsistency between the astronaut and the environment.

Verdict: Wan 2.7 is the clear winner as it provides a much more cinematic and visually rich environment with superior anatomical rendering. While both models followed the specific positioning instruction, Z-Image Turbo produced a flat, dark background and significant distortions in the horse's legs.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

Wan 2.7
Z-Image Turbo
0% wins 0% ties 100% wins

AI Judge Analysis

Wan 2.7

  • + Excellent photorealism in the textures of the capybara's fur and the jacket fabric.
  • + Successfully placed the human passenger in the back seat as requested.
  • + The lighting and reflections on the glass and car body are very realistic for a nighttime city scene.
  • The passenger is sitting in the right-side seat, making it look like she is in the front passenger seat rather than the back.
  • The capybara's claws/paws are slightly distorted as they grip the wheel.

Z-Image Turbo

  • + The capybara's expression and posture perfectly match the 'calm, professional' description.
  • + Good separation of foreground and background with city light bokeh.
  • + Accurate rendering of the driver's cap and uniform jacket.
  • The passenger appears to be in the front passenger seat, failing the 'back seat' part of the prompt.
  • The perspective of the steering wheel and the driver's arms is slightly awkward and lacks depth.
  • Low visibility of the 'New York' setting outside compared to Model A.

Verdict: Wan 2.7 provides a much more detailed and photorealistic image with convincing textures and complex city reflections. While both models failed to clearly place the passenger in the back seat (both appearing to put her in the front next to the driver), Wan 2.7's superior rendering of the taxi's exterior and interior environment makes it the stronger choice.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Perfect text rendering for all requested fields including the specific date and location
  • + Highly detailed and decorative border with thorns, webs, and skulls
  • + Excellent composition with a focused central illustration and balanced layout
  • The style leans more toward vintage illustration than cinematic realism
  • The 'Est. 1847' text was not requested in the prompt

Z-Image Turbo

  • + Captures a more cinematic and moody lighting style
  • + The torn parchment aesthetic provides a nice sense of depth and texture
  • Misspelled location as 'The Archves'
  • The scroll banner text is placed at the top rather than near the banner element
  • The overall layout feels a bit cluttered with overlapping elements

Verdict: Wan 2.7 is the clear winner as it followed every instruction perfectly, including complex text rendering with zero spelling errors. While Z-Image Turbo captured the cinematic lighting well, it failed on a key detail by misspelling the location name and lacked the polished, professional layout found in Wan 2.7.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
Wan 2.7
Before After
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent adherence to the request for a full, thick head of hair.
  • + Highly realistic texture and natural integration with existing facial hair.
  • + Perfect preservation of the facial features, glasses, and background.
  • None notable.

Z-Image Turbo

  • + Successfully modified the hairstyle to a buzzed look.
  • Failed to provide a 'full, thick head of hair' as requested.
  • Removed the subject's glasses which was not requested.
  • Significantly altered facial features and eye squint, failing to preserve the person's likeness.

Verdict: Wan 2.7 perfectly executed the complex task of adding realistic hair while maintaining a 1:1 match of the original person's face, glasses, and the surrounding environment. Z-Image Turbo failed on almost every instruction, delivering a short buzz cut instead of thick hair, and completely changing the person's facial structure and removing their eyewear.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent text rendering and alignment.
  • + Correct flag icon for Japan.
  • + Complex and varied 3D sushi models with high-quality textures.
  • The white napkin/paper details on the left look slightly flat.

Z-Image Turbo

  • + Good 3D soft textures and lighting.
  • + Clean isometric diorama base.
  • Incorrectly used the flag of China instead of Japan.
  • Text alignment is slightly off-center.
  • Very simplistic content with only one piece of sushi.

Verdict: Wan 2.7 is the clear winner as it followed all instructions, including the correct flag and high-quality text placement. Z-Image Turbo failed a key geographical requirement by placing a Chinese flag in a scene labeled 'Japan', and the single piece of sushi felt sparse compared to Wan 2.7's full platter.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent adherence to the 'caricature' style with exaggerated features and comic aesthetics.
  • + Successfully incorporates all requested elements: TV anchor (studio setting/microphone), dogs, and hockey (puck/mini-rink).
  • + Maintains the recognizable composition and basic facial structure of the subject in an artistic way.
  • Minor spelling error in the speech bubble text ('Rolee' instead of 'Rule').

Z-Image Turbo

  • + High fidelity to the original source image's realism.
  • + Subtly adds a small dog in the background.
  • Completely fails the core instruction to create an exaggerated caricature.
  • Missing the hockey and TV anchor profession elements entirely.
  • The changes made are so minimal they are hard to notice.

Verdict: Wan 2.7 followed the complex editing instructions perfectly, transforming the photo into a vibrant, humorous caricature that included all three specific thematic elements requested. In contrast, Z-Image Turbo failed to provide a caricature and missed nearly all of the requested content, only adding a small dog to the background of the original photo.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent depiction of god rays and sunrise lighting consistent with the prompt.
  • + Sharp, high-resolution fur textures on all animals.
  • + Included all four requested animals with distinct, accurate features.
  • The kitten's pose and tongue look slightly unnatural and AI-generated.
  • The butterflies appear flattened against the background.

Z-Image Turbo

  • + Successfully captured the 'tumbling together' aspect of the prompt with physical interaction between the animals.
  • + Faces are more expressive and traditionally 'cute'.
  • + Better balance of the animals as one cohesive group.
  • Lacks the 'god rays' lighting effect requested in the prompt.
  • The fox kit has a slightly distorted front leg/paw structure.
  • Less overall detail in the background wildflowers compared to Model A.

Verdict: Wan 2.7 is the winner for its superior technical execution of the lighting effects and 8K detail, particularly in the grass and fur. While Z-Image Turbo captured a more playful and physically connected composition, it failed to incorporate the specific 'god rays' lighting requested and had more noticeable anatomical artifacts.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent adherence to the 'hand-painted' and 'watercolor' Ghibli aesthetic.
  • + Maintains the composition and iconic poses of the original meme perfectly.
  • + Good use of soft pastel colors and warm lighting for a nostalgic feel.
  • The woman in the foreground is rendered in full detail, losing the depth-of-field blur present in the original.

Z-Image Turbo

  • + Preserved the depth-of-field blur on the foreground character.
  • + Successfully changed the faces while maintaining high resolution.
  • Completely failed the stylistic instruction to create a Ghibli-inspired illustration.
  • The output is essentially just another photograph with slightly different faces.

Verdict: Wan 2.7 successfully executed the complex stylistic transformation, turning the meme into a beautiful watercolor illustration that reflects the requested Ghibli aesthetic. In contrast, Z-Image Turbo largely ignored the creative part of the prompt, returning a standard photograph that lacks any illustrative quality.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
Wan 2.7
Before After
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent preservation of the subject's face and original features
  • + Realistic blowing hair effect that matches the wind direction of the leaves
  • + Successfully added a large number of falling leaves to create a dynamic feel
  • Some leaves appear a bit blurry or low-resolution compared to the rest of the scene

Z-Image Turbo

  • + Successfully added falling leaves and some movement to the hair
  • + Maintains a clean aesthetic comparable to the original image
  • Noticeably changed the woman's face and features from the source image
  • Altered the position of the dog's leash and the dog's tail structure
  • Hair movement is less dramatic and 'dynamic' than Model A

Verdict: Wan 2.7 is the clear winner as it successfully applied all requested edits—dynamic hair and flying leaves—while perfectly preserving the woman's face and the dog's appearance from the source image. Z-Image Turbo failed at source preservation, significantly altering the woman's facial features and making unnecessary changes to the leash and dog.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent typography rendering with the requested accent mark.
  • + Comprehensive inclusion of all elements like the cloche, steam, banner, and vintage texture.
  • + Strong vintage emblem aesthetic with balanced decorative elements.
  • Small spelling error in the primary text ('Florion' instead of 'Florian').
  • The composition is a bit crowded for a 'minimalist' request.

Z-Image Turbo

  • + Perfect spelling of all requested text.
  • + Captures the 'minimalist' aspect of the prompt much better than Model A.
  • + Clean vector style that is appropriate for a modern logo application.
  • Missed the 'banner' requirement for the 'Est. 1720' text.
  • The cloche illustration is very basic compared to the vintage aesthetic requested.

Verdict: Wan 2.7 provides a much more detailed and aesthetically pleasing 'vintage' design that feels authentic to the time period, despite a small typo in the name and a slightly busy layout. Z-Image Turbo captures the minimalism and spelling perfectly but fails to include the requested banner and feels a bit plain. Wan 2.7 is the preferred choice for its superior texture, composition, and adherence to the 'vintage emblem' style.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

Wan 2.7
Z-Image Turbo

AI Judge Analysis

Wan 2.7

  • + Excellent adherence to the infographic structure, including all 6 requested steps.
  • + Very clean typography with surprisingly good spelling for an AI (Tranquility, Armstrong, Aldrin, Collins).
  • + Strong vector aesthetic with a professional vertical layout.
  • One minor spelling error in 'DESCRIPT' instead of Descent.
  • The background stars are slightly cluttered, making some text a bit harder to read.

Z-Image Turbo

  • + Features a distinct and clean flat-vector illustration style.
  • + The colors are bright and adhere to the requested NASA-inspired palette.
  • Fails to include several of the requested steps (Launch icon is present but the overall flow is missing).
  • Poor spelling including 'APOLIO E 11', 'Translurian', and 'Descenty'.
  • Weak composition with icons floating awkwardly in white space without a clear timeline.

Verdict: Wan 2.7 produced a comprehensive and highly professional infographic that perfectly followed all the requested steps in a logical sequence. Z-Image Turbo failed on both layout and text accuracy, containing multiple spelling errors and missing several steps of the mission timeline. Wan 2.7 is the clear winner for its superior composition and information hierarchy.

Next steps

Explore each model