Vyro AI's professional-grade text-to-image model delivering photorealistic output with accurate text rendering and typography precision for commercial workflows
Settled by community votes across 7 shared challenges, with an AI judge weighing in on each.
ImagineArt 1.5 (Preview)
#6 of 62 in Text-to-Image
Wan 2.5 (Preview)
#28 of 62 in Text-to-Image
Where the votes landed
ImagineArt 1.5 (Preview)
0.0%
win rate
Ties
0.0%
Wan 2.5 (Preview)
100.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent handling of glass refraction and reflections on the table.
- + Highly realistic texture on the wooden table and the vintage book.
- − The plant is positioned above the cube rather than clearly behind/visible through it.
- − The glass cube looks more like a solid block of glass rather than a hollow container.
Wan 2.5 (Preview)
- + Perfect adherence to the spatial prompt, with the plant clearly visible through the glass.
- + The lighting is more atmospheric with visible dust motes and soft window shadows.
- − The blue sphere looks somewhat flat and lacks realistic material depth compared to the surroundings.
- − The red book appears to be slightly floating or not fully weighted on the glass surface.
Verdict: Wan 2.5 (Preview) followed the complex spatial instructions much better, correctly placing the plant behind the glass and balancing all elements according to the prompt. While ImagineArt 1.5 (Preview) produced a more photorealistic wood texture and more convincing glass refraction, it failed to properly place the plant 'behind' the cube as requested.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent skin texture with realistic age spots and wrinkles
- + Strong colors with the red bicycle appearing vibrant against the wet pavement
- + Composition feels intimate and candid
- − Anatomical failure with the hands, including extra fingers and merged digits
- − Mechanical logic of the tools and bicycle parts is slightly nonsensical
Wan 2.5 (Preview)
- + Accurately depicts light rain and reflections on the pavement
- + Excellent shallow depth of field and framing that captures the environment
- + Reasonable anatomical correctness in the hands compared to the competitor
- − The visible rain looks somewhat like static lines rather than natural droplets
- − Bicycle geometry is a bit tangled near the kickstand and gears
Verdict: Wan 2.5 (Preview) produces a superior cinematic image that captures the atmosphere of light rain and street reflections perfectly. While ImagineArt 1.5 (Preview) has impressive skin textures, it suffers from severe anatomical issues with the subject's hands that break the realism required by the prompt.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent facial textures and realistic skin imperfections.
- + Stronger 'battle-worn' appearance with a more mature, weathered character.
- + Natural integration of the torchlight reflecting on the chest plate.
- − The braids are a bit messy and less defined than in Model B.
- − The bokeh background is slightly more distracting due to larger light clusters.
Wan 2.5 (Preview)
- + Ornate engravings on the pauldrons are sharp and intricate.
- + Hair braids and beads are very clearly defined and numerous.
- + Higher contrast and cleaner detail on the leather straps and buckle.
- − The subject looks significantly younger, which slightly clashes with the 'battle-worn' descriptor.
- − The dirt on the face looks a bit like makeup or paint rather than natural grime.
Verdict: Both models captured the prompt details exceptionally well, but ImagineArt 1.5 (Preview) creates a more convincing 'battle-worn' character with realistic skin textures and age. Wan 2.5 (Preview) offers superior detail in the armor engravings and hair styling, though the character appears more like a young squire than a seasoned paladin.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Features more realistic and high-quality food photography
- + Clean layout that resembles a high-end restaurant menu
- − Text is completely illegible and lacks structure
- − Missed specific section head requests like 'pizza'
Wan 2.5 (Preview)
- + Strict adherence to the requested grid structure for food photos
- + Includes specific section headers for appetizers, pizza, and mains
- + Better font clarity and usage of vibrant accents
- − Several spelling errors and gibberish text throughout
- − Photos lack the professional culinary styling found in Model A
Verdict: Wan 2.5 (Preview) followed the prompt instructions more accurately, providing a clear grid layout with the specified categories (Appetizers, Pizza, Mains) and recognizable minimalist typography. While ImagineArt 1.5 (Preview) produced much more appetizing and professional food photography, it failed to incorporate the specific sections requested and the text is entirely nonsensical scribbles.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent variety and realistic rendering of different types of sushi
- + Accurately represents the 45-degree isometric perspective requested
- + High level of detail in the textures of the fish and rice
- − Text is placed in the top-right corner instead of top-center
- − Typography is stylized with outlines rather than being standard 'large bold' text
Wan 2.5 (Preview)
- + Perfect text placement and styling according to the prompt
- + Superior '3D cartoon' aesthetic with soft, refined 3D-render textures
- + Very clean composition and lighting that matches the 'diorama' request
- − Displays only a single piece of sushi rather than a wider variety
- − Perspective is a bit flatter/lower than a true isometric 45-degree angle
Verdict: Wan 2.5 (Preview) better captures the specific aesthetic requested, delivering a clean 3D cartoon diorama with perfect text rendering and placement. While ImagineArt 1.5 (Preview) shows a more impressive variety of sushi with realistic textures, it fails to center the text and feels less like a 'miniature cartoon' and more like a high-end food render.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent typography with a flowing, classic feel
- + Accurate name rendering including the accent on 'Caffè'
- + Good use of warm brown tones for a sophisticated brand identity
- − The 'Est. 1720' text is placed simply inside the emblem rather than a banner
- − The steam effect is very faint and almost unnoticeable
- − Minor clipping on the tail of the 'n' in 'Florian'
Wan 2.5 (Preview)
- + Stronger adherence to 'vector' and 'minimalist' keywords
- + Includes a clear banner for the text as requested
- + Features a more prominent and stylized steam icon
- − Noticeable spelling error in 'Caffè' (missing the second 'f')
- − The texture on the paper background is a bit heavy-handed compared to the 'subtle' request
Verdict: Wan 2.5 (Preview) produced a more balanced and professional layout that strictly followed the banner and vector style requested, but it failed on basic spelling. ImagineArt 1.5 (Preview) handled the typography and spelling of 'Caffè Florian' perfectly but missed the specific 'banner' element, resulting in a slightly less structured logo.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent typography and layout for the main title and crew section.
- + Adheres strictly to the requested NASA-inspired color palette.
- + Includes creative silhouettes for the crew members that fit the vector style.
- − Several spelling errors in technical labels like 'TRANSLACDIC' and 'LANDAR'.
- − The logical flow of the icons is confusing and loops back on itself.
Wan 2.5 (Preview)
- + Features a very clean, high-quality flat vector Saturn V and Lunar Module.
- + Displays a logical flow from launch to landing with clear trajectory lines.
- + Correct spelling for technical terms like 'TRANSLUNAR' and 'TRANQUILITY'.
- − The crew portraits are inconsistently styled, particularly the illustration for Collins.
- − Lacks a main title header at the top of the poster.
Verdict: Both models followed the prompt's aesthetic and content requirements well. ImagineArt 1.5 (Preview) has a superior overall poster layout and better crew iconography, but suffers from significant spelling errors and a disjointed flow. Wan 2.5 (Preview) is the winner because it successfully visualizes a logical mission progression with high-quality icons and accurate spelling, despite the odd crew portrait for Michael Collins.
Explore each model
Alibaba's text-to-image and image-to-image generation model from the Wan AI suite, offering high-quality visual generation capabilities