Vyro AI's professional-grade text-to-image model delivering photorealistic output with accurate text rendering and typography precision for commercial workflows
Settled by community votes across 7 shared challenges, with an AI judge weighing in on each.
ImagineArt 1.5 (Preview)
#6 of 62 in Text-to-Image
Qwen Image 2.0
#34 of 62 in Text-to-Image
Where the votes landed
ImagineArt 1.5 (Preview)
0%
win rate
Ties
0%
Qwen Image 2.0
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent depiction of solid glass physics and refraction.
- + High realism in textures, especially for the wooden table and weathered book.
- + Accurate lighting and shadow work that matches the soft window light prompt.
- − The sphere appears to be fused with or resting on a solid core within the glass rather than floating in a hollow space.
- − The plant is very dense and takes up most of the background, making it less 'partially visible through glass' and more just 'behind'.
Qwen Image 2.0
- + Successfully depicts the sphere inside a hollow glass chamber.
- + Clear implementation of all spatial requirements including behind-the-cube visibility.
- + Clean, modern aesthetic with good color balance.
- − The glass cube has illogical reflections, showing spheres on the outer faces that don't match the interior.
- − The perspective of the glass base is slightly skewed compared to the table surface.
- − The lighting is a bit flat compared to the more dramatic shadows in Image A.
Verdict: Both models followed the prompt instructions perfectly in terms of object placement. ImagineArt 1.5 (Preview) produced a more photorealistic image with superior textures and lighting, while Qwen Image 2.0 did a better job of representing the sphere as an object contained 'inside' a hollow glass structure. ImagineArt 1.5 is the preferred winner due to significantly higher visual quality and more convincing material physics.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent high-resolution skin textures and realistic facial wrinkles
- + Very strong adherence to the 'imperfect framing' prompt with a tight, candid perspective
- + Effective rendering of rain droplets on the bicycle frame
- − Anatomical errors in the hands with extra/merging fingers
- − Lacks the requested motion blur in the background car
- − Bicycle mechanics (the pedal and crank) are physically non-sensical and distorted
Qwen Image 2.0
- + Successfully captured motion blur on the passing vehicle
- + Excellent use of shallow depth of field and street reflections
- + Better overall composition and believable elderly subject
- − Hands are somewhat blurry and lack detail
- − The bicycle chain and sprocket have some structural inconsistencies
Verdict: Qwen Image 2.0 followed the technical photographic instructions much better, specifically capturing the motion blur and atmospheric reflections that ImagineArt 1.5 (Preview) missed. While ImagineArt 1.5 has impressive skin texture, its failure to render coherent hands or bicycle parts makes it less successful as a realistic image.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI judge analysis unavailable for this challenge.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Successfully creates a complex multi-column menu layout
- + Includes a high-quality, large hero image of a pizza
- + Accents and thin rules give it a polished, upscale feel
- − The text is completely illegible and garbled
- − The grid layout for the food photos is irregular rather than a clean grid
Qwen Image 2.0
- + Features a very clean, structured minimalist grid
- + Renders the category titles (Appetizers, Pizza, Mains) with perfect legibility
- + Food photography is well-lit and appetizing
- − The small item descriptions contain significant character distortions
- − The design is a bit repetitive with multiple pizzas shown under various headings
Verdict: Qwen Image 2.0 is the clear winner as it successfully follows the prompt's layout and font requirements, rendering large, legible headlines in a bold sans-serif font. While ImagineArt 1.5 (Preview) creates an interesting menu composition, its total failure at typography and less organized grid makes it less practical for a design task.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent 3D miniature diorama feel with high-quality PBR textures
- + Includes complex sushi varieties with detailed translucency and sheen
- + Good adherence to the 45° isometric perspective requested
- − Text layout is pushed to the top-right corner rather than the requested top-center
- − Stylized font choice for 'SUSHI' is slightly less clean than requested bold text
Qwen Image 2.0
- + Perfect text placement at top-center with clean bold fonts
- + Accurate representation of the small raised diorama base using a wooden texture
- + High clarity and very clean overall composition
- − The sushi looks more like a standard photo than a '3D cartoon scene'
- − The flag icon is quite large and placed to the side rather than with the text vertically
Verdict: ImagineArt 1.5 (Preview) captures the 'miniature 3D cartoon' and 'isometric' aesthetic much better, resulting in a more cohesive diorama. Qwen Image 2.0 followed the text placement instructions more accurately but produced a flatter, more photographic image that lacked the specific 3D-render style requested in the prompt.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent typography integrated into a cohesive badge design
- + Consistent vector emblem style with sophisticated brown and cream tones
- + Correct spelling and placement of the text
- − The 'Est. 1720' text is placed on a solid shape rather than a specific banner
- − The steam effect is very subtle and nearly lost against the background
Qwen Image 2.0
- + Successfully includes the requested banner for the 'Est. 1720' text
- + Distinct steam visualization on the cloche
- + Clean, high-contrast illustration
- − The typography for 'Caffè Florian' is a bit basic and less 'classic' in style
- − Technical glitches where the steam interacts with the cloche rim and the banner is cut off on the right
- − The cloche handle looks slightly off-center
Verdict: ImagineArt 1.5 (Preview) produced a more professional and aesthetically pleasing logo with superior typography and a cohesive color palette. While Qwen Image 2.0 followed the 'banner' part of the prompt more literally, it suffered from technical rendering artifacts and less sophisticated font choices.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent layout with a strong heading and distinct sections.
- + Highly detailed graphics including specific crew silhouettes and names.
- + Professional use of the color palette and background elements like stars.
- − Several spelling errors in technical terms (e.g., 'TRANSLACDXIC', 'LANDAR', 'MAFKER').
- − Flow of the infographic path is slightly confusing and loops back on itself.
Qwen Image 2.0
- + Clear, linear progression of the mission steps which is easy to follow.
- + Good representation of the Lunar Module for 'Descent' and 'Landing'.
- + Correct spelling for almost all labels including 'Tranquility'.
- − Typos in main labels such as 'Translunjar' and 'Lunar Orbt' (where one is partially obscured).
- − Iconography is somewhat inconsistent in scale and style compared to Model A.
Verdict: ImagineArt 1.5 (Preview) creates a more visually sophisticated and balanced poster with high-quality vector aesthetics, but it suffers from significant spelling errors and a confusing logical flow. Qwen Image 2.0 provides a much clearer linear infographic that is easier to read, though it lacks the professional graphic design polish and consistent icon style shown by ImagineArt. ImagineArt 1.5 is the preferred model for its superior layout and adherence to the 'modern infographic' aesthetic.
Explore each model
Alibaba's Qwen Image 2.0 model with enhanced text rendering, supporting both Chinese and English prompts with up to 6 images per request