Vyro AI's professional-grade text-to-image model delivering photorealistic output with accurate text rendering and typography precision for commercial workflows
Settled by community votes across 7 shared challenges, with an AI judge weighing in on each.
ImagineArt 1.5 (Preview)
#5 of 62 in Text-to-Image
Qwen Image 2512
#30 of 62 in Text-to-Image
Where the votes landed
ImagineArt 1.5 (Preview)
62.5%
win rate
Ties
25.0%
Qwen Image 2512
12.5%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent photo-realistic visual quality
- + Natural integration of the plant and window light
- + High level of detail on the book texture and wooden table
- − Failed spatial prompt adherence by placing the cube on the book instead of the book on the cube
Qwen Image 2512
- + Perfect adherence to the spatial requirements of the prompt
- + Accurate representation of window light from the left
- + Clean composition that follows every instruction
- − Glass reflections are slightly confusing with an extra blue sphere appearing on the right
- − The plant behind is very blurry compared to Model A
Verdict: Qwen Image 2512 followed all spatial instructions perfectly, placing the red book on top of the glass cube and the blue sphere inside. In contrast, ImagineArt 1.5 failed the primary spatial challenge by placing the glass cube on top of the book, despite having significantly higher photographic realism and detail.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent depiction of raining texture with visible droplets on the bike.
- + Very realistic skin texture and facial lighting on the man.
- + Captured the 'imperfect framing' prompt with a tight, candid perspective.
- − The man's hand interacting with the bike is physically incoherent.
- − Missing the 'motion blur' for the car in the background.
Qwen Image 2512
- + Strong composition that feels more cinematic and balanced.
- + Good application of shallow depth of field and bokeh.
- + Better integration of the man and the bicycle as a whole.
- − The man is posing for the camera rather than 'repairing' the bike.
- − Skin and hair textures look slightly smoothed and less 'natural' than Model A.
Verdict: ImagineArt 1.5 (Preview) followed the prompt's request for an 'imperfect framing' much better, creating a truly candid feel with highly realistic skin textures and rain details, though it suffered from anatomical issues in the hands. Qwen Image 2512 produced a more aesthetically pleasing, professional-looking photograph with a better 50mm lens look, but it missed the 'candid' and 'repairing' aspect of the prompt, leaning into a portrait style instead. ImagineArt 1.5 is the winner for capturing the specific gritty realism and action requested.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent execution of warm torchlight reflecting off the skin and armor
- + Very clear, detailed engraving on the chest plate
- + Physically plausible leather strap and ring hardware
- − The torch flame is very large and distracting, partially cutting out of frame
- − The skin texture looks slightly more plastic/smooth compared to the other model
Qwen Image 2512
- + Superior battle-worn details with realistic scars and dirt
- + Better adherence to the 'hair braided with small beads' prompt with multiple colored beads
- + Beautiful bokeh sparks that create a nice sense of atmosphere
- − The buckles on the leather straps are slightly distorted and cluttered
- − The hair and beard look slightly over-sharpened compared to the facial skin
Verdict: Both models followed the prompt exceptionally well, but Qwen Image 2512 produces a better 'battle-worn' aesthetic with more distinct scars and a better integration of hair beads and bokeh sparks. While ImagineArt 1.5 (Preview) has very nice lighting on the armor, the overall composition and character detail in Qwen Image 2512 feel more cohesive and professional.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Follows the requested brochure-style layout with a clean, professional aesthetic.
- + Accurately includes and spells the requested section headers like 'Appetizers', 'Pizza', and 'Mains'.
- + Uses a realistic white background with high-quality, vibrant food photography that integrates well with the text.
- − Small body text is illegible placeholder gibberish.
- − The perspective of the trifold mockup makes it harder to see the full design at once.
Qwen Image 2512
- + Strong adherence to the 'grid' requirement for food photos.
- + Effective use of vibrant color-coded accents for different menu categories.
- + Clean, front-facing 2D presentation makes the layout easy to read.
- − Significant spelling errors in every primary header (e.g., 'Appetiizers', 'Piesmanets', 'Means').
- − Overall layout feels slightly more cluttered than the 'minimalist' request.
Verdict: ImagineArt 1.5 (Preview) produced a much more professional and usable design, with correct spelling for the main headers and a high-end casual dining aesthetic. While Qwen Image 2512 followed the grid requirement more strictly, its numerous spelling errors and slightly chaotic layout make it less effective as a design mockup.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent 3D text rendering with great depth and perspective alignment.
- + Very high-quality PBR materials, especially on the fish textures and rice grains.
- + Strong adherence to the isometric perspective with clean, sharp edges.
- − Minor spelling error in the text ('SUSHN' instead of 'SUSHI').
- − The text is slightly off-center to the right rather than 'top-center'.
Qwen Image 2512
- + Perfect text spelling and placement as requested in the prompt.
- + Great diorama base execution with miniature greenery and ginger.
- + Soft, appealing 'cartoon' aesthetic with a polished 3D finish.
- − The text is flat 2D rather than the 3D style suggested by the 'miniature 3D cartoon scene' context.
- − Texture detail on the fish is slightly less refined than Model A.
Verdict: Both models followed the prompt well, but Qwen Image 2512 wins because it delivered perfect text spelling and placement, whereas ImagineArt 1.5 misspelled 'SUSHI' as 'SUSHN'. While ImagineArt 1.5 had more impressive 3D text and realistic materials, Qwen Image 2512's overall composition and diorama base better captured the 'miniature scene' feel requested.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Perfect adherence to the vector emblem style
- + Clean and professional typography layout within a circular frame
- + Accurate rendering of the 'Est. 1720' text
- − The steam is very abstract and lacks the 'retro' charm of the rest of the logo
- − Minimalist style might be slightly too simple for a vintage luxury brand
Qwen Image 2512
- + Excellent vintage illustration style with high-quality shading
- + Beautifully rendered steam and banner element
- + Strong use of texture on the background requested in the prompt
- − The typography is slightly cramped and less readable than Model A
- − The cloche handle is slightly off-center
Verdict: ImagineArt 1.5 (Preview) creates a very clean, professional vector emblem that perfectly balances the requested text and imagery in a functional logo format. Qwen Image 2512 offers a more artistic and textured illustration with superior vintage shading, but ImagineArt 1.5 is preferred for its superior layout and clearer typography which fits the 'minimalist logo' requirement better.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
ImagineArt 1.5 (Preview)
- + Excellent layout that feels like a professional infographic poster.
- + Accurately represents the crew members with clear labels and iconography.
- + Adheres strictly to the requested NASA-inspired color palette.
- − Several spelling errors in technical labels such as 'TRANSLACDIC' and 'LANDAR'.
- − The flow of the infographics logic is a bit tangled with arrows overlapping center icons.
Qwen Image 2512
- + Features more detailed and recognizable illustrations of the Saturn V and Lunar Module.
- + Clean, legible typography for the header and several sub-labels.
- + Good use of vertical space to transition from Earth to the Moon's surface.
- − Logical numbering is broken with repeating step numbers (two step 2s, two step 3s).
- − Severe spelling errors in labels such as 'Desceeint' and 'Translaurtcoit'.
- − The 'Steps stop at landing' text from the prompt was mistakenly included as a literal label.
Verdict: ImagineArt 1.5 (Preview) produced a superior infographic layout that actually looks like a finished poster, including a cohesive crew section and a better color balance. While Qwen Image 2512 has more detailed illustrations, it failed significantly on logical numbering and included prompt instructions as literal text on the poster.
Explore each model
Improved version of Alibaba's Qwen image model with better text rendering, finer natural textures, and more realistic human generation.