Reve AI's text-to-image generation model with strong aesthetic quality, accurate text rendering, and detailed instruction following capabilities
Settled by community votes across 8 shared challenges, with an AI judge weighing in on each.
Reve Image 1.0
#40 of 62 in Text-to-Image
Vidu Q2
#42 of 62 in Text-to-Image
Where the votes landed
Reve Image 1.0
0%
win rate
Ties
0%
Vidu Q2
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
Reve Image 1.0
- + Excellent adherence to lighting instructions (soft light from the left)
- + Clean, modern aesthetic with realistic glass refraction
- + The blue sphere appears gracefully suspended
- − The plant is more to the left rather than directly behind the cube as requested
- − The red book looks more like a small notepad
Vidu Q2
- + Perfect positioning of the plant behind the cube
- + Highly detailed textures on the wooden table and book spine
- + Excellent shadow and reflection work inside the cube
- − Lighting is harsher and more golden than the 'soft window light' requested
- − The sphere appears to be floating unnaturally right near the bottom with a double reflection
Verdict: Both models followed the complex spatial instructions well. Vidu Q2 followed the positioning of the plant better and provided much richer textures, but Reve Image 1.0 captured the specific 'soft window light' atmosphere more accurately and had a cleaner composition. Vidu Q2 is the winner due to superior detail and better adherence to the 'behind the cube' spatial requirement.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
Reve Image 1.0
- + Excellent depiction of cinematic torchlight and bokeh sparks in the background.
- + Highly intricate engraving details on the plate armor that look historically inspired.
- + Authentic skin texture and lifelike eye reflection.
- − The character looks very young, which may slightly clash with the 'battle-worn' description.
- − The dirt on the face appears a bit like digital specks rather than smudged grime.
Vidu Q2
- + Perfect interpretation of a 'battle-worn' character with visible scars and grit.
- + Superior rendering of leather straps and buckles over the armor.
- + Strong composition that captures the paladin's stern expression and braids with beads well.
- − The engraving on the armor is slightly less detailed compared to Reve Image 1.0.
- − The background lighting feels a bit more generic and less atmospheric than the torchlight in model A.
Verdict: Reve Image 1.0 excels in cinematic lighting and the sheer detail of the armor's engravings, giving it a very high-budget film look. However, Vidu Q2 better captures the essence of a 'battle-worn paladin' through its character design, including more convincing scars, leather textures, and a mature facial structure. Vidu Q2 is the preferred output for more accurately balancing all elements of the prompt.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
Reve Image 1.0
- + Excellent source preservation, keeping the face and lighting nearly identical to the original.
- + Follows the prompt for a full, thick head of hair.
- − The hair texture looks somewhat like a wig and lacks a natural, feathered transition at the forehead.
- − The volume of the hair is slightly disproportionate to the head shape.
Vidu Q2
- + Features a very realistic and modern hair texture with stray strands for added authenticity.
- + The transition at the hairline is more believable and integrated.
- − Slightly alters the facial features, making the eyes and brow look a bit different from the source image.
Verdict: Both models successfully added a full head of hair while maintaining the overall scene. Vidu Q2 produces a more realistic and stylish hair texture, though it slightly alters the subject's facial appearance, whereas Reve Image 1.0 preserves the original face perfectly but adds hair that looks somewhat artificial.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
Reve Image 1.0
- + Excellent adherence to the 'soft refined textures' and 'miniature 3D cartoon' prompt.
- + Very clean, minimal, and professional presentation.
- + Perfect isometric 45-degree angle.
- − The sushi looks a bit overly simplified, like plastic toys rather than food.
- − Text is slightly off-center relative to the composition.
Vidu Q2
- + Superior detail in the sushi models, accurately depicting rice grains and varied toppings.
- + Creative placement of the flag icon on a stand.
- + Better technical execution of the 'realistic PBR materials' requested.
- − The white rice grains look slightly lumpy or chaotic compared to a professional 3D render.
- − The text 'SUSHI' is not perfectly centered under 'JAPAN'.
Verdict: Both models followed the complex prompt very well, but Reve Image 1.0 captured the 'cartoon' and 'soft' aesthetic more effectively. However, Vidu Q2 is the winner because its sushi models are much more sophisticated and detailed, successfully balancing the 'realistic PBR' and 'miniature' aspects of the prompt better than the overly-simplified blocks in Reve Image 1.0.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
Reve Image 1.0
- + Excellent facial preservation from the source image.
- + Creative humor with the dog wearing a headset and tie.
- + Clearly incorporates all elements: hockey (stick, jersey, puck), TV anchor (news desk, mic), and dogs.
- − The hand holding the microphone is poorly rendered with anatomical errors.
- − Caricature style is a bit safe, focusing more on a composite than an exaggerated anatomical caricature.
Vidu Q2
- + Strong 'caricature' aesthetic with more exaggerated features and a professional vector illustration style.
- + Successfully incorporates secondary elements like the hockey rink background and multiple dogs.
- + Great preservation of the subject's clothing style (denim shirt over black top).
- − The microphone and the hand holding it are poorly proportioned and lack structural coherence.
- − Distorted finger rendering on the subject's left hand.
Verdict: Reve Image 1.0 does a superior job of capturing the subject's likeness while effectively merging the three requested themes (hockey, news, and dogs) into a cohesive scene. Vidu Q2 offers a more traditional 'caricature' drawing style and better preserves the original clothing, but it suffers from more significant anatomical distortion in the hands and a less focused composition.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
Reve Image 1.0
- + Perfectly captures the Studio Ghibli art style with thick line work and painterly backgrounds.
- + Preserves the exact composition and posing of the original meme flawlessly.
- + Achieves a beautiful hand-painted watercolor texture.
- − The face of the woman in red is slightly too blurry compared to the rest of the scene.
Vidu Q2
- + Excellent preservation of the original subjects' facial features while stylizing them.
- + High resolution and clean linework.
- − The style leans more toward modern digital anime/manhwa than the specific 'Studio Ghibli' look requested.
- − The lighting feels a bit too bright and lacks the 'dreamy, nostalgic' mood asked for.
Verdict: Reve Image 1.0 is the clear winner as it masterfully adapts the source image into a specific Studio Ghibli aesthetic, including the characteristic background painting style and character design. Vidu Q2 creates a high-quality illustration, but it feels more like a generic filter or modern anime style and fails to capture the requested 'warm, nostalgic' mood as effectively as Reve Image 1.0.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
Reve Image 1.0
- + Excellent addition of dynamic hair movement that feels natural to the source pose
- + Motion-blurred leaves create a strong sense of speed and wind
- + Near-perfect preservation of the original subjects and background
- − Some leaves appear overly large and slightly overlap the subject in an intrusive way
Vidu Q2
- + Successfully added a large number of autumn leaves to the scene
- + Changes the hair effectively to show a blowing effect
- + Maintains the identity and features of the woman and dog well
- − The added leaves lack motion blur, appearing like static stickers floating in the air
- − The hair edit is slightly less fluid and dynamic compared to the competitor
Verdict: Reve Image 1.0 is the clear winner because it accurately applies motion blur to the blowing leaves, whereas Vidu Q2 adds sharp, static leaves that fail to convey 'dynamic motion'. Reve Image 1.0 also creates a more energetic and convincing wind effect in the woman's hair while perfectly preserving the source image's integrity.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
Reve Image 1.0
- + Closer adherence to the requested NASA-inspired navy palette.
- + More accurate spelling of technical terms like 'Earth Orbit' and 'Landing'.
- + Clearer progression of steps representing the mission stages.
- − Missed step 3 (Translunar) entirely in the sequence.
- − Significant spelling errors in the main title and 'Lunnr' / 'Dessent'.
- − The icons for descent and landing are repetitive and lack distinct visual storytelling.
Vidu Q2
- + Includes a wider variety of custom icons including astronauts and trajectories.
- + Captures the 'light gray' aspect of the palette well for the background.
- + Better consistency in the illustrative style across the different icons.
- − Complete failure on text legibility and spelling for almost every word.
- − Layout is cluttered and lacks the 'clean, modern' infographic feel requested.
- − Icons for the lunar module are overly complex and lose the 'flat vector' look.
Verdict: Reve Image 1.0 is the preferred model because it successfully creates a usable infographic layout with mostly legible text, whereas Vidu Q2 fails significantly on typography and information hierarchy. Reve Image 1.0 followed the specific color palette instructions more effectively, creating a cohesive NASA-inspired design despite missing one of the requested steps.
Explore each model
ShengShu Technology's text-to-image and reference-to-image model with support for character consistency and multi-reference image processing