Head to head
Esc

Models · slot A

to navigate to pick

OmniGen v2 VectorSpaceLab Vidu Q2 ShengShu Technology

Settled by community votes across 15 shared challenges, with an AI judge weighing in on each.

OmniGen v2

16.9 arena score

#55 of 62 in Text-to-Image

Skill signature · Text-to-Image

Vidu Q2

19.8 arena score

#42 of 62 in Text-to-Image

Vote tally

Where the votes landed

OmniGen v2

0.0%

win rate

Ties

0.0%

Vidu Q2

100.0%

win rate

0.0% 0.0% ties 100.0%
Shared challenges 15

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

OmniGen v2
Vidu Q2
0% wins 0% ties 100% wins

AI Judge Analysis

OmniGen v2

  • + Reflections in the sphere are glossy and realistic.
  • + Adheres perfectly to the lighting direction requested.
  • + The glass cube looks thick and physical.
  • The plant is highly blurred and lacks detail.
  • The book is slightly too generic/minimalist.

Vidu Q2

  • + Excellent texture on the red book cover and gold lettering.
  • + Realistic plant detail with sharp fronds.
  • + Intricate shadows and reflections on the wooden surface.
  • Lighting is harsher than the 'soft light' requested.
  • Perspective on the bottom glass plane is slightly skewed.

Verdict: Both models followed every instruction in the prompt perfectly. OmniGen v2 produced a cleaner image with better soft lighting, while Vidu Q2 excelled in rendering fine textures like the leather on the book and the leaves of the plant. Vidu Q2 is slightly preferred for its superior detail and realistic glass-on-wood shadows.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Excellent reflections on the wet pavement
  • + Accurate representation of light rain and atmospheric depth
  • + Good full-body composition and color balance
  • The man is standing next to the bike rather than actively repairing it
  • The bicycle design has minor structural inconsistencies near the chain guard
  • Lacks the requested motion blur from passing cars

Vidu Q2

  • + Captured the 'imperfect framing' and 'repairing' action much more accurately
  • + Highly detailed natural skin texture on the hands and face
  • + Effective shallow depth of field with realistic background bokeh
  • Anatomical errors with extra hands visible near the handlebars
  • The bicycle chain and frame geometry are physically impossible and messy
  • Over-sharpening artifacts are visible on the skin and metal textures

Verdict: OmniGen v2 produces a cleaner, more pleasant image that captures the atmosphere of rain and wet pavement well, but it fails the specific action of 'repairing'. Vidu Q2 follows the 'candid' and 'imperfect framing' prompt much better and shows a man actually working on the bike, but it suffers from severe anatomical glitches and structural nonsense in the bicycle's mechanical parts.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Excellent photographic clarity and face rendering
  • + Warm, atmospheric lighting that realistically interacts with the subject
  • + Intricate engraving on the pauldrons
  • The character looks too clean and 'pristine' for the 'battle-worn' prompt requirement
  • Dirt on face looks like distinct spots rather than realistic grime

Vidu Q2

  • + Strong adherence to the 'battle-worn' prompt with visible scars and weathered armor
  • + Superior detail on leather straps, buckles, and cloth underlayers
  • + More complex and realistic hair braiding with beads
  • The lighting is a bit flat compared to the dramatic glow in the other version
  • Slight anatomical awkwardness where the neck meets the armor

Verdict: While OmniGen v2 produces a more beautiful and cleaner portrait, Vidu Q2 is the clear winner for prompt adherence. Vidu Q2 accurately captures the 'battle-worn' aesthetic with realistic scars, weathered textures, and highly detailed leather and cloth components, whereas OmniGen v2 feels like a studio photoshoot of a pristine character.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Layout follows a clean, professional grid system.
  • + Bold sans-serif fonts are very clear even if the text is misspelled.
  • + Composition effectively uses white space for a high-end feel.
  • Includes significant spelling errors like 'RESTAURATED MENTS' and 'PIZZZZAN'.
  • The food photography looks somewhat artificial and lacks detail.

Vidu Q2

  • + Provides a more realistic layout with prices, mimicking an actual menu structure.
  • + Food photos are vibrant and appear more appetizing and varied.
  • + Excellent use of colorful accents and icon elements to define sections.
  • The text rendering is highly garbled with non-standard characters.
  • The grid feels a bit cluttered compared to the minimalist request.

Verdict: OmniGen v2 provides a superior minimalist layout with a very clean grid, though it suffers from obvious spelling mistakes. Vidu Q2 offers more realistic food photography and better menu functional details like pricing columns, but the text is largely unreadable. OmniGen v2 is the likely winner for better adhering to the 'minimalist' and 'bold sans-serif' aesthetic requested.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Excellent text legibility and graphic design integration
  • + Vibrant colors and clean, professional appearance
  • Failed the main request for an exploded/deconstructed burger view
  • Graphic design style feels more illustrated than photorealistic

Vidu Q2

  • + Successfully followed the instruction for an exploded, deconstructed burger
  • + Features a more realistic texture on the food and background
  • + Text effects match the 'fiery' prompt perfectly
  • The currency symbol is incorrect (rendered as an 'E' variant instead of €)
  • The composition feels slightly more cluttered compared to model A

Verdict: Vidu Q2 is the winner as it accurately fulfilled the core 'exploded burger' requirement which OmniGen v2 ignored. While OmniGen v2 produced a cleaner graphic, Vidu Q2 captured the sense of motion, the photorealistic textures, and the fiery text effects specified in the prompt.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Successfully rendered the date accurately into the title.
  • + The text is very legible in the lower half of the board.
  • + Captures a clean chalkboard aesthetic with a wooden frame.
  • Large spelling error in the word 'SPECIALS' (rendered as SPECALS).
  • Severe overlapping and garbled text in the first two menu items.
  • The handwriting style for the title is blocky rather than the requested elegant cursive.

Vidu Q2

  • + Excellent authentic chalk texture with realistic smudges and pressure variations.
  • + The title font adheres well to the request for 'elegant cursive' style.
  • + Better spatial layout that feels like a real café menu board.
  • Significant spelling errors throughout the menu items (e.g., 'Musshoom', 'Octopd', 'Lemepun').
  • Incorrect pricing for the first item ($34 instead of $24).
  • Text becomes very messy and illegible toward the bottom of the board.

Verdict: Both models struggled with the complex task of rendering specific long-form strings of text without spelling errors. OmniGen v2 has better legibility in the bottom half but suffers from serious line overlapping and a major typo in the header, while Vidu Q2 much more accurately captures the artistic 'chalk' aesthetic and cursive requested, despite its phonetic spelling failures.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Clean, vector-like aesthetic with clear subjects
  • + Accurate rendering of basic astronaut gear and horse anatomy
  • Completely failed the negative constraint to have the horse on top
  • The background is very basic and lacks the cinematic quality requested
  • Composition is static and lacks the surreal atmosphere requested

Vidu Q2

  • + Dynamic, cinematic composition with high level of detail
  • + Rich, vibrant cosmic background that fits the space theme well
  • + Incredible texture on the horse that incorporates a galaxy-skin effect
  • Completely failed the logic-defying spatial instruction of having the 'horse on top'
  • Minor artifact where the reins blend into the astronaut's hand

Verdict: Both models failed the specific spatial reasoning challenge of placing the horse on top of the astronaut, instead defaulting to the standard 'astronaut on horse' trope. However, Vidu Q2 is significantly better in terms of visual quality, cinematic detail, and creative interpretation of 'surreal' through the galaxy-patterned horse hide, whereas OmniGen v2 produced a flat, generic image.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Strong image clarity and sharp textures on the capybara's fur.
  • + Good color vibrance in the yellow taxi and blue jacket.
  • The capybara has realistic human hands and arms, which is anatomically incorrect and unsettling.
  • The woman is sitting in the passenger seat next to the driver instead of the back seat.
  • The capybara's head is merged awkwardly with the jacket collar.

Vidu Q2

  • + Accurately placed the woman in the back seat as requested.
  • + The capybara uses its actual paws to grip the steering wheel.
  • + The composition and perspective feel like a more authentic taxi interior.
  • The woman's hands and phone are slightly blurry/distorted.
  • The capybara's paws on the wheel have some minor clipping issues.

Verdict: Vidu Q2 adhered much better to the spatial requirements of the prompt by placing the businesswoman in the back seat and correctly rendering capybara paws instead of human hands. OmniGen v2 failed the primary composition by placing the passenger in the front and produced a jarring effect by giving the capybara human limbs.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Near perfect rendering of the main title text.
  • + Clean, polished graphic design style with a clear border.
  • + Excellent central jack-o-lantern design that pops against the background.
  • Confusing layout for the bottom event details with text overlapping.
  • The scroll banner contains gibberish text instead of the requested phrase.
  • Small location name spelling error ('Arcas' instead of 'Arches').

Vidu Q2

  • + Superior parchment texture and 'vintage gothic' aesthetic.
  • + Includes impressive vine and spiderweb border details as requested.
  • + Realistic rendering of the jack-o-lantern with atmospheric lighting.
  • Several spelling errors in the title and banner text.
  • Incorrect date provided (2025 instead of 2026).
  • Composition feels slightly cluttered with the overlapping thorns.

Verdict: Both models struggled with the complex multi-line text requirements, but OmniGen v2 produced a cleaner, more legible main title. However, Vidu Q2 captured the 'vintage gothic parchment' aesthetic much more effectively with its detailed borders and realistic textures, despite the spelling errors. Choosing a winner depends on whether one prioritizes the clean layout of OmniGen v2 or the superior atmospheric detail of Vidu Q2.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Excellent 3D typography and drop shadows
  • + Perfectly executes the isometric diorama base concept
  • + High clarity and clean design aesthetic
  • The flag icon is a generic bicolor rectangle instead of the Japanese flag
  • Sushi anatomy is slightly strange, merging nigiri with a shrimp tail

Vidu Q2

  • + Accurate Japanese flag icon
  • + Higher variety of sushi types on the plate
  • + Soft, clay-like textures fit the miniature cartoon theme well
  • Text is somewhat flat and lacks the 3D depth of Model A
  • The 'diorama base' is less distinct and looks more like a tray
  • Small artifacts present on the corners of the base

Verdict: OmniGen v2 produces a much cleaner, more professional-looking 3D graphic with superior typography and a better realization of the isometric diorama request. Vidu Q2 followed the 'Japanese flag' instruction more accurately and included more diverse sushi, but it lacked the polished render quality and bold compositional impact of OmniGen v2.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Bright and colorful visual appeal
  • + High clarity on the central subjects
  • Failed to include the bunny
  • Characters look like 3D animation/cartoons rather than the requested hyper-photorealistic style
  • Static posing fails to capture the 'playfully chasing' and 'tumbling' action from the prompt

Vidu Q2

  • + Captured all requested animals including the puppy, kitten, bunny, and fox kit
  • + Strong adherence to the 'playfully chasing' and 'tumbling' action
  • + Achieved a much more photorealistic texture and lighting style
  • Anatomical issues with the bunny's extra legs
  • Some butterflies have distorted wing structures

Verdict: Vidu Q2 is the clear winner as it successfully incorporated all four animals and captured the energetic action of the prompt, whereas OmniGen v2 missed the bunny and provided a static, cartoonish image. While Vidu Q2 has some anatomical artifacts, its adherence to the photorealistic style and complex scene description is far superior.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Excellent character translation into a clean, modern anime aesthetic.
  • + Simplified, bold compositional elements that enhance the iconic meme readability.
  • + Clean line work and consistent cell shading.
  • Fails to capture the 'hand-painted textures' and 'soft pastel' requirements of the Studio Ghibli prompt.
  • Completely changes the facial expressions, losing the original 'distracted boyfriend' narrative tension.

Vidu Q2

  • + Strong adherence to the Studio Ghibli requested style with watercolor textures and soft pastel color grading.
  • + Preserves the original facial expressions and emotional context perfectly.
  • + Maintains the specific details of the source image like the plaid shirt patterns and background architecture.
  • The line work on the man's face is slightly jittery compared to the smoothness of the women.
  • Slightly less 'clean' than image A, though this is part of the painterly aesthetic requested.

Verdict: Vidu Q2 is the clear winner as it successfully applied the specific 'Studio Ghibli' style elements like hand-painted textures and soft pastels while perfectly preserving the source image's character expressions and composition. OmniGen v2 created a generic modern anime illustration that ignored the texture requirements and changed the emotional context of the scene.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
OmniGen v2
Before After
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Successfully applied blowing hair effect
  • + Added flying leaves as requested
  • Significantly altered the woman's facial features and the dog's appearance
  • The overall image style became more painterly and less photographic than the source
  • Poor source preservation with background elements shifting

Vidu Q2

  • + Excellent source preservation, keeping the woman, dog, and background almost identical to the original
  • + Highly effective 'hair in the wind' effect that blends naturally
  • + Abundant flying leaves create a strong sense of dynamic motion
  • Some leaves in the foreground are a bit blurry, though this adds to the motion feel

Verdict: Vidu Q2 is the clear winner as it successfully applied all requested edits while maintaining near-perfect consistency with the source image. OmniGen v2 essentially regenerated the entire scene, resulting in a different looking person and dog, failing the primary goal of image editing.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Excellent vector-style cleanliness with bold lines
  • + Effective use of the requested banner element
  • + Good adherence to the warm brown and cream color palette
  • Spelling error in the main name 'CAFFFLORIN'
  • The steam element is very basic and small

Vidu Q2

  • + Beautiful subtle texture and shading on the cloche
  • + Elegant illustrative style for the steam
  • + Accurate representation of the 'vintage minimalist' aesthetic
  • Significant text rendering failures throughout the image
  • Redundant and nonsensical text lines added at the bottom
  • Missed the accent on 'Caffè'

Verdict: OmniGen v2 produces a much cleaner and more professional vector-like logo, though it suffers from a minor spelling error. Vidu Q2 captures a nicer vintage atmosphere and texture, but the text is completely garbled and redundant, making it unusable as a logo. OmniGen v2 is the winner for its superior composition and legibility.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

OmniGen v2
Vidu Q2

AI Judge Analysis

OmniGen v2

  • + Strong adherence to the requested color palette (navy, white, red).
  • + Clean, modern layouts that feel like a professional infographic template.
  • Incorrect mission number (Apolo 17 instead of 11).
  • Nonsensical text and icons that do not clearly follow the requested 6-step sequence.
  • Visual symbols are generic and repetitive rather than specific mission stages.

Vidu Q2

  • + Successfully included specific mission elements like the Saturn V and Lunar Module.
  • + Good illustration of the Earth and Moon with orbit rings as requested.
  • + Organized layout that flows through the stages of the mission.
  • Text is largely gibberish and misaligned with the intended step names.
  • Aesthetics are slightly more 'cartoonish' than the requested 'clean flat-vector' infographic style.
  • Background elements like the four-pointed stars feel a bit generic.

Verdict: Vidu Q2 is the winner because it actually attempts to illustrate the specific steps requested, including recognizable icons for the Saturn V, Lunar Module, and trajectory orbits. In contrast, OmniGen v2 provides a generic template with high-quality lines but fails to represent the Apollo 11 mission accurately, even getting the mission number wrong.

Next steps

Explore each model