Head to head
Esc

Models · slot A

to navigate to pick

LongCat-Image Meituan OmniGen v2 VectorSpaceLab

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

LongCat-Image

12.9 arena score

#61 of 62 in Text-to-Image

Skill signature · Text-to-Image

OmniGen v2

16.8 arena score

#57 of 62 in Text-to-Image

Vote tally

Where the votes landed

LongCat-Image

0%

win rate

Ties

0%

OmniGen v2

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent adherence to the 'plant behind the cube visible through the glass' instruction
  • + Realistic materials including convincing glass thickness and reflections
  • + Accurate lighting coming from the window on the left
  • The glass cube appears to have a mirrored base not explicitly requested
  • The blue sphere is slightly off-center compared to Model B

OmniGen v2

  • + Clean, minimalist aesthetic with vibrant colors
  • + The blue sphere is perfectly centered and has a high-quality finish
  • Fails significantly on the spatial requirement of the plant being 'behind the cube', placing it to the side instead
  • The glass cube has an illogical floating appearance for the red book due to lack of visible top thickness
  • The red book looks more like a plastic box than a book

Verdict: LongCat-Image followed every instruction in the prompt, particularly the complex visual task of showing the plant through the glass. In contrast, OmniGen v2 failed the spatial positioning of the plant and created a less convincing material for the book, although it produced a very clean image.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent adherence to the 'imperfect framing' and 'motion blur' instructions.
  • + Strong cinematic atmosphere with realistic lighting and reflections.
  • + The elderly man's pose and skin texture look very natural and candid.
  • The bicycle rendering has minor structural issues, with the fork appearing disconnected or doubled.
  • The man's hands lack fine detail where they interact with the metal.

OmniGen v2

  • + High resolution and clean rendering of the central subject.
  • + Beautiful and clear reflections on the wet pavement.
  • Fails to include transition motion blur or imperfect framing; feels too static and centered.
  • The bicycle is missing its chain or any visible drivetrain connection.
  • The man is standing next to the bike rather than actively repairing it as requested.

Verdict: LongCat-Image adheres much better to the prompt, capturing the specific 'candid' and 'motion blur' requirements that give a street photography feel. While OmniGen v2 has cleaner reflections, it fails significantly on the technical details of the bicycle (missing chain) and does not show the man 'repairing' the bike, instead showing him just standing with it.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent adherence to the 'battle-worn' descriptor with visible wounds and dirt.
  • + Highly detailed engraved plate armor and textured leather straps.
  • + Captures the braided hair with small beads perfectly as requested.
  • The bokeh sparks appear a bit generic and flat in the upper right.

OmniGen v2

  • + Beautiful cinematic lighting and shallow depth of field.
  • + Lifelike eyes with clear reflections.
  • + Very clean, aesthetic composition.
  • Misses the 'small beads' in the braids, showing only a few white pins.
  • The character looks too clean and manicured for a 'battle-worn' paladin.
  • The armor engraving is much simpler than requested.

Verdict: LongCat-Image is the clear winner for its superior prompt adherence, particularly in capturing the gritty, 'battle-worn' aesthetic with detailed wounds and intricate armor engravings. While OmniGen v2 produced a beautiful, clean portrait, it failed to incorporate the specific 'beads' in the braids and felt too polished for the requested theme.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + High-quality food photography with realistic textures
  • + Vibrant pink accents create a distinct visual identity
  • + Good use of white space and professional margins
  • Text is largely nonsensical gibberish
  • The layout feels slightly cluttered with inconsistent font sizes
  • Missing specific requested sections like 'Mains'

OmniGen v2

  • + Cleaner, more minimalist layout following a clear grid
  • + Better text approximations for headings like 'Apptetizes' and 'Pizzzas'
  • + Strict adherence to the requested grid-based food photo style
  • Food photography looks more synthetic and generic
  • Significant spelling errors in every heading
  • Layout feels a bit sparse in the bottom sections

Verdict: OmniGen v2 provides a much stronger layout that aligns with the 'modern minimalist' request, utilizing a clean grid and organized sections that better approximate a real menu. While LongCat-Image has superior food photography, its text is completely illegible and the overall design feels less professional for a casual dining establishment.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent typography with a realistic fiery glow effect.
  • + High-quality photorealistic textures on the charcoal and ingredients.
  • + Strong adherence to the price and secondary text requirements.
  • Failed the 'exploded' instruction, showing the burger almost entirely assembled.

OmniGen v2

  • + Clean layout with a bright, energetic composition.
  • + Good integration of the fire motif into the background burst.
  • Failed the 'exploded' instruction, showing a static stacked burger.
  • Text rendering is slightly clipped and lacks the requested fiery glowing effect compared to Model A.
  • The burger looks more like a 3D render than a photorealistic image.

Verdict: Both models failed the specific 'exploded view' instruction, instead opting for a floating but assembled burger. LongCat-Image is the clear winner due to its superior text rendering, highly detailed photorealistic textures (especially the charcoal), and better execution of the fiery glow effect requested in the prompt.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent chalk texture with realistic dust and smudging on the board surface.
  • + Provides a convincing background environmental context of a cozy café.
  • + The handwriting has a strong, authentic chalk stroke feel.
  • Text is largely nonsensical and fails almost every spelling requirement.
  • Composition is messy with prices detached from their items.

OmniGen v2

  • + Successfully spelled most of 'TODAY SPECIALS' and the date correctly.
  • + Closer adherence to the menu items list including the 'Brown butter' request (rendered as Brown Chocolat Chip).
  • + Text is much more legible and follows a standard layout.
  • The chalk texture is overly clean and look more like a digital font.
  • The frame and board lack the realistic depth and texture of the first image.
  • Multiple spelling errors in the menu line items.

Verdict: Both models struggled with the complex text and specific handwriting requirements. LongCat-Image produced a much more realistic and aesthetically pleasing environment with authentic chalk textures, but the text was unreadable. OmniGen v2 followed the prompt's content much better and provided legible text, though it sacrificed the realistic 'handwritten' and environmental quality to do so.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent cinematic lighting and environmental detail.
  • + Dynamic composition with a sense of motion.
  • + Rich background elements including satellites and planetary bodies.
  • The horse has five legs, which is a significant anatomical error.
  • The prompt specified 'horse on top', but the astronaut is riding the horse.

OmniGen v2

  • + Clean, high-contrast visual style.
  • + Anatomically correct horse structure.
  • + Clear and well-defined astronaut suit.
  • Completely failed the negative constraint to have the horse on top of the astronaut.
  • The composition is static and lacks the cinematic depth requested.
  • Background is very simple and lacks detail.

Verdict: Both models failed the specific spatial logic instruction to have the horse on top of the astronaut. LongCat-Image provided a much more cinematic and detailed environment, but suffered from a major anatomical glitch with the horse's legs. OmniGen v2 produced a cleaner image with correct anatomy, but it was visually underwhelming compared to the requested cinematic style.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Features a highly realistic capybara texture and anatomy.
  • + Accurately places the capybara's front paw on the steering wheel.
  • + Includes multiple passengers with bored expressions consistent with the prompt.
  • The taxi sign placement on top of the car is slightly misaligned with the roof structure.

OmniGen v2

  • + Clearly depicts a yellow cap and a dark jacket as requested.
  • + Lighting on the car exterior is consistent and clean.
  • The driver has human hands instead of capybara paws.
  • The human passenger is partially holding the steering wheel, creating a logical error.

Verdict: LongCat-Image adheres much better to the prompt's unusual requirements by depicting the capybara with its own paws on the wheel. OmniGen v2 suffers from a major anatomical failure by giving the capybara human hands and placing the passenger in a position where she is overlapping with the driving controls.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent rendering of the primary title text
  • + Highly detailed and realistic 3D textures on the pumpkin and thorns
  • + Accurate depiction of the thorny border and parchment style
  • Small text at the bottom contains several spelling errors and repetitions
  • The layout feels slightly crowded with the large pumpkin in the center

OmniGen v2

  • + High contrast, clean graphic design style
  • + Strong spooky atmosphere with the moon and tree silhouettes
  • + Balanced layout with clear sections for text
  • Significant text errors, including gibberish in the scroll banner and misspelled location
  • The pumpkin is a flat illustration compared to the cinematic lighting of the other image
  • Failed to include thorns in the border as requested

Verdict: LongCat-Image provides a much more polished and 'cinematic' result with high-quality textures and an impressive thorny border, though it struggles with the smaller event details. OmniGen v2 has a nice compositional balance and clean typography for the title, but fails on the specific border requirements and contains more significant textual errors in the body.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent text rendering and placement that follows the prompt exactly.
  • + High-quality 3D clay-like textures with impressive subsurface scattering on the salmon.
  • + Accurately depicts the Japanese flag next to the text.
  • The 'diorama base' is a bit loose in interpretation, appearing more like a standard wooden tray.

OmniGen v2

  • + Successfully captures the 45-degree isometric diorama base specified in the prompt.
  • + Clean, vibrant cartoon aesthetic with balanced composition.
  • The flag icon is incorrect, showing three horizontal stripes instead of the Japanese flag.
  • The sushi design is anatomically confusing, blending nigiri-style tops with maki-style rolls.

Verdict: LongCat-Image provides superior adherence to the specific details of the prompt, particularly the flag icon and the text layout. While OmniGen v2 captures the isometric diorama base more literally, LongCat-Image's rendering of textures and correct iconography makes it the more polished and accurate result.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent realization of the 'god rays' and 'dew sparkles' requested in the prompt.
  • + Detailed fur texture and realistic lighting across all subjects.
  • + Captures the meadow atmosphere with high fidelity and depth of field.
  • Anatomical failure on the kitten, which has been given long rabbit ears, merging it with the bunny.
  • The four distinct animals requested are not all present as individuals (missing the bunny as a standalone creature).

OmniGen v2

  • + Better count of animals, though still missing a distinct fourth species.
  • + Vibrant colors and a high level of cuteness that fits a 'wholesome vibe'.
  • Missing the bunny entirely; only shows a dog, cat, and fox.
  • The art style leans toward CGI/illustration rather than the requested 'hyper-photorealistic' masterpiece.
  • Butterflies have simplified, non-photorealistic textures.

Verdict: LongCat-Image provides much higher technical quality and lighting effects, perfectly capturing the requested god rays and photorealistic fur, but it unfortunately fuses the cat and bunny into one creature. OmniGen v2 fails to meet the photorealistic requirement, producing a generic 3D-render aesthetic, and also fails to include all four requested animals. LongCat-Image is preferred for its superior visual fidelity despite the anatomical error.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Excellent typography rendering with the correct spelling and accent mark.
  • + Wonderful hand-drawn vintage texture and character.
  • + Creative integration of the cloche and banners into a cohesive emblem.
  • The word 'Caffè' is repeated twice unnecessarily.
  • The steam lines are slightly chaotic and overlap the top text.

OmniGen v2

  • + Successfully achieves a minimalist, clean vector style.
  • + Accurate interpretation of the 'Est. 1720' request on a plain background.
  • Failed the spelling of 'Caffè Florian', merging it into 'CAFFFLORIN'.
  • Lacks the 'vintage' character or texture requested in the prompt, appearing too modern.

Verdict: LongCat-Image captured the 'vintage' and 'subtle texture' requirements perfectly, delivering a high-quality emblem despite the redundant text. OmniGen v2 followed the minimalist instruction but failed significantly on the brand name spelling and lacked the artistic character expected from the prompt.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

LongCat-Image
OmniGen v2

AI Judge Analysis

LongCat-Image

  • + Stronger adherence to the specific icon requests like a large lunar module on the surface.
  • + Effective use of the NASA-inspired color palette with high-contrast outlines.
  • + Captures the 'flat-vector' style with more detailed and recognizable illustrations.
  • Text is largely nonsensical and jumbled.
  • The layout is a bit chaotic and doesn't clearly follow the 1-6 step progression requested.

OmniGen v2

  • + Layout more closely resembles a professional modern infographic with clean grids.
  • + Text rendering is slightly more legible, though still contains many typos.
  • + Accurately applies the minimalist 'clean vector' aesthetic.
  • Incorrectly identifies the mission as 'APOLO 17' instead of Apollo 11.
  • Failed to include most of the specific icons requested for the steps (e.g., missing Saturn V and trajectory arc).
  • The icons are generic and repetitive rather than descriptive of the mission stages.

Verdict: LongCat-Image provides much better illustrations that actually follow the prompt's content requirements, such as the lunar module and Earth, though the layout is less like a standard infographic. OmniGen v2 has a superior 'modern' layout and cleaner typography, but it completely fails the prompt instructions by referencing Apollo 17 and ignoring the specific step-by-step icon requirements.

Next steps

Explore each model