Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1 OpenAI OmniGen v2 VectorSpaceLab

Settled by community votes across 15 shared challenges, with an AI judge weighing in on each.

GPT Image 1

23.2 arena score

#28 of 62 in Text-to-Image

Skill signature · Text-to-Image

OmniGen v2

16.8 arena score

#57 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1

100.0%

win rate

Ties

0.0%

OmniGen v2

0.0%

win rate

100.0% 0.0% ties 0.0%
Shared challenges 15

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent adherence to the 'partially visible through the glass' instruction for the plant.
  • + Highly realistic glass textures and refractions around the edges.
  • + Accurate lighting direction and soft shadows.
  • The glass cube appears to have an open front or missing face rather than being a solid closed object.

OmniGen v2

  • + Successfully includes all requested elements in a clean composition.
  • + Beautiful glossy texture on the blue sphere with realistic reflections.
  • + Sharp focus on the central subject.
  • The plant is positioned entirely above the cube, missing the requested effect of being visible through the glass.
  • The glass cube looks a bit more like a thick acrylic box than glass.

Verdict: GPT Image 1 is the winner because it successfully captured the complex spatial relationship of seeing the plant through the glass cube, whereas OmniGen v2 placed the plant entirely behind and above the cube. GPT Image 1 also demonstrated superior realism in the glass refraction and the texture of the book pages.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent skin texture and realistic, aged features on the subject.
  • + Captures the requested 'imperfect framing' and 50mm feel perfectly.
  • + Authentic atmosphere with subtle rain, wet textures, and natural lighting.
  • The rear bicycle structure is a bit anatomically messy around the chain guard area.
  • Missing significant motion blur from the passing cars as requested.

OmniGen v2

  • + Strong reflections on the wet pavement.
  • + Clearer representation of a full bicycle.
  • Skin texture appears smoothed and lacks the 'natural skin texture' requested.
  • Composition feels like a staged studio portrait rather than a 'candid street photo'.
  • Rain effect looks like a simple digital overlay of streaks rather than environmental moisture.

Verdict: GPT Image 1 is the superior image as it adheres closely to the requested photographic style, providing a gritty, candid, and realistic atmosphere with impressive facial detail. OmniGen v2 fails to capture the 'no stylization' and 'candid' aspects, instead producing a clean, artificial-looking image with flat textures and poorly integrated rain effects.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent depiction of a battle-worn character with realistic skin texture and grit
  • + Highly detailed and realistic engraving on the plate armor
  • + Atmospheric lighting that feels natural to the environment
  • The beads in the hair are very dark and blend into the shadows

OmniGen v2

  • + Clearly visible beads in the braided hair as requested
  • + Clean and vibrant bokeh effect in the background
  • + Good focus on the lifelike quality of the eyes
  • The character looks too clean and polished for a 'battle-worn' description
  • Dirt marks look like flat digital face paint rather than actual grime
  • Armor texture looks slightly plastic compared to Model A

Verdict: GPT Image 1 captures the 'battle-worn' essence of the prompt far more effectively, featuring realistic skin imperfections and intricately textured, weathered armor. While OmniGen v2 displays the hair beads more clearly, its subject appears too pristine and the armor lacks the convincing metallic depth found in GPT Image 1.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent typography and readability with clean sans-serif fonts
  • + High-quality, appetizing food photography that looks professional
  • + Clear layout with relevant sections that match the prompt
  • The placeholder text contains spelling errors like 'descrigion'
  • The grid is a bit tight with limited white space between the photos and text

OmniGen v2

  • + Successfully creates a double-page spread menu layout
  • + Uses vibrant color blocks and accents as requested
  • + Better representation of a 'grid' layout for the overall design
  • Text is largely nonsensical gibberish or has severe spelling errors
  • The resolution of the food photos is lower compared to Model A
  • The fonts used are more stylistic/rustic rather than a 'bold sans-serif'

Verdict: GPT Image 1 is the superior choice because it provides professional-grade food photography and legible, structured typography that fits a modern minimalist aesthetic. While OmniGen v2 captures the 'grid' and 'vibrant accents' aspects of the prompt well, its failure to generate readable text and its lower image clarity make it less useful as a design mock-up.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent photorealistic texture on the meat and bun
  • + Successfully rendered all requested text elements with the corect font style
  • + Captures the 'exploded' suspended animation look perfectly
  • The price text is missing the leading '6' in the starburst

OmniGen v2

  • + Bold, high-contrast graphic design suitable for a cartoonish ad
  • + Corretly includes the full '6.99' numerical value in the starburst
  • Failed the 'exploded burger' requirement; the burger is fully assembled
  • Text is partially cut off on the left side
  • Lacks the photorealism requested in the prompt, appearing more like a 3D render

Verdict: GPT Image 1 followed the complex layout instructions much better, delivering a truly 'exploded' burger with high photorealism and consistent fiery text effects. OmniGen v2 failed to separate the burger components and had significant layout issues with the secondary text being cut off, while also opting for a more plastic, non-photorealistic art style.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent text legibility and spelling throughout the entire menu.
  • + Realistic chalk texture with convincing grainy artifacts.
  • + Consistent handwriting style that feels authentic to a café setting.
  • The title is in a print-style sans-serif rather than the requested 'elegant cursive'.
  • The layout is a bit sparse with significant empty space at the sides.

OmniGen v2

  • + Includes a wooden frame which adds to the 'cozy café' atmosphere.
  • + Attempts a more varied handwriting style for the title and body text.
  • Numerous spelling errors and gibberish text throughout the list.
  • Poorly rendered letterforms with overlapping lines and messy corrections.
  • Failed to follow the specific menu item text precisely.

Verdict: GPT Image 1 is the clear winner as it successfully rendered almost all the requested text with near-perfect spelling and a very realistic chalk aesthetic. While OmniGen v2 included a nice frame, the actual content of the board became an illegible mix of typos and artifacts that failed the core prompt requirements.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent cinematic lighting and textures.
  • + Highly detailed rendering of the spacesuit and horse anatomy.
  • Failed the core prompt instruction to have the horse on top of the astronaut.

OmniGen v2

  • + Clean, vibrant colors.
  • + High contrast and sharp lines.
  • Failed the core prompt instruction to have the horse on top of the astronaut.
  • Visual style is more illustrative than 'cinematic' or 'surreal'.

Verdict: Both GPT Image 1 and OmniGen v2 failed the specific spatial logic request of placing the horse on top of the astronaut, instead defaulting to the common trope of an astronaut riding a horse. GPT Image 1 is the superior image due to its atmospheric cinematic quality and realistic textures, whereas OmniGen v2 looks more like a standard digital illustration.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent photorealism in texture and lighting
  • + Captures the professional, calm expression of the capybara perfectly
  • + Accurate depiction of a capybara's anatomy for the paws on the wheel
  • The view is from outside looking in, while the prompt requested a scene 'inside'

OmniGen v2

  • + Dynamic composition and angle
  • + Good interpretation of the driver cap
  • Anatomical failure where human hands are growing out of the capybara's arms
  • Perspective issues where the capybara seems to be floating in front of the seat
  • The businesswoman is in the front passenger seat rather than the requested back seat

Verdict: GPT Image 1 far exceeds OmniGen v2 in both realism and prompt adherence. While GPT Image 1 successfully creates a photorealistic and believable scene with correct animal anatomy, OmniGen v2 contains significant errors, most notably the jarring human hands attached to the capybara and placing the passenger in the front seat.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent typography with perfect spelling of all requested text.
  • + Sophisticated, authentic vintage gothic aesthetic with moody lighting.
  • + Clear and balanced composition that functions well as a real invitation.
  • The parchment texture is a bit dark, making the tree silhouettes blend into the background.

OmniGen v2

  • + High contrast lighting makes the jack-o-lantern and moon pop.
  • + Creative use of a deckled-edge parchment border effect.
  • Significant spelling errors and garbled text on the banner and bottom details.
  • The visual style is more like a modern digital illustration than a vintage gothic poster.
  • Poor layout of event details, resulting in a cluttered and confusing bottom section.

Verdict: GPT Image 1 is far superior for this task, as it followed all text instructions perfectly with high-quality, readable typography. OmniGen v2 struggled with the text rendering, producing several illegible words and failing to capture the 'vintage gothic' mood, opting instead for a bright, cartoony style. GPT Image 1 successfully created a polished, cinematic piece that looks like a professional invitation.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent typography with clean, bold text and a correct Japanese flag icon.
  • + Higher quality 3D textures with realistic soft-touch PBR materials.
  • + Superior object modeling, especially in the rice grains and sushi garnish.
  • The perspective is slightly more focused/closer than a true 'miniature' diorama look.

OmniGen v2

  • + Stronger adherence to the 'isometric' diorama layout with a distinct raised base.
  • + Clean, bright color palette that pops against the blue background.
  • Incorrect flag icon that does not resemble the Japanese flag.
  • Visible AI artifacts in the sushi details, such as the shrimp tail merging into nigiri.
  • Text has slight rendering errors and inconsistent shadows.

Verdict: GPT Image 1 is the clear winner due to its professional-grade typography and high-fidelity 3D rendering. While OmniGen v2 captures the requested isometric diorama layout better, it fails on key details like the flag icon and exhibits significant anatomical errors in the food items.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1
OmniGen v2
100% wins 0% ties 0% wins

AI Judge Analysis

GPT Image 1

  • + Excellent adherence to the 'hyper-photorealistic' instruction with realistic fur textures and lighting.
  • + Correctly includes all four requested animals: puppy, kitten, bunny, and fox kit.
  • + Dynamic composition that captures the 'tumbling' and 'chasing' actions described.
  • One butterfly has slightly clipped wings in the upper left.
  • The kitten's anatomy is a bit merged with the puppy's leg area.

OmniGen v2

  • + Bright, vibrant colors that fit a 'wholesome' vibe.
  • + Clear, centered composition suitable for a greeting card.
  • Failed to include the bunny, only showing three animals.
  • Style is closer to 3D animation/clipart rather than the requested 'hyper-photorealistic' look.
  • The animals are sitting still rather than 'chasing' or 'tumbling'.

Verdict: GPT Image 1 successfully followed all aspects of the prompt, including the specific list of four animals and the requested photorealistic style. OmniGen v2 missed one of the animals (the bunny) and produced a stylized, cartoon-like image that ignored the 'hyper-photorealistic' requirement.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent adherence to the 'hand-painted texture' and 'dreamy' lighting request
  • + Perfectly captures the Studio Ghibli watercolor aesthetic
  • + Preserves the characters' expressions and the specific composition of the original meme.
  • Image looks slightly washed out compared to modern digital anime styles.

OmniGen v2

  • + Clean, high-resolution digital illustration style
  • + Preserves details like the plaid pattern on the shirt and the background bus very well.
  • Fails to capture the 'Studio Ghibli' look, opting for a generic modern anime style
  • Loses the key narrative of the meme by making the girlfriend look happy instead of angry
  • Does not follow the request for soft pastel colors or hand-painted textures.

Verdict: GPT Image 1 successfully transformed the photo into a Studio Ghibli-inspired piece by perfectly capturing the soft watercolor textures and nostalgic mood requested. In contrast, OmniGen v2 produced a generic digital anime style that missed the specific textural requirements and altered the characters' emotions, ruining the context of the original 'distracted boyfriend' meme.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
GPT Image 1
Before After
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent source preservation, maintaining the woman's face and dog's appearance perfectly.
  • + Logical hair movement that flows naturally from the head.
  • + Leaves appear integrated into the original scene's lighting.
  • The wind effect on the dog's fur is slightly subtler than requested.

OmniGen v2

  • + Successfully adds a clear wind-blown effect to the hair.
  • + Leaves are bright and highly visible.
  • Failed to preserve the woman's facial identity, significantly changing her appearance.
  • Leaves look like stickers or icons rather than natural elements of the scene.
  • Noticeable loss of detail and texture in the dog's fur compared to the source.

Verdict: GPT Image 1 is the clear winner as it successfully applies the requested motion and leaves while maintaining the high quality and identity of the source image. OmniGen v2 significantly alters the woman's face and background details, and the added leaves have a flat, artificial appearance that lacks integration.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent typography with perfect spelling and correct accent marks.
  • + Clean vector-style execution with a subtle, professional texture.
  • + Balanced layout that feels authentic to the vintage prompt.
  • Ignored the 'light background' instruction, delivering a black background.

OmniGen v2

  • + Followed the light/cream background instruction correctly.
  • + Good use of the banner element for the 'Est. 1720' text.
  • + Clean, minimalist line work on the cloche.
  • Major spelling error in the main brand name ('CAFFFLORIN').
  • Typographic hierarchy is slightly cluttered with overlapping elements.
  • Steam effect is very thin and lacks visual weight compared to the rest of the logo.

Verdict: GPT Image 1 failed the background color instruction but delivered a much higher quality logo with perfect spelling and superior typography. OmniGen v2 followed the color scheme but failed significantly on the primary brand name spelling and overall composition.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1
OmniGen v2

AI Judge Analysis

GPT Image 1

  • + Excellent text rendering with names and phases correctly spelled
  • + Followed the flat-vector style and specific NASA color palette perfectly
  • + Logical layout that visualizes the progression of the mission mission phases
  • Included a typo in the bottom bar ('EARLLUNAR')
  • Disconnected 'Lunar Orbit' and 'Descent' icons from their labels

OmniGen v2

  • + Clean icon designs within circular frames
  • + Modern, bold typographic header
  • Extremely poor text rendering with numerous gibberish words and spelling errors
  • Incorrect mission number (Apollo 17 instead of 11)
  • Failed to follow the specific 6-step sequence requested in the prompt

Verdict: GPT Image 1 followed the complex multi-step prompt almost perfectly, providing a clear infographic with legible text and accurate historical names. OmniGen v2 struggled significantly with text generation, provided fewer steps than requested, and misidentified the mission as Apollo 17.

Next steps

Explore each model