Head to head
Esc

Models · slot A

to navigate to pick

GPT Image 1 Mini OpenAI Imagen 4.0 Ultra Generate 001 Google

Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.

GPT Image 1 Mini

24.9 arena score

#13 of 62 in Text-to-Image

Skill signature · Text-to-Image

Imagen 4.0 Ultra Generate 001

21.9 arena score

#33 of 62 in Text-to-Image

Vote tally

Where the votes landed

GPT Image 1 Mini

0%

win rate

Ties

0%

Imagen 4.0 Ultra Generate 001

0%

win rate

Shared challenges 13

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent adherence to the 'plant behind the cube' instruction with realistic distortion
  • + High visual quality with realistic textures on the book and sphere
  • + Natural-looking soft window lighting
  • The sphere appears to be floating unnaturally without a clear support
  • The glass cube looks more like an empty frame or thin container than a solid object

Imagen 4.0 Ultra Generate 001

  • + Accurate rendering of a solid glass cube with realistic refractive properties
  • + Crisp text rendering on the book spine
  • + Dynamic lighting and shadows on the wooden table
  • The plant is more to the side than 'behind' the cube, missing the requested visual interaction
  • The blue sphere has a strange double-reflection/ghosting artifact
  • The sphere is floating in the center of a solid glass block, which is physically impossible

Verdict: GPT Image 1 Mini followed the spatial instructions much better, correctly placing the plant behind the cube so it is visible through the glass. While Imagen 4.0 Ultra Generate 001 produced a more beautiful solid glass texture and sharp text, it failed to place the plant behind the cube and included a distracting visual artifact next to the blue sphere.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent shallow depth of field and bokeh
  • + Strong cinematic mood with realistic lighting and reflections
  • + Coherent bicycle structure with logical components like the kickstand and basket

Imagen 4.0 Ultra Generate 001

  • + Exceptional skin texture and facial detail
  • + Dynamic composition with visible tools and better sense of action
  • + Captures 'motion blur from passing cars' more effectively
  • The red bicycle frame has anatomical issues, such as the down tube missing and pedals/gears being strangely placed
  • The white flecks on the ground look more like petals than rain reflections

Verdict: GPT Image 1 Mini produces a more coherent and aesthetically pleasing 'cinematic' image with realistic bicycle geometry, though it lacks the fine skin detail of the competition. Imagen 4.0 Ultra Generate 001 provides incredible detail in the subject's face and hands, but the bicycle is structurally nonsensical and the background blur feels less natural than the 50mm lens look achieved by GPT.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent moody lighting and atmospheric bokeh sparks.
  • + Superior skin texture with realistic dirt and subtle scarring.
  • + The armor engraving is intricate and feels integrated into the metal material.
  • Failed to include the beads in the hair braids as requested.
  • The leather straps are partially obscured and lack the specific detail requested.

Imagen 4.0 Ultra Generate 001

  • + Perfectly captured the beads in the braids component of the prompt.
  • + Highly detailed leather straps with visible texture and buckles.
  • + Balanced composition with the torch directly visible to explain the lighting.
  • The scars look like clean red lines rather than realistic, healed wounds.
  • The facial skin texture is slightly smoother and less 'battle-worn' than Model A.
  • The 'bokeh' effect is less pronounced, with sparks appearing as sharp dots.

Verdict: GPT Image 1 Mini produces a more evocative and cinematic portrait with superior skin and lighting realism, capturing the 'battle-worn' feel excellently but missing the hair beads. Imagen 4.0 Ultra is much more literal with prompt adherence, including every specific detail like the beads and leather straps, though the skin and scars feel slightly more artificial. Overall, GPT Image 1 Mini is preferred for its artistic quality and texture realism.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Perfect text rendering for section headers
  • + Extremely clean and organized grid layout
  • + Clear white background with high-contrast bold typography
  • Lack of descriptive menu item text or pricing
  • Left side of the layout feels slightly empty with just lines

Imagen 4.0 Ultra Generate 001

  • + Includes pricing and item names for a more realistic menu feel
  • + Highly vibrant and detailed food photography
  • + Good use of the grid across the entire page width
  • Text consists of nonsensical gibberish
  • The 'Pizza' header is centered while 'Appetizers' is left-aligned, creating imbalance
  • Some font rendering is blurry or inconsistent

Verdict: GPT Image 1 Mini creates a much cleaner, professional-looking template with perfect typography, though it lacks specific item details. Imagen 4.0 Ultra Generate 001 provides a more detailed restaurant simulation with prices and varied photos, but the text is garbled and the layout is less cohesive. GPT Image 1 Mini is preferred for its design clarity and adherence to the minimalist aesthetic.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography with a glowing neon texture
  • + Clean and photorealistic burger ingredients
  • + Strong contrast with the dark background
  • The 'exploded' effect is a bit static with minimal motion blur
  • The starburst shape is somewhat basic and geometric

Imagen 4.0 Ultra Generate 001

  • + Dynamic sense of motion with swirling background embers
  • + High level of detail in the food textures and sauce
  • + Creative use of a fiery, jagged starburst for the price
  • Main title text is slightly less readable than the other version
  • The composition feels a bit cramped at the top

Verdict: GPT Image 1 Mini provides a clean, professional ad layout with superior text rendering and clarity. However, Imagen 4.0 Ultra Generate 001 captures the requested 'sense of motion' much better through the swirling background and more complex lighting on the ingredients.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI judge analysis unavailable for this challenge.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent cinematic lighting and atmosphere
  • + Highly detailed textures on the spacesuit and horse hair
  • + Captures a subtle, surreal mood through monochromatic toning
  • Completely failed the negative constraint to put the horse on top of the astronaut
  • Common interpretation of the concept rather than the specific request

Imagen 4.0 Ultra Generate 001

  • + Beautifully vibrant colors and complex space background
  • + Creative horse-specific space gear like the visor and glowing horseshoes
  • + High resolution and clear composition
  • Failed the specific positional constraint to have the horse on top of the astronaut
  • Literal interpretation of riding rather than the requested inverted surrealism

Verdict: Both models completely failed the negative constraint and logic-defying request to have the horse on top of the astronaut. GPT Image 1 Mini creates a more moody, cinematic image, while Imagen 4.0 Ultra Generate 001 provides more detail in the equipment and a more vibrant planetary backdrop. Imagen 4.0 is slightly preferred for adding logical gear for the horse, even though both missed the core surreal instruction.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent photorealistic texture on the capybara's fur
  • + Cinematic lighting that feels like a real night-time taxi ride
  • + The human passenger's expression and activity match the prompt perfectly
  • The capybara only has one paw visible on the steering wheel instead of two

Imagen 4.0 Ultra Generate 001

  • + Successfully places both front paws on the steering wheel
  • + Very clear rendering and vibrant colors
  • + Detailed taxi interior and cap
  • The passenger is in the front seat instead of the back seat as requested
  • The capybara's paws look more like monster claws than natural capybara anatomy
  • The passenger's expression looks slightly distressed rather than bored and normal

Verdict: GPT Image 1 Mini captures the requested vibe and composition much better, placing the passenger correctly in the back seat and achieving a high level of photorealism. While Imagen 4.0 Ultra follows the specific 'two paws' instruction, it fails the basic composition by putting the passenger in the front and produces slightly jarring claw-like hands for the animal.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent vintage parchment texture and aesthetic.
  • + Consistent and sophisticated gothic typography.
  • + Central glowing jack-o-lantern has a dark, moody feel.
  • The 'dark' lighting makes some details like the border and trees harder to see.
  • The date uses a comma instead of a period (30.10.2026 became 30.10,2026).

Imagen 4.0 Ultra Generate 001

  • + Perfect text accuracy for all required fields.
  • + Highly detailed border featuring webs and thorns as requested.
  • + Vibrant composition with a clear moon and cinematic lighting.
  • The style feels more like a modern digital illustration than 'vintage gothic'.
  • The scroll banner is slightly warped on the left side.

Verdict: GPT Image 1 Mini captured the requested 'vintage gothic' mood and dark parchment texture much more effectively, creating a cohesive atmospheric piece. However, Imagen 4.0 Ultra Generate 001 followed the text instructions perfectly and included a more literal interpretation of the thorn-and-web border, though it resulted in a cleaner, more modern look.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent 3D miniature toy-like aesthetic with soft, rounded textures
  • + Clean, bold typography that integrates well with the overall design
  • + Precise isometric angle and very tidy composition
  • Simple sushi variety compared to what is possible with the prompt

Imagen 4.0 Ultra Generate 001

  • + Features a more diverse array of sushi types including rolls and roe
  • + Excellent lighting and material definition on the sushi pieces
  • + Followed the request for a diorama base more literally
  • The 'SUSHI' text is much smaller and less bold than requested
  • Floating plate artifact creates a disconnected visual between the plate and the base

Verdict: GPT Image 1 Mini provides a more cohesive 'graphic design' look with superior typography and a very consistent cartoon aesthetic. While Imagen 4.0 Ultra Generate 001 offers better food variety and lighting details, its layout suffers from a floating plate artifact and weaker adherence to the text hierarchy instructions. GPT Image 1 Mini is the better overall image for its clean, professional execution of the requested isometric style.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent photorealism with natural fur textures and lighting.
  • + Dynamic and convincing movement, capturing the animals in a playful romp.
  • + Subtle and realistic integration of dew sparkles and god rays.
  • The kitten has a slightly strange fifth leg or tail-like protrusion between its front legs.
  • Butterflies are less varied and fewer in number compared to the other model.

Imagen 4.0 Ultra Generate 001

  • + Perfect adherence to all requested animals with distinct, expressive poses.
  • + Vibrant and colorful composition with a high variety of butterflies and flowers.
  • + Very clear 'dew sparkles' on the grass as requested in the prompt.
  • Leans toward a digital illustration or 'hyper-real' CGI style rather than true photorealism.
  • The fox's anatomy, particularly the front paws, feels a bit stylized and stiff.

Verdict: GPT Image 1 Mini produces a much more convincing photorealistic image with beautiful natural lighting, though it suffers from a minor anatomical glitch on the kitten. Imagen 4.0 Ultra Generate 001 creates a charming, storybook-like scene that perfectly captures every element of the prompt but lacks the realistic depth and texture of the former, appearing more like a high-end 3D render.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent typography rendering with the correct accent on 'Caffè'.
  • + Accurate date and text inclusion on the banner.
  • + Good use of gold/brown tones with a textured gold-leaf effect.
  • Ignored the request for a 'light background', opting for solid black instead.
  • The banner style is a bit chunky and less elegant than typical vintage emblems.

Imagen 4.0 Ultra Generate 001

  • + Perfectly followed the request for a light background with subtle texture.
  • + Clean vector-style execution with elegant line work.
  • + Sophisticated banner design and balanced composition.
  • Missing the grave accent on 'Caffè' (rendered as 'Caffe').
  • Slightly less 'warmth' in the color palette compared to the prompt's request for warm brown tones.

Verdict: Imagen 4.0 Ultra followed the layout and background requirements of the prompt much better than GPT Image 1 Mini, which completely ignored the instruction for a light background. While GPT Image 1 Mini handled the character accent in 'Caffè' more accurately, Imagen 4.0 Ultra's overall aesthetic is more aligned with the 'vintage minimalist' and 'vector emblem' style requested.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

GPT Image 1 Mini
Imagen 4.0 Ultra Generate 001

AI Judge Analysis

GPT Image 1 Mini

  • + Excellent text legibility and accuracy for all listed steps.
  • + Clean, modern vector aesthetic perfectly matches the flat-vector style requested.
  • + Precise adherence to the requested NASA-inspired color palette.
  • The 'Translunar' iconography is a bit abstract and loopy compared to the others.
  • Layout is a bit sparse with significant empty space at the bottom.

Imagen 4.0 Ultra Generate 001

  • + Sophisticated composition and layout that feels like a professional poster.
  • + Includes a clear title and header as part of the infographic design.
  • + Good use of the color palette across the entire canvas.
  • Severe 'hallucination' of text, creating nonsensical words like 'BEOMBERS' and 'SPONDRES'.
  • Failed to follow the specific 6-step sequence requested in the prompt.
  • The iconography is messy and indistinct compared to the clean vector request.

Verdict: GPT Image 1 Mini is the clear winner because it successfully followed the complex instructional sequence of the prompt with perfect text rendering. While Imagen 4.0 Ultra produces a more visually dense and 'designed' layout, its text is completely illegible and it failed to include the specific icons and steps requested.

Next steps

Explore each model