OpenAI's previous image generation model that accepts both text and image inputs and produces image outputs
Settled by community votes across 16 shared challenges, with an AI judge weighing in on each.
GPT Image 1
#32 of 62 in Text-to-Image
HiDream I1 Full
#60 of 62 in Text-to-Image
Where the votes landed
GPT Image 1
0%
win rate
Ties
0%
HiDream I1 Full
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to lighting instructions with a soft window light from the left.
- + Highly realistic textures on the wooden table and the book cover.
- + Clean, believable glass physics and reflections.
- − The sphere appears to be floating without a clear reason or shadow beneath it.
- − The glass cube has thick, greenish edges that look more like an aquarium frame than a solid glass cube.
HiDream I1 Full
- + Excellent depiction of the plant seen through the glass, effectively showing refraction.
- + The sphere has a nice pearlescent texture.
- + Creative use of shadows and light patches on the table surface.
- − The sphere is floating unnaturally in the center of the cube.
- − The perspective of the cube is slightly skewed, making it look tilted relative to the table.
- − Duplicate blue reflections or orbs appear on the left and right sides of the cube, which are physically confusing.
Verdict: Both models followed the complex spatial prompt accurately. GPT Image 1 is preferred because its composition feels more grounded and the photographic quality of the textures is superior. While HiDream I1 Full handles refractive properties well, it suffers from strange artifacts like ghost spheres and a slightly distorted cube shape.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1
- + Excellent natural skin texture and facial realism
- + Realistic lighting and wet reflections on the pavement
- + Strong adherence to the shallow depth of field and color palette
- − The bike anatomy is slightly illogical where the hand is working
- − Missing significant motion blur on passing cars
HiDream I1 Full
- + Good full-body composition and environment wide shot
- + Beautiful reflections on the wet asphalt
- − Physical scale error where the man is larger than the bike
- − Anatomical defects in the hands
- − Skin texture appears overly smoothed and artificial
Verdict: GPT Image 1 significantly outperforms HiDream I1 Full in terms of realism and photographic quality. While both models struggled to include the specific 'motion blur' requested, GPT Image 1 provides a much more convincing 'no stylization' look with natural skin and better lighting, whereas the subject in HiDream I1 Full looks poorly integrated into the scene and out of scale with the bicycle.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1
- + Exceptional skin texture with realistic dirt and faint scarring
- + Subtle and sophisticated torchlight lighting that interacts naturally with the environment
- + Extremely detailed engraving on the plate armor that looks aged and worn
- − The 'small beads' in the hair are few and very understated
- − Slightly less 'paladin' aesthetic and more of a gritty mercenary feel
HiDream I1 Full
- + Clearer adherence to the 'beads' part of the hair prompt
- + Vibrant use of bokeh sparks and a visible torch in the background
- + Good focus on leather strap details as requested
- − The scars look like artificial red lines painted on the face
- − Strong over-sharpening artifacts around the hair and eyes
- − Anatomically awkward eyes with messy iris patterns
Verdict: GPT Image 1 is the superior image due to its incredible photorealism and sophisticated lighting, capturing a truly 'battle-worn' look without appearing cartoonish. While HiDream I1 Full followed specific prompts like beads more literally, the over-processed facial features and artificial-looking scars significantly detract from its visual quality.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1
- + Excellent high-resolution food photography with appealing colors.
- + Strong grid layout that feels professionally balanced.
- + Accurate text placement for categories and prices.
- − Nonsense filler text in the descriptions ('Apperoiation descrigion').
- − Missing a menu title at the top.
HiDream I1 Full
- + Includes a clear central title for the document.
- + Follows the white background and minimalist aesthetic.
- − Repetitive food photos, showing four very similar pizzas instead of a variety of dishes.
- − Lower visual quality on text rendering with significant artifacts.
- − Nonsense words and inconsistent pricing formatting.
Verdict: GPT Image 1 is the clear winner as it provides high-quality, diverse food photography that matches the 'appetizers/pizza/mains' prompt, despite using filler text. HiDream I1 Full fails on variety, repeating nearly identical pizza images in every section and producing garbled text with more artifacts.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1
- + Excellent text accuracy following the specific menu items prompt
- + Realistic chalk texture with visible grain and dust
- + Coherent handwriting style across all lines
- − The 'elegant cursive' request for the title was not fully met, as it appears more like block print
- − The layout is a bit cramped with text cutting off at the bottom edges
HiDream I1 Full
- + Successfully applied a cursive style to the title text
- + Good environmental atmosphere and framing of the chalkboard
- + Chalk illustration of the coffee cup is high quality
- − Failed almost entirely on the specific menu items, producing nonsensical gibberish
- − Prices are repetitive and illogical (many $50 items)
- − The text looks more like a digital font overlay than hand-drawn chalk
Verdict: GPT Image 1 followed the complex text instructions perfectly, rendering the specific menu items and prices requested with highly realistic chalk physics. While HiDream I1 Full captured a more 'elegant cursive' title and better framing, it failed the core task of rendering legible and accurate menu content, resulting in gibberish text.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1
- + Excellent texture on the astronaut's suit and the horse's coat.
- + Cinematic lighting and high-quality artistic rendering of the background space.
- − Failed the core prompt instruction to have the horse on top of the astronaut.
- − Anatomical issues with the horse's back hooves.
HiDream I1 Full
- + Bright, clear colors and sharp contrast.
- + Dynamic sense of motion in the horse's pose.
- − Failed the core prompt instruction to have the horse on top of the astronaut.
- − Visible artifacts around the horse's tail and hooves.
- − General 'ai-glossy' look lacks the requested cinematic detail.
Verdict: Both GPT Image 1 and HiDream I1 Full failed to follow the specific surreal instruction for the horse to be 'on top' of the astronaut, instead providing the standard astronaut-on-horse interpretation. GPT Image 1 is the better image overall due to its superior lighting, texture, and sophisticated cinematic composition, whereas HiDream I1 Full has noticeable artifacts and a flatter aesthetic.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1
- + Excellent photorealistic texture on the capybara fur.
- + Better lighting coherence between the interior and the blurry NYC cityscape.
- + The capybara's expression and paw placement on the wheel feel very natural and professional.
- − The passenger's hand holding the phone looks slightly distorted.
- − The yellow cap is a bit oversized for the capybara's head.
HiDream I1 Full
- + Successfully includes the requested beige/tan coat for the businesswoman.
- + Stronger contrast and brighter colors make the image pop.
- + Accurate placement of the paws on the steering wheel.
- − The capybara's head looks 'copy-pasted' onto a human body rather than integrated naturally.
- − The passenger is sitting in the middle of the back seat rather than behind the driver, making the perspective feel off.
- − The lighting on the capybara's face is overly bright compared to the dark cabin.
Verdict: GPT Image 1 is the superior image due to its consistent cinematic lighting and realistic integration of the animal into the scene. While HiDream I1 Full followed the 'coat' instruction well, its composition feels fragmented and less photorealistic than GPT Image 1, which perfectly captures the requested 'bored' atmosphere.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent text rendering with almost no spelling errors
- + Strong cinematic lighting and vintage gothic atmosphere
- + Perfect adherence to the requested layout including the scroll banner
- − Incorrectly labeled the location line with the word 'TIME'
- − The color palette is very dark, making some background details hard to see
HiDream I1 Full
- + High contrast makes the central elements pop
- + Good interpretation of the torn parchment aesthetic
- − Significant text hallucinations and gibberish at the bottom
- − Failed to include the specific date, time, and location requested
- − The art style is more like a modern clip-art vector than a vintage cinematic poster
Verdict: GPT Image 1 followed the prompt's instructions for text and atmospheric details much more effectively than HiDream I1 Full, which suffered from significant text errors and missed the specific event details. While GPT Image 1 made a minor labeling error on the last line, it successfully captured the 'vintage gothic' mood, whereas HiDream I1 Full felt like a generic digital illustration.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1
- + Successfully added a full head of hair that follows the head shape.
- + Mostly preserved the facial identity and lighting of the subject.
- + Maintained the original background and clothing perfectly.
- − The added hair has a slightly unnatural, wig-like texture at the crown.
- − There is a slight modification to the forehead wrinkles that looks a bit smeared.
HiDream I1 Full
- + Preserved the aesthetic style of the original image.
- − Failed to add a full head of hair, only adding growth to the back like a mullet.
- − Completely changed the background from a desert to a farm with wind turbines.
- − Added strange artifacts, including a black object protruding from the beard and a large backpack not in the original.
- − Significantly altered the subject's face and removed his glasses.
Verdict: GPT Image 1 followed the instructions nearly perfectly, adding a full head of hair while preserving the identity and composition of the source image. In contrast, HiDream I1 Full failed on almost every metric, hallucinating a new background, adding strange objects, and failing the primary request to provide a full head of hair.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent text rendering with 'JAPAN' and 'SUSHI' clearly legible.
- + Includes the requested small flag icon.
- + Higher quality PBR material feel with subtle textures on the base and food.
- − The salmon nigiri texture looks slightly more like plastic than soft food.
HiDream I1 Full
- + Good variety of sushi types on the diorama.
- + Accurate isometric perspective.
- − Missed the request for a small flag icon.
- − The text 'SUSHI' is significantly smaller and less balanced compared to the first image.
- − Presence of minor artifacts on the left edge of the frame.
Verdict: GPT Image 1 followed the prompt more precisely by including all elements, including the small flag icon and the specific text hierarchy. HiDream I1 Full produced a nice variety of sushi but failed to include the flag and had slightly less refined typography. GPT Image 1's use of PBR-style textures and clean composition makes it the more polished and accurate response.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to all prompt elements including news anchor desk, dog, and hockey references.
- + Effective use of caricature style with exaggerated facial features while staying recognizable.
- + Clear, high-quality watercolor-like illustration style.
- − The hands holding the papers are somewhat anatomically messy.
HiDream I1 Full
- + Successfully maintains the subject's likeness in a digital illustration style.
- + Integrates a dog and a subtle hockey reference in the background screen.
- − Weak execution of the 'caricature' aspect compared to typical expectations.
- − Fails to clearly represent the 'tv show anchor' profession, showing more of a living room setting.
- − The dog appearing to sit inside the jacket looks physically awkward.
Verdict: GPT Image 1 is the superior caricature as it fully embraces the requested theme, placing the subject at a news desk with clear hockey and dog integration. HiDream I1 Full provides a more conservative digital portrait that largely ignores the professional 'tv show anchor' setting and only subtly touches on the hockey interest.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1
- + Excellent action-oriented composition that captures the requested 'tumbling' and 'chasing' behavior.
- + Realistic fur textures and anatomical proportions for all four distinct animals.
- + Strong adherence to lighting requirements with visible god rays and backlighting.
- − The fox kit has dark paws that look slightly muddy/blurry compared to the rest of the body.
HiDream I1 Full
- + High saturation and vibrant colors that enhance the 'joyful' vibe.
- + Very clear, large eyes on the puppy and fox.
- − Failed to include the requested baby bunny, instead generating two kittens.
- − The pose is static and seated, missing the 'chasing' and 'tumbling' action requested.
- − The kittens have slightly distorted facial features and lack the realism found in Model A.
Verdict: GPT Image 1 is the clear winner as it correctly followed the prompt to include all four specific animals (puppy, kitten, bunny, and fox) and successfully captured the dynamic energy of them playing. HiDream I1 Full failed on prompt adherence by omitting the bunny and providing a static, posed shot rather than an action scene.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the original poses and facial expressions.
- + Strong adherence to the 'hand-painted textures' and 'pastel colors' request.
- + Accurately captures the 'distracted boyfriend' meme composition while applying the Ghibli style.
- − The characters' eyes are a bit simplified, losing some of the Ghibli-specific character design nuances.
- − Background is very blurry/dreamy, lacking the detailed environments typical of Ghibli films.
HiDream I1 Full
- + High-quality character designs that closely mimic the specific modern Ghibli aesthetic.
- + Rich, detailed background with Japanese signage that adds to the 'Studio Ghibli' atmosphere.
- − Fail to preserve the core narrative of the source image; the man is smiling at the woman in the red dress rather than looking back in secret.
- − The woman on the right has lost the 'jealous/angry' expression central to the meme's meaning.
- − The man's beard and clothing pattern are departures from the source image's identity.
Verdict: GPT Image 1 is the clear winner because it successfully transforms the style while preserving the specific poses and 'story' of the source image. HiDream I1 Full produces a high-quality illustration in the correct style, but it completely changes the character interactions and loses the 'distracted boyfriend' narrative requested by the edit prompt.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the subject's identity, clothes, and overall scene layout.
- + Perfectly adds wind-blown hair and flying leaves while keeping the original dog and background elements.
- − The leaf rendering is a bit sparse and lacks motion blur to signify speed.
HiDream I1 Full
- + Captures a very high level of energy and 'dynamic' motion as requested.
- + The leaves have a nice variety in size and color.
- − Fails as an edit by completely changing the person's face and features.
- − Completely removes the dog from the scene, failing source preservation.
Verdict: GPT Image 1 succeeded in the editing task by modifying the existing elements (hair and environment) while keeping the subject and the dog perfectly intact. HiDream I1 Full failed the core requirement of image editing by generating a completely new person and removing the dog, though it did capture a more energetic feel.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1
- + Accurately includes all requested text including 'Caffè Florian'
- + Excellent vintage typography and proper use of the accent mark
- + Strong minimalist vector aesthetic with a subtle grain texture
- − Failed the light background instruction, providing a black background instead
- − The steam effect is very simple compared to the cloche detail
HiDream I1 Full
- + Adhered to the light cream background and warm brown tone instructions
- + Clean vector illustration style
- + Correctly included the 'Est. 1720' banner
- − Completely omitted the main brand name 'Caffè Florian'
- − Included an extra coffee cup element not requested in the prompt
- − The banner and cloche are slightly misaligned
Verdict: GPT Image 1 followed the complex text requirements perfectly and captured a superior vintage typographic feel, though it failed to provide the light background requested. HiDream I1 Full adhered to the color palette and background instructions but failed the most critical part of the prompt by omitting the restaurant name entirely. GPT Image 1 is the winner as its high-quality typography and complete text adherence make it a functional logo, whereas the other is an incomplete design.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1
- + Follows the specific request for icons for each of the six steps fairly well.
- + Excellent adherence to the color palette and flat-vector style requirements.
- + Clear, legible typography for the names of the astronauts and the mission phases.
- − The layout is somewhat cluttered and non-linear, making the flow of steps hard to follow.
- − Contains a spelling error ('EARLLUNAR').
HiDream I1 Full
- + Stronger overall graphic design and composition with a centered hero element.
- + Perfectly clean vector aesthetics and high-quality title typography.
- + Good use of negative space and balance.
- − Failed to include all requested mission steps, missing several icon phases.
- − Included internal prompt text in the final design (e.g., '+ ICON', '+ UNG').
- − Misinterpreted the translunar step as a shuttle-style craft which is historically inaccurate.
Verdict: GPT Image 1 followed the complex multi-step instructions much more accurately, attempting to depict all six phases of the mission despite some layout issues and a minor typo. HiDream I1 Full produced a more professional-looking graphic from a pure design standpoint, but it failed to follow the content requirements, hallucinating text from the prompt and omitting several requested steps. GPT's adherence to the logical progression of the mission makes it the more successful tool for this specific task.
Explore each model
HiDream AI's 17B parameter text-to-image model using sparse diffusion transformer with mixture of experts, achieving state-of-the-art image generation quality with strong prompt following