OpenAI's previous image generation model that accepts both text and image inputs and produces image outputs
Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.
GPT Image 1
#28 of 62 in Text-to-Image
HiDream I1 Fast
#51 of 62 in Text-to-Image
Where the votes landed
GPT Image 1
0%
win rate
Ties
0%
HiDream I1 Fast
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to lighting instructions with clear soft window light from the left.
- + Highly realistic textures on the wood table, glass edges, and matte sphere.
- + Perfect spatial arrangement of the plant behind and visible through the glass.
- − The glass cube has a metallic base plate that wasn't explicitly requested.
HiDream I1 Fast
- + Successfully includes all requested elements in the frame.
- + The blue sphere has realistic glossy reflections.
- + Great bokeh effect on the background plant.
- − The glass cube is missing its back right vertical edge, making it look physically impossible.
- − Lighting direction is flat and does not clearly originate from the left as requested.
- − The sphere appears to be floating unnaturally without a point of contact.
Verdict: GPT Image 1 significantly outperforms HiDream I1 Fast by following lighting directions and maintaining structural integrity. GPT Image 1's cube is physically coherent and the textures are crisp, whereas HiDream I1 Fast fails to render all edges of the glass cube and misses the directional lighting requirement.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1
- + Excellent skin texture and realistic, candid facial expression.
- + High-quality rendering of raindrops on the bicycle and wet pavement reflections.
- + Superior adherence to the 50mm shallow depth of field request.
- − The bike structure becomes a bit abstract and nonsensical in the lower chain area.
- − Missing significant motion blur from cars, though lights suggest background traffic.
HiDream I1 Fast
- + Great overall street composition and environment mood.
- + Better depiction of the red bicycle as a whole object.
- + Captures the 'light rain' atmosphere well with reflective ground.
- − The man appears to be floating or sitting on air rather than the bike or a seat.
- − Skin texture looks slightly smoothed and less 'natural' than Model A.
- − Lacks the requested 'imperfect framing' or 'motion blur' from cars.
Verdict: GPT Image 1 is the winner due to its superior photographic realism and adherence to the character-specific details like natural skin texture and a candid feel. While HiDream I1 Fast creates a nice atmospheric scene, the man's seating position is physically impossible, appearing to hover over the frame of the bike, which breaks the realism requested.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1
- + Naturalistic textures on the skin and armor
- + Excellent implementation of warm torchlight and subtle reflections
- + Mood matches the 'battle-worn' description perfectly
- − The beads in the hair are very subtle and blend in
- − Lacks significant visible leather or cloth layers
HiDream I1 Fast
- + Very clear and colorful beads in the braids
- + Strong bokeh effect with visible torches
- + High level of ornamental detail on the armor
- − The face looks overly airbrushed and 'clean' for a battle-worn character
- − The scars look like red markings rather than realistic old wounds
- − Skin texture lacks the lifelike pores and flaws requested
Verdict: GPT Image 1 is the clear winner for its superior realism and adherence to the 'battle-worn' aesthetic, featuring lifelike skin textures and nuanced lighting. HiDream I1 Fast delivers a more stylized, almost comic-book appearance that feels less gritty and fails to render realistic scars or skin texture.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1
- + Excellent typography with legible bold sans-serif fonts.
- + High-quality, realistic food photography that looks professional.
- + Clean and authentic minimalist layout that strictly follows the prompt.
- − Minor spelling errors in placeholder text like 'descrigion'.
- − The grid structure is a bit basic.
HiDream I1 Fast
- + Includes a header section which provides more of a 'complete' document feel.
- + Better variety in the number of images within the grid.
- − Severe text rendering issues with illegible gibberish.
- − The food photography looks somewhat blurry and low-quality.
- − Contains significant artifacts in the font and graphic elements.
Verdict: GPT Image 1 is the clear winner as it produces a professional, high-fidelity menu that closely adheres to the modern minimalist aesthetic. While it has minor spelling errors, HiDream I1 Fast fails significantly in visual quality, producing distorted text and low-resolution food imagery that is unusable for a professional project.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1
- + Excellent typography with perfect spelling and fiery effects.
- + Perfect execution of the 'exploded' burger concept with clear separation of layers.
- + High photorealistic quality on the food textures like the patty and lettuce.
- − The price in the starburst is missing the '6' digit, showing only '€.99'.
HiDream I1 Fast
- + Dynamic fire at the base of the image adds to the atmosphere.
- + Included the full price '€6.99' correctly inside a starburst.
- − Failed the 'exploded' layout, showing a mostly assembled burger with floating onion rings.
- − Secondary text is cluttered and contains spelling errors like 'TIMIK'.
- − The food rendering looks more like plastic or low-quality CGI than photorealistic.
Verdict: GPT Image 1 is the superior choice for marketing as it successfully captures the 'exploded' burger aesthetic and features clean, professional-grade typography. While it missed a digit in the small price tag, HiDream I1 Fast failed the primary compositional request of an exploded burger and produced garbled text with significantly lower visual realism.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1
- + Excellent text rendering with perfect spelling across all items.
- + Accurate chalk texture that looks realistically hand-applied.
- + Clean and balanced composition that focuses on the prompt requirements.
- − The 'elegant cursive' requested for the title is more of a print-cursive hybrid.
- − Lacks the 'cozy café' background context, focusing only on the board.
HiDream I1 Fast
- + Successfully captures the 'cozy café' atmosphere with a nice depth of field.
- + Follows the 'elegant cursive' instruction for the title more closely.
- − Frequent spelling errors and character overlaps (e.g., 'Risote', 'and Her28', 'Buter').
- − Messy text layering where prices and words collide.
- − Failed to successfully complete the final menu item text.
Verdict: GPT Image 1 significantly outperforms HiDream I1 Fast by delivering perfectly legible and accurately spelled text, which is essential for a menu prompt. While HiDream I1 Fast provides a better environment and more cursive-style heading, the numerous typographic glitches and overlapping letters make it functionally inferior to the clean, professional execution of GPT Image 1.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1
- + Excellent cinematic lighting and texture on the horse and spacesuit.
- + Captures the 'in space' setting as requested.
- − Failed the negative constraint; the astronaut is riding the horse.
- − The reigns appear to be held by the horse's neck rather than the astronaut's hands.
HiDream I1 Fast
- + Natural lighting and clean rendering of the astronaut and horse equipment.
- + High clarity and sharp focus on the subjects.
- − Failed the negative constraint; the astronaut is riding the horse.
- − Failed the environmental setting; the horse is in a desert, not in space.
Verdict: Both GPT Image 1 and HiDream I1 Fast failed the specific spatial logic of the prompt, which requested the horse to be on top of the astronaut. However, GPT Image 1 is the superior image as it followed the 'in space' instruction and provided a much more cinematic and detailed aesthetic, whereas HiDream I1 Fast placed the scene in a terrestrial desert.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1
- + Excellent photorealistic texture on the capybara's fur.
- + Accurate composition with the passenger in the back seat as requested.
- + Realistic cinematic lighting and depth of field.
- − The capybara's paws look slightly more like primate hands than rodent paws.
HiDream I1 Fast
- + Bright, vibrant colors in the background bokeh.
- + The capybara's head is well-rendered.
- − Severe anatomy failure with human hands emerging from the capybara's sleeves to drive.
- − Anatomy/perspective error with the passenger appearing to sit in the front passenger seat next to the driver instead of the back.
- − The capybara's head is not properly attached to the body.
Verdict: GPT Image 1 followed the prompt requirements much more accurately, placing the passenger in the back seat and maintaining a consistent photorealistic style. HiDream I1 Fast failed significantly on anatomy by giving the capybara human hands and placing the businesswoman in the front seat, which contradicts the prompt's instructions.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent typography with perfect spelling in all three requested text areas.
- + Consistent vintage aesthetic with a dark, moody parchment texture.
- + Highly professional layout that feels like a real cinematic poster.
- − The parchment is very dark, making the thorns and webs in the border a bit hard to see.
HiDream I1 Fast
- + Strong contrast with the light parchment makes the central illustration pop.
- + Included all requested elements like bats, twisted trees, and webs.
- + Clear, legible date at the bottom.
- − Significant text rendering issues on the banner and location details.
- − The illustration style is more cartoonish and less 'polished cinematic' than requested.
- − Layout feels cluttered with repeating/overlapping text at the bottom.
Verdict: GPT Image 1 followed the instructions perfectly, delivering flawless text rendering and a sophisticated gothic aesthetic that aligns with the 'cinematic' and 'polished' requirements. HiDream I1 Fast struggled with the text on the banner and the specific event details, resulting in a cluttered and less professional appearance. GPT Image 1 is the clear winner for its superior composition and accuracy.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent text rendering with clean, bold typography.
- + Superior PBR material rendering with a soft, clay-like matte finish.
- + Perfectly follows the raised diorama base requirement with a cohesive aesthetic.
- − The flag icon is stylized as a card rather than a simple graphic icon.
HiDream I1 Fast
- + Good use of multiple sushi types including nigiri and gunkan.
- + Accurately captures the solid light blue background and isometric perspective.
- − Contains several visual artifacts, particularly a floating garbled icon next to the text.
- − The sushi design is anatomically confusing, featuring a nigiri piece with a maki-roll filling inside the rice.
- − Text rendering is slightly less refined compared to the competitor.
Verdict: GPT Image 1 is the clear winner for its high-quality rendering of materials and flawless text. While HiDream I1 Fast attempts a more complex scene, it suffers from significant logical errors (like a maki roll hidden inside a nigiri piece) and floating artifacts near the typography.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to the 'tumbling' and 'chasing' action in the prompt.
- + Included all four requested animals correctly: puppy, kitten, bunny, and fox.
- + Superior lighting with clearly defined god rays and dew sparkles.
- − The fox's front right paw is anatomically slightly blurred/unclear.
- − The kitten's mouth and teeth look slightly unnatural upon close inspection.
HiDream I1 Fast
- + Features very sharp, expressive eyes on all animals.
- + Soft, pleasing bokeh in the background and foreground.
- − Failed to include the requested baby bunny.
- − The animals are mostly sitting still, missing the 'chasing' and 'tumbling' action requested.
- − Included two kittens instead of one, deviating from the prompt.
Verdict: GPT Image 1 significantly outperformed HiDream I1 Fast by accurately including all four requested species and capturing the dynamic energy of the animals playing. While HiDream I1 Fast produced a cute image, it failed to include the bunny and ignored the specific motion-based instructions of the prompt, opting for a static pose instead.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1
- + Excellent typography and spelling throughout.
- + Strong minimalist vector aesthetic that fits a logo design.
- + Clean, centered composition with a professional layout.
- − Generated a black background instead of the requested light background.
- − The steam element is a bit overly simplified.
HiDream I1 Fast
- + Followed the color palette and background request much better.
- + Includes the requested 'subtle texture' on the background.
- + Good use of the cloche dome design with more dynamic steam.
- − Typos in both the name and the date ('CAFFÉ' instead of 'Caffè' and '17210' instead of '1720').
- − The text placement inside the cloche causes readability issues.
Verdict: GPT Image 1 followed the technical vector style and text requirements perfectly, though it failed to provide the requested light background. HiDream I1 Fast captured the requested color scheme and texture much better, but failed on the basic execution of the text by adding an extra digit to the date and misspelling the restaurant name. Because logical text adherence is critical for a logo, GPT Image 1 is the superior output despite the background color error.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1
- + Excellent text legibility and correct spelling for almost all labels.
- + Clean, professional flat-vector aesthetic that perfectly matches the 'modern infographic' request.
- + Logical layout that follows a sequence from launch to landing.
- − Includes a spelling error with 'EARLLUNAR' in the bottom footer.
- − Icon placement is a bit scattered, making the flow of the six steps slightly confusing to follow.
HiDream I1 Fast
- + Strong title presence with 'APOLLO 11' clearly displayed.
- + Good use of the requested NASA-inspired color palette.
- − Significant text rendering issues with numerous gibberish characters and misspellings.
- − Icons are messy and do not clearly represent the specific mission phases requested.
- − Composition feels cluttered and lacks the 'crisp lines' requested in the prompt.
Verdict: GPT Image 1 significantly outperforms HiDream I1 Fast by delivering a professional, clean infographic style that follows the prompt's request for crisp lines and specific phases. While GPT Image 1 has one spelling error, HiDream I1 Fast suffers from widespread text corruption and poor iconography that fails to clearly communicate the mission steps.
Explore each model
Distilled version of HiDream AI's 17B parameter text-to-image model