OpenAI's previous generation image model with higher quality than DALL-E 2 and support for larger resolutions
Settled by community votes across 12 shared challenges, with an AI judge weighing in on each.
DALL-E 3
#40 of 62 in Text-to-Image
HiDream I1 Full
#60 of 62 in Text-to-Image
Where the votes landed
DALL-E 3
0%
win rate
Ties
0%
HiDream I1 Full
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
DALL-E 3
- + High resolution and artistic lighting effects
- + Rich wood and paper textures
- + Creative interpretation of the sphere containing a landscape
- − Failed the spatial part of the prompt by placing the book inside the cube
- − Failed to place the book on TOP of the cube
- − Cube design includes a wooden frame not requested
HiDream I1 Full
- + Perfect adherence to all spatial instructions
- + Accurate rendering of light coming from the left window
- + Correct placement of the red book on top of the glass cube
- − The blue sphere appears to be floating unnaturally without support
- − The lighting on the cube's base has a slightly messy reflection
Verdict: HiDream I1 Full followed every instruction in the prompt perfectly, including the specific spatial relationships between the objects and requested light source. DALL-E 3 produced a high-quality image but failed the core challenge by placing the red book inside the cube instead of on top of it.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
DALL-E 3
- + Excellent composition using foreground framing and reflection to create a cinematic feel
- + Strong interpretation of 'imperfect framing' with a candid, fly-on-the-wall perspective
- + Accurate depiction of a red bicycle and wet pavement reflections
- − Anatomical issues with the man's feet and the connection of his arm to his body
- − High contrast and sharp lighting lean more toward 'stylized' than 'no stylization'
HiDream I1 Full
- + Highly realistic skin textures and clothing details that avoid an AI-generated look
- + Perfectly captures the lighting and atmosphere of a rainy day in a natural way
- + Accurate representation of a 50mm shallow depth of field
- − The man appears to be sitting on or straddling the bicycle rather than 'repairing' it
- − Lack of 'motion blur from passing cars' as the background vehicles appear static
Verdict: Both models followed the prompt well, but HiDream I1 Full won on photographic realism and natural textures, whereas DALL-E 3 struggled with human anatomy and felt more like a digital painting. While DALL-E 3 had a more creative layout, HiDream I1 Full adhered better to the 'no stylization' and 'natural skin texture' requirements.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
DALL-E 3
- + Excellent lighting that realistically reflects off the metal engravings.
- + Highly detailed facial skin texture and lifelike eyes.
- + Strong professional shallow depth of field effect.
- − Missed the request for hair braided with small beads.
- − The scars look like surface-level marks rather than deep, battle-worn history.
HiDream I1 Full
- + Successfully included braided hair with visible beads as requested.
- + Detailed metal armor with appropriate dirt and blood splatter for a 'battle-worn' look.
- + Stronger adherence to the specific character accessories mentioned in the prompt.
- − The eyes look slightly artificial and glass-like compared to the other model.
- − The lighting is flat and lacks the dramatic warm glow of the torchlight requested.
Verdict: While DALL-E 3 produces a more cinematic and technically superior photographic image with better lighting, HiDream I1 Full followed the prompt's structural details more accurately by including the braided hair and beads. HiDream I1 Full is the winner for comprehensive prompt adherence despite the less realistic lighting.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
DALL-E 3
- + Excellent grid composition that effectively mixes photography and layout blocks
- + Wide variety of food imagery that looks professional and appetizing
- + Superior use of color accents and vibrant visual balance
- − Text is largely gibberish or decorative rather than readable
- − Layout is more of a mood board style than a functional single-page menu
HiDream I1 Full
- + Text is highly readable and uses a bold sans-serif font as requested
- + Clear hierarchical sections for pizza and mains
- + Functional and clean professional layout
- − Repetitive food imagery with four very similar pizzas instead of varied food types
- − Less visually creative and lacks the 'vibrant accents' requested in the prompt
- − Slightly dated design aesthetic compared to Model A
Verdict: DALL-E 3 (Image A) captures the 'modern minimalist' aesthetic with much better composition and professional food photography integration, though the text is unreadable. HiDream I1 Full (Image B) is much more functional as a menu and follows the text instructions better, but it fails on variety by showing four nearly identical pizzas and lacks the artistic vibrancy of the first image. DALL-E 3 is preferred for its superior design and creative interpretation of the prompt.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
DALL-E 3
- + Excellent authentic chalk texture with realistic smudging and dust
- + Visual aesthetic perfectly matches the 'cozy cafe' description
- + Captures the cursive style requested in the header
- − Significant spelling errors throughout the menu items such as 'Occtus' and 'Grililled'
- − Prices shown do not match the specific numbers requested in the prompt
- − Text becomes cluttered and illegible toward the bottom
HiDream I1 Full
- + Clean layout with higher legibility for the header text
- + Correctly identifies the date and year from the prompt
- + Good use of illustrations such as the coffee cup
- − Text appears as a digital font rather than authentic hand-drawn chalk
- − Completely ignored the specific menu items like Truffle Mushroom Risotto
- − Nonsensical text and repetitive pricing throughout the board
Verdict: DALL-E 3 captures the soul of the prompt with beautiful, authentic chalk textures and a cozy atmosphere, even though it struggles significantly with spelling and price accuracy. HiDream I1 Full produces much cleaner text but fails specifically because the handwriting looks like a synthetic digital font, and it ignores the requested menu items entirely in favor of gibberish. DALL-E 3 is the winner for adhering to the requested artistic style and attempting the specific menu items.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
DALL-E 3
- + Excellent cinematic lighting and atmosphere
- + Beautiful and detailed nebula background
- + Creative use of glowing clouds to ground the surreal concept
- − Failed the specific spatial instruction (horse is on bottom, not top)
- − The horse's legs have slightly anatomical artifacts at the hooves
HiDream I1 Full
- + Sharp image quality and clear details on the astronaut suit
- + Good sense of motion and dynamic posing
- + Clean composition with a nice planetary curve
- − Failed the specific spatial instruction (horse is on bottom, not top)
- − The horse's harness merges into the astronaut's leg
- − Less 'surreal' than the other version, appearing more like a standard montage
Verdict: Both DALL-E 3 and HiDream I1 Full failed the specific spatial logic prompt to put the 'horse on top' of the astronaut, both reverting to the standard image of an astronaut riding a horse. DALL-E 3 is the preferred choice as it better captured the 'surreal' and 'cinematic' style requested, whereas HiDream I1 Full felt more like a generic composite with some minor limb and harness clipping.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
DALL-E 3
- + Excellent interior detail with realistic dashboard and lighting
- + Creative inclusion of a 'CAPYBARA' sign in the background
- + High visual quality and atmospheric lighting
- − Failed to include the human passenger in the back seat
- − Clipped composition with the front passenger seat taking up much of the foreground
HiDream I1 Full
- + Successfully followed all instructions including the human passenger
- + Excellent pose and expression for both the capybara and the woman
- + Paws are correctly positioned on the steering wheel
- − The transition between the capybara's fur and the jacket collar is slightly anatomically awkward
- − The lighting on the woman's hair is a bit inconsistent with the dark taxi interior
Verdict: While DALL-E 3 produced a more visually stunning and detailed interior, it completely missed the requirement for a passenger. HiDream I1 Full successfully followed the entire prompt, capturing the humorous juxtaposition of the bored businesswoman and the capybara driver perfectly.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
DALL-E 3
- + Expertly captures the vintage gothic aesthetic with dark, moody lighting and intricate 3D-like textures.
- + Includes all requested elements like twisted trees, webs, and a central glowing jack-o-lantern in a cohesive composition.
- + The overall artistic quality is high, resembling a real physical prop.
- − Text rendering is largely illegible gibberish, failing to provide the specific event details requested.
- − Does not include the requested scroll banner.
HiDream I1 Full
- + Renders the primary scroll banner text 'You are invited to a night of frights' perfectly.
- + Follows the layout instructions for a poster with distinct sections for title, banner, and details.
- − The 'event details' text at the bottom is factually incorrect and full of nonsense words like 'Yootmber'.
- − The visual style is more like a flat clip-art illustration than a 'polished, cinematic' vintage gothic poster.
- − Failed to include the specific date, time, and location requested in the prompt.
Verdict: DALL-E 3 produces a far superior visual image that perfectly matches the requested 'vintage gothic' and 'cinematic' mood, though it fails significantly on text legibility. HiDream I1 Full manages to render one specific line of text correctly but fails on all other text data and lacks the artistic depth and atmospheric quality of the first model. DALL-E 3 is preferred for its adherence to the stylistic and compositional prompts, despite the text issues common in some generators.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
DALL-E 3
- + Excellent 3D rendering with soft, toy-like textures and PBR materials.
- + Highly creative interpretation of the sushi as a singular cube-like structure.
- + Great use of the diorama base to incorporate the requested text and flag.
- − Failed to place the text at the top-center as requested.
- − Omitted the specific word 'SUSHI' from the text elements.
- − Rice geometry looks more like pearls than realistic grains.
HiDream I1 Full
- + Followed text placement instructions perfectly with 'JAPAN' and 'SUSHI' at top-center.
- + Accurate isometric perspective and clean display on the diorama base.
- + Offers a more recognizable and diverse variety of sushi types.
- − Completely missed the 'small flag icon' requirement.
- − Lighting is a bit flatter compared to the soft shadows in Model A.
- − The background shows some subtle painting-like artifacts/textures instead of being a solid flat color.
Verdict: HiDream I1 Full is the winner because it adhered much more strictly to the layout instructions, specifically the text content and top-center placement. While DALL-E 3 produced a more visually pleasing 3D render with superior lighting and materials, it failed to include all the requested text and placed the available text on the base rather than at the top of the frame.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
DALL-E 3
- + Excellent depiction of god rays and sunrise lighting consistent with the prompt
- + Accurately included all four requested animals: puppy, kitten, bunny, and fox
- + Strong emotive composition that captures the 'joyful wholesome vibe'
- − The butterflies have weirdly distorted hybrid wings with furry creature bodies
- − Leans heavily into a stylized 3D animation look rather than the requested hyper-photorealistic style
HiDream I1 Full
- + Better realism in the fur textures and animal features
- + Clean lighting that feels organic to the meadow scene
- − Failed to include the baby bunny, providing two kittens instead
- − Composition is more static and less 'playful' or 'tumbling' than requested
- − Missing the specific 'god rays' lighting visual mentioned in the prompt
Verdict: DALL-E 3 followed the complex prompt requirements much better, successfully including all four specific animal types and the atmospheric lighting effects. While HiDream I1 Full has a slightly more realistic photographic texture, it ultimately failed to generate the bunny and ignored the specific 'god rays' instruction, making DALL-E 3 the more accurate choice despite its more 'Pixar-like' aesthetic.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
DALL-E 3
- + Excellent use of texture and vintage stippling effects
- + Rich, complex composition with professional vector emblem aesthetics
- + Clear rendered date and steam elements
- − Failed to include the primary requested text 'Caffè Florian', replacing it with 'COFFEE HOUSE'
HiDream I1 Full
- + Successfully adheres to the minimalist vector style requested
- + Clean, balanced layout with a clear focal point
- + Accurate rendering of the banner and date
- − Completely missing the primary brand name 'Caffè Florian'
- − The cloche dome design is slightly awkward with the coffee cup integrated inside it
Verdict: Both models failed to include the specific name 'Caffè Florian' from the prompt. DALL-E 3 produced a much more visually compelling and authentic vintage emblem with superior texture and detail, whereas HiDream I1 Full provided a very basic minimalist icon. DALL-E 3 is the preferred choice for its artistic quality and adherence to the 'vintage' and 'cloche' aspects of the prompt, despite the text substitution.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
DALL-E 3
- + Features a sophisticated, professional layout with a consistent color palette.
- + Includes high-quality textures and detailed lunar illustrations.
- + Successfully captures the vintage-modern NASA aesthetic requested.
- − Fails to follow the specific 6-step chronological structure requested.
- − Text is largely illegible gibberish.
- − The rocket designs resemble space shuttles rather than the Saturn V.
HiDream I1 Full
- + Adheres much better to the requested flat-vector style.
- + Attempts to follow the specific step-by-step instructions for iconography.
- + Better text rendering, despite some spelling errors.
- − Includes prompt technical terms (like '+ ICON' and '+ UNG') directly in the final image labels.
- − The composition is unbalanced and cluttered.
- − Rockets look more like cartoons than technical Saturn V icons.
Verdict: Both models struggled with the specific technical instructions. DALL-E 3 produced a far more beautiful and professional-looking poster, but completely ignored the requested 6-step structure. HiDream I1 followed the prompt's logical structure much more closely, but the inclusion of prompt text in the final labels and the less-polished art style make it less effective as a final piece. DALL-E 3 is the winner for visual quality and style adherence, despite the content mismatch.
Explore each model
HiDream AI's 17B parameter text-to-image model using sparse diffusion transformer with mixture of experts, achieving state-of-the-art image generation quality with strong prompt following