OpenAI's state-of-the-art image generation model with better instruction following and adherence to prompts
Settled by community votes across 8 shared challenges, with an AI judge weighing in on each.
GPT Image 1.5
#7 of 62 in Text-to-Image
Reve Image 1.0
#44 of 62 in Text-to-Image
Where the votes landed
GPT Image 1.5
72.7%
win rate
Ties
0.0%
Reve Image 1.0
27.3%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1.5
- + Excellent adherence to the glass cube geometry and material properties.
- + Very high detail on the red book's texture and the blue sphere's reflection.
- + Succesfully captures the plant being visible 'through' the glass as requested.
- − The blue sphere is quite large relative to the cube, rather than 'small' as prompted.
Reve Image 1.0
- + Accurately renders a 'small' blue sphere relative to the cube size.
- + Good composition with a clear soft window light effect casting a shadow on the table.
- − The blue sphere appears to be floating mid-air inside the cube without support.
- − The perspective of the cube is slightly warped, and the glass lacks the thickness and caustic realism seen in Model A.
Verdict: GPT Image 1.5 is the superior image due to its exceptional material realism and detail, particularly in how the glass cube interacts with its environment and the red book. While Reve Image 1.0 followed the size constraint for the sphere better, the sphere's floating state and the less convincing glass rendering make it the weaker choice.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1.5
- + Exceptional textural detail on the skin, leather straps, and engraved armor.
- + Strong adherence to the 'battle-worn' aesthetic with realistic dirt and scars.
- + Powerful composition with high-quality bokeh and cinematic lighting.
- − The hair beads are somewhat merged with the hair strands in a less distinct way.
Reve Image 1.0
- + Excellent representation of the beaded braids requested in the prompt.
- + Ornate engraving on the plate armor is clearly visible and well-executed.
- + Good use of warm torchlight and shallow depth of field.
- − The character looks very young and lacks the 'battle-worn' texture requested.
- − The skin texture is too smooth and the dirt looks like individual isolated dots rather than realistic grime.
Verdict: GPT Image 1.5 is the superior image due to its incredible attention to detail, specifically the lifelike eyes and the realistic 'battle-worn' texture of the skin and armor which perfectly captures the prompt's mood. While Reve Image 1.0 followed the specific request for hair beads more literally, it failed to deliver on the gritty, weathered aesthetic required for a battle-worn paladin.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1.5
- + Natural integration of the hair with the existing sideburns and beard
- + High level of preservation for facial features and clothing
- + Realistic texture and density that matches the lighting of the scene
- − The hair slightly changes the upper forehead shape and skin texture
Reve Image 1.0
- + Excellent preservation of the background and original clothing
- + Follows the prompt for a 'full' head of hair
- − The hair volume appears unnaturally high and lacks a realistic silhouette
- − The blending at the temples and sideburns is messy and less coherent than the other model
- − The image overall appears slightly softened/less sharp than the source
Verdict: GPT Image 1.5 provides a much more convincing and realistic edit, blending the new hair seamlessly into the subject's existing beard and sideburns while maintaining the photo's original sharp quality. Reve Image 1.0 creates an unnaturally large, 'wig-like' volume of hair that doesn't integrate well with the subject's facial structure or the surrounding environment.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1.5
- + Excellent PBR materials and textures, especially on the wood and ceramics.
- + Clear, bold typography that is perfectly integrated into the scene.
- + High level of detail including realistic food textures and atmospheric lighting.
- − The scene is a bit more crowded than the 'minimal garnish' request specified.
- − The chopsticks are not resting on a chopstick rest correctly (floating/clipping).
Reve Image 1.0
- + Follows the 'miniature 3D cartoon' style very accurately with softened shapes.
- + Matches the 'minimal garnish' and 'solid background' instructions perfectly.
- + Clean, simple composition that emphasizes the diorama base.
- − The text is placed to the side rather than 'top-center'.
- − The materials appear less 'realistic PBR' and more like simple plastic/clay.
- − Lack of complex textures makes the rice look like generic white blocks.
Verdict: GPT Image 1.5 produced a much higher quality render with sophisticated materials and professional typography, though it leaned more toward realism than a cartoon style. Reve Image 1.0 captured the 'cartoon' aesthetic well but failed on the specific layout instructions for the text and lacked the textural depth requested in the 'realistic PBR' prompt.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1.5
- + Excellent caricature style with exaggerated features that still look like the person.
- + Superb prompt adherence by including multiple dogs, hockey players/gear, and a detailed news set.
- + Very clean text rendering and high-quality artistic detail.
- − The small dog's helmet and stick are slightly messy in terms of fine detail.
Reve Image 1.0
- + Strong resemblance to the original person's facial features.
- + Creative elements like the dog wearing headphones and a tie.
- + Good clarity and clean composition.
- − The hockey stick is just floating/propped up awkwardly without much context.
- − Less 'exaggerated' as a caricature compared to Model A.
Verdict: GPT Image 1.5 followed the prompt much more effectively, creating a busy, humorous, and highly detailed caricature that perfectly blended the news anchor role with dogs and hockey themes. Reve Image 1.0 felt a bit more like a standard illustration with fewer elements, and the facial proportions were less 'exaggerated' in the traditional caricature sense.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1.5
- + Excellent preservation of the source image's composition and poses.
- + Beautiful, warm color palette that fits the 'nostalgic' and 'dreamy' prompt keywords.
- + Maintains the distinct facial expressions of the original meme perfectly in a new style.
- − The art style leans more towards modern shoujo anime or digital watercolor than the specific hand-drawn Ghibli aesthetic.
- − Overuse of soft glow/haze can obscure some of the fine linework.
Reve Image 1.0
- + Very accurate capture of the Studio Ghibli '90s cel-shaded look and line weights.
- + Excellent hand-painted background textures that feel grounded in the requested style.
- + Effective use of the 'out of focus' effect for the foreground character to match the source photo's depth of field.
- − The man's facial expression is a bit more neutral/blank compared to the source's 'pucker' expression.
- − The woman in red's face is a bit generic and lacks the personality of the source.
Verdict: Both models did an excellent job translating the 'Distracted Boyfriend' meme into an illustration. GPT Image 1.5 followed the mood and lighting instructions perfectly, creating a beautiful image, but Reve Image 1.0 was much more successful at specifically capturing the Studio Ghibli art style and 'painted' background aesthetic requested.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1.5
- + Excellent preservation of the original person's face and clothing.
- + Subtle and realistic outward-blowing hair effect.
- + The leaves are integrated into the scene's lighting well.
- − The leaf density is a bit high, which can feel slightly cluttered.
- − The leash handle area has some minor artifacting where it meets the hand.
Reve Image 1.0
- + Strong sense of dynamic motion with larger, blurred foreground leaves.
- + Very energetic hair-blowing effect that matches the wind direction of the leaves.
- + Good preservation of the dog and background elements.
- − The larger leaves in the foreground are a bit blurry and distract from the subject.
- − The hair edit looks slightly more 'painted on' compared to the original texture.
Verdict: Both models did an excellent job of following the edit instructions while preserving the source image. GPT Image 1.5 is preferred for its superior preservation of the subjects' fine details and more natural integration of the leaves, whereas Reve Image 1.0 has slightly more aggressive motion blur that obscures the foreground too much.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1.5
- + Excellent text rendering with correct spelling for all steps and names.
- + All six requested stages are clearly represented with logical, high-quality icons.
- + Consistent flat-vector style with a professional NASA-inspired palette.
- − The layout is a bit cramped at the top, cutting off the main title.
- − The 'Launch' section includes a globe icon that might be better suited for orbit.
Reve Image 1.0
- + Clean, minimalist layout with good use of negative space.
- + Accurately captures the requested flat-vector aesthetic.
- − Several spelling errors in icons (LUNNR, DESSENT, APQUO).
- − Missing the 'Translunar' step entirely, skipping from step 2 to 4.
- − Numbering is inconsistent and non-sequential.
Verdict: GPT Image 1.5 followed the prompt instructions perfectly, including all six specific mission steps with correct spelling and highly detailed vector icons. Reve Image 1.0 struggled with text accuracy and omitted the requested 'Translunar' stage, resulting in an incomplete and confusing infographic.
Explore each model
Reve AI's text-to-image generation model with strong aesthetic quality, accurate text rendering, and detailed instruction following capabilities