OpenAI's previous image generation model that accepts both text and image inputs and produces image outputs
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
GPT Image 1
#29 of 62 in Text-to-Image
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Where the votes landed
GPT Image 1
0.0%
win rate
Ties
0.0%
GPT Image 1 Mini
100.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1
- + Excellent handling of glass thickness and refractive physics
- + Stronger, more detailed textures on the book and sphere
- + Higher overall resolution and sharpness
- − The sphere appears slightly disconnected at its base from the floor of the cube
GPT Image 1 Mini
- + Successfully includes all requested elements in a balanced composition
- + Good adherence to lighting and placement instructions
- − The glass box lacks structural integrity, with edges looking like thin wires rather than glass panes
- − The image is noticeably softer and lower in resolution compared to Model A
Verdict: Both models followed the complex spatial instructions perfectly. GPT Image 1 is the winner due to superior visual quality, showing much more realistic glass properties and fine textures on the red book, whereas GPT Image 1 Mini's output looks significantly blurrier and less detailed.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the car's front-end design and trim
- + Accurate hair style for the man
- + Dynamic motion blur on the wheels and road adds realism
- − The man's clothing is significantly changed from the source image
- − The man appears slightly small/distant within the cabin
GPT Image 1 Mini
- + Perfectly preserves the man's plaid coat and black scarf from the source
- + High facial similarity to the source subject
- + Beautiful lighting and scenic composition
- − The car's front hood and headlight design is altered from the source Rolls-Royce
- − Perspective issues between the man and the car's interior geometry
Verdict: GPT Image 1 (Model A) does a much better job of maintaining the specific look and details of the Rolls-Royce from the source image, whereas GPT Image 1 Mini (Model B) modifies the car's design. However, Model B is much more successful at preserving the man's distinct outfit, specifically the plaid pattern of the coat and the texture of the scarf, making it feel more like a composite of the two source images. Model A belongs to a higher technical class for its realistic motion blur and car accuracy, while Model B wins on character consistency.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1
- + Excellent depiction of rainy conditions with visible droplets in the air and on the seat
- + Strong lighting with realistic bokeh from car headlights
- + Deep, natural-looking shadows and skin texture
- − The hand interacting with the wheel is anatomically muddled with the bicycle parts
- − Background car lacks the requested motion blur, appearing more static despite the soft focus
GPT Image 1 Mini
- + Superb facial details and natural skin texture that looks very photographic
- + Better hand/finger definition than the competitor
- + Successful use of shallow depth of field to draw focus to the subject
- − The bicycle frame geometry is slightly warped near the rear rack
- − The background car lacks any 'motion blur' requested in the prompt, appearing quite sharp and static
Verdict: Both models captured the requested mood and aesthetic perfectly, opting for a cinematic, high-quality photograph style. GPT Image 1 (Model A) did a slightly better job illustrating the 'light rain' and environmental atmosphere through droplets and lighting, while GPT Image 1 Mini (Model B) excelled in the sharp detail of the subject's face. Model A is the slight winner for better integrating the subject with the wet environment.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI judge analysis unavailable for this challenge.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1
- + Excellent high-resolution food photography that looks appetizing.
- + Includes functional text elements like pricing and item descriptions.
- + Effective use of bold sans-serif fonts and primary color accents.
- − Text contains frequent misspellings and gibberish (e.g., 'Apperoiation descrigion').
- − The grid layout is slightly uneven due to the large scale of the images.
GPT Image 1 Mini
- + Strict adherence to the 'grid' layout requested in the prompt.
- + Perfectly clean and minimalist aesthetic with no spelling errors in headers.
- + The composition feels more like a complete menu template with a clear header.
- − Lacks individual item names, descriptions, or prices, leaving large empty spaces.
- − Food photos are smaller and have less detail than the other model.
Verdict: GPT Image 1 provides much better visual quality for the food itself and feels like a more 'functional' menu despite the gibberish text. GPT Image 1 Mini creates a superior structural layout that perfectly matches the minimalist grid request, but it fails to include any specific menu items or pricing. GPT Image 1 is preferred because the high-quality photography and complete textual structure are more useful for a design concept.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1
- + Excellent photorealistic texture on the patty and bun.
- + Vibrant lighting with clear sauce droplets adding to the motion.
- + Text rendering for titles is clean and well-integrated.
- − Failed the price accuracy, rendering '€.99' instead of '€6.99'.
- − The starburst shape is slightly generic compared to Model B.
GPT Image 1 Mini
- + Perfect text accuracy, including the specific price '€6.99'.
- + More dynamic starburst effect for the price tag.
- + Good vertical composition and floating effect.
- − The burger ingredients look slightly less detailed and more 'plastic' than Model A.
- − The lettuce and tomato seem a bit small in proportion to the bun.
Verdict: GPT Image 1 has superior visual textures and lighting, making the food look more appetizing, but it fails on the specific text element for the price. GPT Image 1 Mini captured all text requirements perfectly, including the €6.99 price, though its food rendering is slightly less realistic. GPT Image 1 Mini is the winner for total prompt adherence.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1
- + Excellent chalk texture that genuinely mimics physical chalk on a board.
- + Strong adherence to the specific text and numbers in the prompt.
- + Atmospheric lighting that enhances the cafe feel.
- − The title is in all-caps rather than the 'elegant cursive' requested.
- − Some text is slightly cropped at the bottom edge.
GPT Image 1 Mini
- + Perfect text accuracy including all specific menu items and prices.
- + Clean composition with the full chalkboard frame visible.
- + Natural variation in letter sizing and baseline.
- − Failed the negative constraint regarding 'no printed or digital fonts' as the bottom text looks like a standard digital font.
- − Failed the request for 'elegant cursive' for the title.
Verdict: Both models struggled with the specific request for 'elegant cursive' handwriting for the title, both opting for all-caps sans-serif styles instead. GPT Image 1 is preferred because its chalk texture is significantly more realistic and it avoids the computer-generated font appearance found at the bottom of the GPT Image 1 Mini output.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
GPT Image 1
- + Successfully replicates the complex leg-crossing pose from Image 1.
- + Maintains the checkered scarf and sunglasses from the character reference.
- + Captures the tilted head angle and perspective of the original pose.
- − The facial features are significantly distorted and unrecognizable.
- − The hands are poorly rendered with missing/merged fingers.
- − The scarf physics and integration with the torso are messy.
GPT Image 1 Mini
- + Excellent facial likeness and preservation of character details from Image 2.
- + High visual quality with clean textures and consistent lighting.
- + Good hand anatomy and clear representation of the clothing.
- − Fails to replicate the specific leg-crossing pose, opting for a simpler knee-up position.
- − The scarf ends abruptly behind the arm rather than hanging naturally.
Verdict: GPT Image 1 attempted the difficult task of replicating the exact leg-crossing pose but failed significantly in terms of facial likeness and anatomical detail. GPT Image 1 Mini produced a much cleaner, more recognizable version of the character with better overall quality, though it compromised on the exactness of the pose. GPT Image 1 Mini is preferred because the character's face is actually clear and looks like the source, whereas GPT Image 1 is plagued by artifacts.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1
- + Excellent color contrast with the brown horse against the dark space background.
- + Highly detailed texture on the space suit and the horse's coat.
- + Strong cinematic composition with the planet curvature at the bottom.
- − Failed the negative constraint to have the horse on top of the astronaut.
GPT Image 1 Mini
- + Atmospheric lighting that blends the horse and astronaut into the scene well.
- + Clean rendering of the lunar surface in the background.
- + Detailed harness and rein work.
- − Failed the negative constraint to have the horse on top of the astronaut.
- − Lower color vibrancy compared to the other image.
Verdict: Both GPT Image 1 and GPT Image 1 Mini failed the specific prompt instruction to place the horse on top of the astronaut, instead opting for the more conventional 'astronaut riding a horse' interpretation. GPT Image 1 is the preferred winner due to significantly better visual quality, richer colors, and more intricate details in the space suit and horse's mane.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1
- + Maintains the facial features and unique skin patterns of the original person very accurately.
- + Successfully replicates the specific plaid pattern and coat style from the reference image.
- + Matches the lighting and depth of field of the beach background well.
- − The transition between the neck and the clothing is slightly blurred.
- − The watch style is different from the one seen in Image 2.
GPT Image 1 Mini
- + Strong full-body composition that accurately incorporates the jeans and belt from Image 2.
- + Excellent integration of the vitiligo patterns on the hands, maintaining consistency with the person's identity.
- + High visual clarity and sharp textures on the clothing.
- − Noticeably changes the facial structure and head shape, losing the likeness of the original person.
- − The scarf pattern is slightly less accurate to the reference than Model A's.
Verdict: GPT Image 1 (Model A) is the winner because it prioritizes the core instruction of keeping the person's exact face and hair unchanged while still providing a high-quality clothing transfer. While GPT Image 1 Mini (Model B) does a better job of capturing the full outfit (jeans and shoes) and maintaining skin details on the hands, it fails the primary constraint by significantly altering the subject's facial appearance and head shape.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1
- + Excellent fur texture and realistic capybara anatomy
- + Includes clear 'TAXI' text on the driver's cap
- + Both paws are distinctly visible on the steering wheel as requested
- − The passenger's facial features are slightly distorted and blurry
- − The lighting on the capybara's head is slightly harsh for a night scene
GPT Image 1 Mini
- + Atmospheric cinematic lighting that feels more natural for a night scene
- + The human passenger is rendered with much higher detail and a clearer expression
- + Includes a classic checkerboard pattern on the taxi hat
- − Only one paw is clearly visible on the steering wheel, missing the 'both front paws' prompt instruction
- − The hat lacks the 'TAXI' text seen in the other version
Verdict: Both models followed the prompt well, but GPT Image 1 followed the specific instruction to have both paws on the wheel more accurately. While GPT Image 1 Mini produced a more realistically rendered passenger and better integration of lighting, GPT Image 1 provides a sharper main subject with better prompt adherence regarding the capybara's pose.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent typography with a classic gothic aesthetic
- + Clean and readable text rendering
- + Great use of depth with the foggy background and moon
- − Information error: combined the 'Time' and 'Location' lines into one confusing entry
- − Jack-o-lantern takes up less of the frame than requested for a 'central' focus
GPT Image 1 Mini
- + Perfect adherence to all text fields including the 7pm time
- + Superior layout with the scroll banner correctly placed under the image
- + Atmospheric use of vintage-style grain and texture
- − The font for the bottom details is a bit more modern compared to the gothic header
- − Very dark overall lighting makes some of the border details hard to see
Verdict: GPT Image 1 Mini emerges as the winner because it successfully included all requested information, whereas GPT Image 1 made a mistake with the event details by omitting '7pm' and mislabeling the location as 'Time'. Additionally, the composition in GPT Image 1 Mini feels more balanced and visually like a professional invitation.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the original facial features and expression.
- + Seamlessly integrates the hair with the original lighting and background texture.
- − The hairline is a bit high and stiff, appearing slightly unnatural.
GPT Image 1 Mini
- + Natural, thick hair texture with realistic volume and messy strands.
- + Provides a more modern, flattering hairline compared to the other model.
- − Significantly alters the person's facial structure, making them look like a different individual.
- − Loss of the original rough skin texture and specific eye/nose shape from the source image.
Verdict: GPT Image 1 is the superior edit because it successfully fulfills the prompt while preserving the identity of the person in the source image. GPT Image 1 Mini creates a more aesthetically pleasing head of hair, but it fails as an edit by fundamentally changing the subject's face and removing original character details.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent typography layout and centering
- + Sophisticated clay-like PBR materials and textures
- + Highly accurate 45° isometric perspective
- − The wasabi and ginger are slightly large relative to the sushi pieces
GPT Image 1 Mini
- + Beautiful color variety in the sushi selection
- + Clean rendering and soft lighting
- + Good adherence to the diorama base requirement
- − The text 'JAPAN' is not perfectly centered with the rest of the elements
- − Small flag icon is placed to the side rather than centered as requested
Verdict: Both models followed the prompt very closely, producing high-quality isometric dioramas. GPT Image 1 is slightly superior because of its perfect alignment of the text and flag, whereas GPT Image 1 Mini's text layout feels slightly off-balance.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1
- + Excellent caricature style with watercolor-like textures and shading.
- + Includes detailed background storytelling like the dog hockey player on the monitor.
- + Maintains character identity through specific clothing details like the double-pocket denim shirt.
- − The facial exaggeration is slightly more aggressive/creepy compared to Model B.
GPT Image 1 Mini
- + Cleaner line work with a classic colored pencil caricature feel.
- + Clearer representation of the news anchor role through a microphone and large graphic.
- + Great preservation of the golden retriever's features often associated with 'love for dogs'.
- − The arm/hand holding the dog appears anatomically awkward and detached.
- − A bit simpler in terms of background detail compared to the other model.
Verdict: Both models did an excellent job translating the source photo into a caricature while hitting all prompt requirements (news anchor, dogs, hockey). GPT Image 1 showcases more artistic depth with watercolor textures and a clever monitor graphic, while GPT Image 1 Mini feels more like a traditional boardwalk caricature with bold lines and clear icons. GPT Image 1 is slightly preferred for its more sophisticated composition and better integration of the hockey elements.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1
- + Excellent depiction of god rays and warm golden backlight.
- + High-energy composition that successfully captures the 'tumbling' and 'chasing' aspect of the prompt.
- + Superior rendering of whiskers and fine fur textures on the cat and bunny.
- − The fox's front right paw is awkwardly shaped and lacks clear definition.
- − One butterfly near the center is quite blurry and indistinct.
GPT Image 1 Mini
- + Clean, well-defined subjects with distinct silhouettes.
- + Natural light interaction with the meadow, showing realistic dew sparkles.
- + Great anatomical consistency for all four animals while maintaining the cute aesthetic.
- − The composition feels slightly more staged and less dynamic than Model A.
- − The kitten's tail looks slightly disconnected from its body due to being obscured by fur.
Verdict: Both models followed the prompt exceptionally well, capturing all four requested animals and the specific lighting conditions. GPT Image 1 (Model A) is chosen as the winner for its more dynamic 'tumbling' composition and more pronounced atmospheric effects like the god rays, which better fit the '8K masterpiece' requirement. GPT Image 1 Mini (Model B) is also high quality but feels a bit more like a static portrait in comparison.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1
- + Successfully captures the specific Studio Ghibli character aesthetic with simplified features and large expressive eyes.
- + Includes hand-painted textures and a soft watercolor wash consistent with the requested style.
- + Preserves the composition and character dynamics of the original meme perfectly.
- − The blurry background is a bit too abstract compared to Ghibli's detailed scenery.
- − The man's shirt pattern is simplified into stripes rather than the original plaid.
GPT Image 1 Mini
- + Excellent retention of the specific plaid pattern on the man's shirt.
- + Preserves the likeness and facial structures of the original subjects more accurately than Model A.
- + Nice colored-pencil or crayon-like texture across the illustration.
- − The style leans more towards generic Western children's book illustration than Studio Ghibli's specific anime style.
- − The lighting feels a bit flat compared to the 'dreamy' atmospheres often found in Ghibli films.
- − The woman in the foreground has a face that looks a bit more realistic/uncanny than the stylized characters in the back.
Verdict: GPT Image 1 captured the requested 'Studio Ghibli' aesthetic much more effectively, using the specific line work and facial styles characteristic of the studio. GPT Image 1 Mini produced a high-quality illustration that preserved clothing details like the plaid shirt better, but it lacks the iconic anime feel requested in the prompt. GPT Image 1 is the winner for better style adherence.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1
- + Successfully added energetic motion to the hair with realistic flyaway strands.
- + Preserved the original appearance and identity of the woman very well.
- + Added many small, dynamic leaves that blend well with the scene's lighting.
- − The dog's tail has a slightly unnatural, choppy texture in the new pose.
- − The leash handle morphs awkwardly into the woman's hand compared to the original.
GPT Image 1 Mini
- + Added large, clearly visible leaves that definitely meet the prompt requirements.
- + Preserved the composition and lighting of the background elements well.
- + Successful hair motion that looks natural and windswept.
- − Noticeably changed the woman's facial features, losing the likeness of the source image.
- − The hand holding the leash has become distorted with merged fingers.
- − The dog's face has changed significantly, looking less like the original golden retriever.
Verdict: GPT Image 1 is the superior edit because it successfully adds the requested motion and leaves while maintaining excellent source preservation of the woman and the dog. In contrast, GPT Image 1 Mini fails the source preservation aspect by significantly altering the facial features of both the human and the animal, despite succeeding with the motion effects.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1
- + Excellent typography with clean, consistent serif lettering
- + Professional banner design with balanced negative space
- + Accurate rendering of the accent mark in 'Caffè'
- − Failed the light background prompt requirement by using a black background
- − The steam effect is a single, somewhat thick line
GPT Image 1 Mini
- + Good use of cross-hatching to create a vintage textured feel
- + Dynamic steam illustration above the cloche
- + Clear and readable text placement
- − Failed the light background prompt requirement by using a black background
- − Inconsistent letter sizing, particularly the 'R' and 'O' in 'FLORIAN'
- − The 'Est. 1720' text is slightly off-center within the banner ribbons
Verdict: Both models failed to adhere to the 'light background' instruction, delivering high-contrast dark backgrounds instead. GPT Image 1 is the superior logo due to its superior typography and cleaner vector-style lines, whereas GPT Image 1 Mini has slight alignment and font-weight issues that detract from its professionalism.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1
- + Excellent typography rendering with almost no spelling errors.
- + Sophisticated composition with a clear background/foreground relationship.
- + Adheres strictly to the requested NASA-inspired color palette.
- − The 'Lunar Orbit' step is missing/merged into a confusing 'EARLLUNAR' label.
- − The layout is somewhat fragmented and difficult to read as a sequence.
GPT Image 1 Mini
- + Perfectly follows the requested 6-step sequence with logical numbering.
- + Creative use of a trajectory line to connect the infographic elements.
- + Clear, consistent flat-vector iconography.
- − The 'Translunar' icon is an abstract scribbled loop that doesn't fit the professional style.
- − The crop at the bottom cuts off the crew silhouettes and names.
Verdict: GPT Image 1 Mini is the better infographic because it captures all six requested steps in a logical, numbered sequence that is easy to follow. While GPT Image 1 has better text rendering and a more polished single-frame look, it fails to include all specific mission phases and creates nonsensical words like 'EARLLUNAR'.
Explore each model
OpenAI's cost-effective image generation model for when image quality isn't the top priority