OpenAI's cost-effective image generation model for when image quality isn't the top priority
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Grok Imagine Image
#27 of 62 in Text-to-Image
Where the votes landed
GPT Image 1 Mini
66.7%
win rate
Ties
0.0%
Grok Imagine Image
33.3%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent adherence to the glass cube geometry with clear, defined edges.
- + Very high-quality texture on the book and wooden table.
- + Perfectly captured the soft window lighting from the left.
- − The blue sphere is quite large, whereas the prompt asked for a 'small' one.
- − The plant is more to the side than strictly 'behind' the cube.
Grok Imagine Image
- + Accurately depicted a 'small' blue sphere as requested.
- + The plant is positioned directly behind the cube according to the prompt.
- + The lighting and shadows on the table are very realistic.
- − The 'glass cube' appears more like a rectangular prism or tall block than a cube.
- − The edges of the book and glass contain some minor artifacts where they meet.
Verdict: Both models followed the complex spatial instructions well. GPT Image 1 Mini produced a more visually pleasing image with superior textures and a perfect cube, though it ignored the 'small' descriptor for the sphere. Grok Imagine followed the 'small' and 'behind' prompts more accurately, but the central object is not a cube and the overall composition feels slightly more cluttered. GPT Image 1 Mini is the winner for its superior aesthetic quality and structural accuracy.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully integrated both source images into one scene
- + Preserved the man's identity, hairstyle, and clothing with high accuracy
- + Maintained the specific car's model details and color
- − The man's hands on the steering wheel have minor anatomical blending issues
- − Scale of the man relative to the car is slightly large
Grok Imagine Image
- + Beautiful landscape and lighting composition
- + Dynamic motion blur on the wheels and road
- − Completely failed to use the man from the source image
- − Generated a generic older man instead of the requested subject
Verdict: GPT Image 1 Mini was the only model to successfully perform the image editing task by combining the two provided source images. While Grok Imagine produced a high-quality photo of a car on a coastline, it ignored the specific instruction to use the man from the second source image, whereas GPT Image 1 Mini captured his likeness and clothing perfectly.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent skin texture and natural facial features on the man.
- + Very realistic depiction of wet asphalt and rain reflections.
- + High visual quality with a convincing shallow depth of field.
- − The white car in the background is static, missing the requested motion blur.
- − The bike's kickstand and frame geometry are physically impossible/nonsense.
Grok Imagine Image
- + Perfectly captures the 'motion blur from passing cars' request.
- + The 'imperfect framing' feels much more authentic to a candid street photo.
- + Film-like aesthetics that match the 50mm lens and no-stylization request.
- − The man's face is obscured and less detailed than in the other image.
- − The bicycle spokes and frame details are a bit messy upon close inspection.
Verdict: GPT Image 1 Mini produces a more detailed and aesthetically pleasing portrait, but it fails to incorporate the requested motion blur for the cars. Grok Imagine captures the requested 'candid' and 'motion blur' elements much more effectively, resulting in a more convincing street photography look despite the lower detail on the subject's face.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent weathered texture on the plate armor engravings
- + Strong cinematic lighting consistent with torchlight
- + Realistic facial skin texture and natural eyes
- − Failed to include the specific request for beads in the hair
- − Braids look more like dreadlocks or matted hair than clean braids
Grok Imagine Image
- + Perfectly adhered to the request for beads in the braided hair
- + Intense, lifelike eyes with clear reflections
- + Excellent representation of the leather straps and cloth underlayer
- − The torch flame in the background is a bit distracting and less out-of-focus than the bokeh sparks requested
- − The dirt on the face looks slightly more like makeup or paint than natural battle grime
Verdict: Grok Imagine is the superior choice for this prompt as it successfully included almost every specific detail, including the hair beads and the leather/cloth underlayers which were largely missing or obscured in the other image. While GPT Image 1 Mini produced a very cinematic and gritty texture on the armor, it failed to render the specific decorative elements requested for the character's hair.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text legibility and clean font choices
- + Perfectly maintains a strict grid layout for food photos
- + Highly minimalist and modern aesthetic
- − The menu content is entirely empty (placeholders only)
- − A bit too sterile, feels more like a template than a finished design
Grok Imagine Image
- + Provides full menu content with dish names and descriptions
- + Dynamic use of color and food photography to fill space
- + Effective use of vibrant accents and typographic hierarchy
- − Text includes several AI hallucinations and spelling errors
- − The grid layout isn't strictly followed compared to Model A
- − Many repeated items in the list (Steak Frites, Grilled Salmon)
Verdict: GPT Image 1 Mini creates a superior template with perfect clarity and a strong grid, though it fails to include any actual menu items. Grok Imagine feels more like a finished product with rich content and vibrant visuals, but it suffers from significant spelling errors and repetitive text. GPT Image 1 Mini is the better choice for a professional design foundation, while Grok Imagine captures the 'casual dining' energy better.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with a cohesive glowing ember effect
- + Clean and realistic food photography style
- + Well-balanced vertical composition
- − The 'exploded' effect is a bit static and vertically aligned
- − Missing some of the requested 'fiery background' detail compared to the other model
Grok Imagine Image
- + Highly dynamic and energetic 'exploded' effect with many suspended components
- + Strong adherence to the fiery background and glowing embers requirement
- + Vibrant colors and high sense of motion
- − The price tag starburst is missing the fiery/glowing effect requested
- − Composition is a bit cluttered with overlapping elements
Verdict: GPT Image 1 Mini produces a more professional and polished advertisement with superior typography and lighting. However, Grok Imagine Image much better captures the 'exploded' and 'dynamic' aspects of the prompt, as well as the fiery background, making it the more exciting interpretation of the request.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text legibility and accuracy
- + Uniform chalk texture across all characters
- + Balanced and centered composition
- − Text looks slightly too clean, bordering on a digital font filtered to look like chalk
- − Title is in print capitals rather than the requested 'elegant cursive'
Grok Imagine Image
- + Highly authentic chalk texture with realistic smudges and dust on the board
- + Handwriting has more natural variations in slant and pressure
- + Captured the 'elegant cursive' elements better in certain parts of the script
- − Slight alignment issues with some letters and prices
- − The background is slightly more distracting compared to the clean frame of Model A
Verdict: Both models followed the complex text prompt exceptionally well with zero spelling errors. GPT Image 1 Mini provides a cleaner, more legible result but feels slightly more artificial, whereas Grok Imagine delivers a much more convincing chalk-on-board texture with realistic smudging and human-like handwriting imperfections.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully integrated the character's clothing, face, and accessories.
- + Maintained the yellow background and red ottoman from Image 1.
- + Matched the lighting style of the environment.
- − Failed to replicate the exact complex leg-crossing pose from Image 1.
- − The character's body orientation is more upright than the requested dynamic lean.
- − Anatomy of the foot on the ottoman is slightly distorted.
Grok Imagine Image
- + Maintained the high-quality resolution of the source image.
- − Completely failed the edit instruction to change the character.
- − Simply output a copy of Image 1 without any character features from Image 2.
- − Zero adherence to the request to swap the person.
Verdict: GPT Image 1 Mini followed the complex multi-image instruction by successfully placing the character from Image 2 into the setting and style of Image 1, even if the exact leg-cross pose was slightly simplified. Grok Imagine Image failed the task entirely, returning an unedited version of Image 1 with no character modifications. GPT Image 1 Mini is the clear winner for actually performing the requested synthesis.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent surface textures on the space suit and horse fur.
- + Moody, cinematic lighting with realistic starfield and moon integration.
- + Subtle and sophisticated composition.
- − Failed the negative constraint; it shows an astronaut riding a horse instead of a horse riding an astronaut.
- − Anatomical issues with the horse's front legs being unevenly sized.
Grok Imagine Image
- + Successfully followed the specific 'horse on top' spatial instruction.
- + Vibrant and surreal color palette with a dynamic nebula background.
- + High level of detail on the astronaut's life support pack and gloves.
- − The horse's back right leg is missing its hoof.
- − The concept of 'riding' is loosely interpreted as the horse is more floating above than riding.
Verdict: GPT Image 1 Mini failed the specific prompt constraint of having the horse on top, providing a literal and cliché interpretation. Grok Imagine followed the difficult spatial instruction correctly and embraced the surreal aspect of the prompt with a more imaginative composition and vibrant colors.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent replication of the specific outfit layers (scarf, coat, belt, and watch).
- + High visual quality with realistic fabric rendering and lighting.
- − Failed to preserve the original person's face, hair, and vitigilo patterns accurately.
- − Changed the composition and background slightly instead of a pixel-perfect overlay.
Grok Imagine Image
- + Perfect preservation of the original person's face, hair, and specific skin vitiligo textures.
- + Kept the background and composition exactly as the source image.
- − Completely ignored the clothing in Image 2, generating a generic 'elaborate' royal outfit instead.
- − The hand integration with the new clothing is physically awkward.
Verdict: Both models failed the complex instruction in different ways. GPT Image 1 Mini was the only one that followed the instruction to use the specific outfit from Image 2, though it failed to preserve the base person's identity. Grok Imagine preserved the person perfectly but completely ignored the provided reference image for the clothing. GPT Image 1 Mini is a slightly better attempt at image-to-image editing, even though it struggled with identity preservation.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent fur texture and photorealistic lighting
- + Cinematic composition with a shallow depth of field
- + Accurate depiction of a professional 'taxi driver' cap and jacket
- − The passenger is very blurry and slightly out of focus
- − Only one paw is clearly on the steering wheel
Grok Imagine Image
- + Perfect adherence to showing both paws on the steering wheel
- + Very clear depiction of the passenger and her expression
- + Strong sense of setting with the visible Manhattan street through the windshield
- − Compositional error places the passenger in the front seat instead of the back seat
- − The capybara's paws look slightly like bird talons or sharp claws
- − The cap looks more like a baseball cap than a traditional taxi driver hat
Verdict: GPT Image 1 Mini produces a much more realistic and cinematic image with superior lighting and texture. While Grok Imagine captures more of the specific prompt details like 'both paws', it fails the core compositional instruction by placing the passenger in the front seat, whereas GPT Image 1 Mini correctly places her in the back.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent layout with a centralized and balanced composition
- + Includes all requested text with perfect spelling
- + Rich, vintage color palette that feels cohesive
- − The 'scroll' banner is somewhat flat compared to the other elements
- − Textures are a bit soft/grainy
Grok Imagine Image
- + Strong 'parchment' aesthetic with torn edges and thorns
- + Dynamic background with atmospheric lighting and a moon
- + High contrast and sharp details on the pumpkin and bats
- − The large 'thorns' border feels a bit cluttered and overlaps the text
- − Font choice for the bottom event details is very basic compared to the gothic header
Verdict: GPT Image 1 Mini delivers a more professional and balanced design that feels like a real invitation, specifically excelling in typography and layout. Grok Imagine Image has more impressive individual assets and a great parchment texture, but the composition feels slightly more cluttered and the font choice for the event details is jarringly modern.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully added a very thick head of hair with high volume.
- + Maintained the general aesthetic and lighting of the scene.
- − Significantly altered the facial features, making the person look like a different individual.
- − The hair texture and edges look slightly airbrushed compared to the grit of the original image.
Grok Imagine Image
- + Excellent source preservation, keeping the original facial features and identity perfectly intact.
- + The hair texture and color match the existing beard perfectly.
- + Natural integration with the glasses and forehead.
- − The hair is somewhat thin at the top, not fully achieving the 'full, thick' request compared to Model A.
Verdict: Grok Imagine succeeded by seamlessly integrating the new hair while preserving the subject's identity and the original image quality. GPT Image 1 Mini provided a thicker volume of hair as requested but failed as an edit by fundamentally changing the subject's face.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography integrated naturally with the flag icon
- + Very clean clay-like 3D textures that match the cartoon scene request
- + Superior framing and adherence to the isometric perspective
- − The diorama base is a bit large compared to the plate
Grok Imagine Image
- + Includes a wider variety of sushi types
- + Good rendering of translucent materials like fish and salmon roe
- − The text 'JAPAN' is smaller and less bold than requested
- − Aspect ratios of objects like the soy sauce bowl feel slightly distorted
- − Visible aliasing and lower clarity on the text and flag edges
Verdict: GPT Image 1 Mini followed the typography instructions perfectly, creating a clean and professional graphic design layout. While Grok Imagine produced a visually appealing variety of sushi, its text rendering and overall composition lacked the 'ultra-clean' polish and bold presence requested in the prompt.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent hand-drawn caricature art style
- + Good preservation of character features from the source image
- + Incorporates all elements in a cohesive, traditional caricature layout
Grok Imagine Image
- + Very creative inclusion of the dog in hockey gear
- + High-quality rendering with a professional TV studio atmosphere
- + Excellent text rendering of 'EVENING NEWS'
- − Character face looks slightly more generic and less like the specific source person compared to Model A
- − The hands on the desk are slightly malformed
Verdict: Both models followed the complex prompt effectively, but GPT Image 1 Mini captured the essence of a 'caricature' better by using a hand-drawn illustration style and maintaining stronger facial recognition from the source image. Grok Imagine Image provided a more humorous interpretation with the skating dog and professional studio setting, though it leans more toward a 'big-head' digital render than a traditional caricature.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1 Mini
- + Natural, dynamic composition that captures the requested 'tumbling' and 'chasing' action.
- + Excellent anatomical realism for all four animal types.
- + Subtle, realistic lighting and soft fur textures that align with the hyper-photorealistic request.
- − The god rays are a bit soft, though they are present.
Grok Imagine Image
- + Strong, dramatic god rays and vibrant golden hour lighting.
- + Lush flower variety in the foreground.
- − Static, posed composition fails to capture the 'chasing' and 'tumbling' action requested.
- − Artificial, doll-like appearance of the animals that borders on 'uncanny' rather than photorealistic.
- − Insects look like generic white blobs rather than clear butterflies.
Verdict: GPT Image 1 Mini is the clear winner as it successfully captures the energy and movement of the animals described in the prompt while maintaining a high level of photorealism. Grok Imagine Image produces a more static, AI-stylized result with animals that look like figurines, and it ignores the 'chasing and tumbling' part of the prompt in favor of a portrait layout.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent preservation of the original subjects' facial expressions and postures
- + Beautiful hand-painted texture resembling colored pencils or pastels
- + Maintains the warm, nostalgic mood requested in the prompt
- − The color palette is a bit monochromatic and lacks the vibrant blues/greens often found in Ghibli films
- − The background is very blurry and lacks the detailed scenery typical of the inspiration style
Grok Imagine Image
- + Successfully captures the iconic Ghibli character design style with clean line work
- + Beautiful background detail with fluffy clouds and charming architecture
- + Excellent use of pastel colors and soft lighting
- − The male subject's beard and facial features are slightly too realistic compared to the anime-style women
- − The woman in the foreground has a somewhat stiff expression compared to the source
Verdict: Grok Imagine is the winner because it successfully transforms both the characters and the environment into the Studio Ghibli aesthetic, including the signature cloud and background style. While GPT Image 1 Mini does a great job of preserving the original image's nuance and applying a nice texture, it feels more like a filter than a full stylistic transformation into the world of Ghibli.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully rendered windswept hair that looks natural for the subject.
- + Preserved the composition of the original scene well while adding dynamic elements.
- − Noticeably changed the woman's facial features and identity during the edit.
- − The flying leaves are slightly blurry compared to the rest of the sharp environment.
Grok Imagine Image
- + Perfectly preserved the facial identity and anatomy of the woman from the source image.
- + Added a high volume of leaves which strongly conveys the requested 'energetic' feel.
- + Included subtle motion on the dog's ears to match the wind effect.
- − The 'blowing hair' effect is less pronounced and looks slightly more static than Model A.
- − A few leaves appear as small orange dots without detailed textures.
Verdict: Grok Imagine is the superior choice for this image editing task because it successfully added all requested dynamic elements while maintaining the subject's identity. GPT Image 1 Mini significantly altered the woman's face, making it look like a different person, whereas Grok Imagine kept the source image's integrity intact while fulfilling the 'wind and leaves' prompt.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography including the correct grave accent on the letter 'è'
- + Follows the request for a banner element naturally
- + Clean, high-contrast vector feel
- − Failed to provide a 'light background' as requested, using black instead
- − The gold/brown tone is a bit grainy and yellowish compared to the prompt
Grok Imagine Image
- + Perfectly adheres to the 'light background' and 'subtle texture' requirements
- + Beautifully integrated vintage warm brown and cream color palette
- + Professional vector nesting of elements
- − Redundant text displaying 'Est. 1720' twice
- − The cloche dome has strange artifacts like a mug handle and spoon protruding from it
Verdict: Grok Imagine adhered much better to the stylistic requirements of the prompt, specifically the light background and color scheme, whereas GPT Image 1 Mini defaulted to a black background. However, GPT Image 1 Mini produced cleaner typography and a more logical icon, avoiding the strange extra appendages found on Grok's cloche.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with zero spelling errors.
- + Clean, consistent vector line-art style following the prompt's constraints.
- + Logical flow of information despite the non-linear layout.
- − The 'Translunar' icon is an abstract scribbled loop that doesn't clearly represent a trajectory.
- − The layout is slightly cluttered in the center.
Grok Imagine Image
- + Stronger infographic layout with a clear title and branded NASA header.
- + Effective use of the 'Tranquility' pin icon and lunar surface background.
- + Better visual representation of the trajectory arc for step 3.
- − Numerous spelling errors including 'Transluiory', '3rajoory', and 'Moom'.
- − Text overlaps icons and elements in several places, reducing readability.
- − The crew names are inconsistent with the icons above them.
Verdict: GPT Image 1 Mini produced a much cleaner and professional-looking graphic with perfect text rendering, whereas Grok Imagine Image struggled significantly with spelling and internal consistency. While Grok Imagine Image had a more ambitious layout and better trajectory iconography, the typos and jumbled text labels make it inferior for a real-world infographic use case.
Explore each model
An image generation model by xAI designed to generate highly aesthetic images from text descriptions.