OpenAI's cost-effective image generation model for when image quality isn't the top priority
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
GPT Image 1 Mini
#13 of 62 in Text-to-Image
Grok Imagine Image Pro
#17 of 62 in Text-to-Image
Where the votes landed
GPT Image 1 Mini
100.0%
win rate
Ties
0.0%
Grok Imagine Image Pro
0.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent photographic realism and texture in the book and sphere.
- + Accurate lighting direction consistent with the prompt.
- + Perfectly clear glass rendering with subtle reflections.
- − The blue sphere is floating or partially intersecting the back glass pane rather than sitting on the surface.
- − The cube structure has an odd, thin metal-like frame rather than being solid glass.
Grok Imagine Image Pro
- + Physically grounded sphere sitting on the base of the cube.
- + Dynamic refraction of the sphere and plant within the thick glass.
- − The glass geometry is warped and inconsistent at the corners.
- − The blue sphere is inexplicably reflected/duplicated on the right side within the glass despite no physical wall being there.
Verdict: GPT Image 1 Mini produces a much more aesthetically pleasing and high-resolution image with superior textures, though the sphere appears to be levitating inside the cube. Grok Imagine Image Pro handles the physical placement of the sphere and the plant's visibility through the glass more realistically, but suffers from significant glass distortion and a strange ghosting artifact of the sphere.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent preservation of the man's identity, including his face, hairstyle, and specific clothing (plaid coat and scarf).
- + High accuracy in maintaining the specific car model and interior details from the source image.
- + Strong dynamic composition with realistic motion blur and lighting.
- − The hands on the steering wheel have some slight anatomical distortion.
Grok Imagine Image Pro
- + Beautiful scenic background that fits the 'California coastline' description perfectly.
- + Good sense of speed and motion through the wheels and road surface.
- − Completely failed to preserve the man's identity, replacing the model with a different person in different clothes.
- − The car's scaling relative to the road feels slightly off in the winding perspective.
Verdict: GPT Image 1 Mini is the clear winner as it successfully performed the complex task of merging two source subjects—the specific man and the specific car—into a new environment while preserving their details perfectly. Grok Imagine Image Pro failed the primary edit instruction by replacing the man with a generic driver, despite creating an aesthetically pleasing background.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent shallow depth of field and bokeh transition
- + Strong adherence to 'imperfect framing' with a close-up, candid feel
- + Atmospheric lighting and realistic ground reflections
- − Anatomical issues with the man's hands merging into the bike spokes
- − Missing visible motion blur from passing cars
- − The bicycle has structural issues where the kickstand and frame meet
Grok Imagine Image Pro
- + Successfully includes motion blur from passing cars
- + Better anatomical correctness in the hands and clear use of a tool
- + Environment feels more like a specific Japanese street scene
- − Depth of field is slightly too deep compared to the 50mm request
- − Street reflections are less pronounced than in Model A
- − The image looks a bit more 'clean' and less like a moody candid shot
Verdict: Grok Imagine Image Pro is the winner for its better handling of complex details like hands and tools, as well as capturing the motion blur requested in the prompt. While GPT Image 1 Mini has a more cinematic and moody aesthetic with superior bokeh, it suffers from significant AI artifacts where the person interacts with the bicycle.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent photographic realism in skin texture and eyes
- + Naturalistic lighting with subtle warm hues
- + Very detailed engraving on the plate armor
- − Failed to include the requested beads in the braided hair
- − Scars are very faint, leaning more towards just dirt
Grok Imagine Image Pro
- + Followed all prompt instructions including hair beads and distinct scars
- + Impressive detail on leather straps and cloth underlayer
- + Dynamic composition with clear Latin text engraving
- − The sparks look like flat digital streaks rather than bokeh light
- − Skin texture looks slightly more airbrushed compared to Model A
Verdict: Grok Imagine Image Pro is the winner because it followed every specific detail of the prompt, including the beads in the hair and the leather straps/cloth textures, which GPT Image 1 Mini missed. While GPT Image 1 Mini has slightly more realistic skin rendering, Grok Imagine Image Pro provides a more complete and visually interesting interpretation of a battle-worn paladin.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with clean, perfectly rendered bold sans-serif fonts
- + Uses a grid layout for food photos as requested
- + Minimalist aesthetic captures the 'casual dining' vibe well
- − The placeholder lines for menu items are empty, making it look like a template rather than a design
- − The food images are slightly inconsistent in lighting and style
Grok Imagine Image Pro
- + Excellent visual hierarchy with photos directly above their corresponding descriptions
- + High-quality, vibrant food photography that looks appetizing
- + Complete design including descriptions, prices, and contact information
- − The font rendering is poor with numerous spelling errors in descriptions (e.g., 'Avucado', 'Pepperani', 'Salmom')
- − The text descriptions for the Appetizers section repeat the pizza description text
Verdict: GPT Image 1 Mini produced a cleaner, more professional-looking template with perfect text rendering, but it failed to include actual menu content. Grok Imagine Image Pro created a much more functional and vibrant menu layout with descriptions and prices, but it suffered from significant spelling errors and repetitive text. GPT Image 1 Mini is preferred for its cleaner design aesthetics, though Grok Imagine Image Pro came closer to a finished marketing piece despite the typos.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with a consistent glowing, fiery effect.
- + Simple and clean layout that feels like a cohesive brand advertisement.
- + Accurate rendering of all requested text elements in the specified styles.
- − The burger feels somewhat static for an 'exploded' view.
- − The lighting on the burger doesn't fully match the intensity of the fiery background.
Grok Imagine Image Pro
- + Dynamic sense of motion with sauce splashes and tilted components.
- + Higher level of textural detail on the meat patty and fresh vegetables.
- + Background embers and smoke create a strong energetic atmosphere.
- − The text doesn't follow the 'fiery, glowing' prompt, looking like flat vector art instead.
- − The starburst element is generic and lacks the glowing effect requested.
- − Text rendering on '6,99' uses a comma instead of a period and looks poorly integrated.
Verdict: GPT Image 1 Mini followed the stylistic prompt for the text much better, creating a professional and cohesive glow across all elements. While Grok Imagine Image Pro captured a more dynamic and appetizing 'exploded' burger, its failure to render the fiery text effects and the low-quality starburst makes it less successful as a complete advertisement according to the prompt.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent chalk texture with grainy, realistic strokes
- + Perfectly centered and clean composition
- + High-fidelity rendering of the wooden frame
- − Failed to provide 'elegant cursive' for the title as requested
- − Text looks a bit too uniform, bordering on a digital font appearance
Grok Imagine Image Pro
- + Successfully used cursive for the menu items to create a more authentic handwritten feel
- + Captures the 'cozy café' atmosphere with a visible background and environmental lighting
- + Variation in handwriting styles makes it look more realistic
- − The title is in print instead of the requested 'elegant cursive'
- − Text layout is slightly less balanced with some spacing issues near the prices
Verdict: Both models followed the complex text instructions very well, with almost no spelling errors. GPT Image 1 Mini produced a cleaner, more readable board with superior chalk texture, but Grok Imagine Image Pro felt more realistic as a photograph because of the café background and the varied handwriting styles. Grok is the narrow winner for capturing the 'handwritten' and 'cozy' spirit of the prompt more effectively despite the title not being cursive.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully transferred the character's facial features, sunglasses, and clothing style
- + Accurately matched the lighting and background color of the first source image
- + Correctly interpreted the identity transformation requested
- − Significantly simplified and altered the specific complex pose from Image 1
- − Anatomical issues in the right hand and feet placement compared to the original pose
Grok Imagine Image Pro
- + Maintained the exact complex pose from Image 1 with high fidelity
- + High visual quality and resolution of the overall image
- − Completely failed to use the character from Image 2 as requested
- − Only slightly modified the face of the original woman from Image 1 instead of replacing her with the man from Image 2
Verdict: GPT Image 1 Mini followed the core instruction of swapping the characters, successfully placing the man from Image 2 into the scene with his original clothing, despite struggling to recreate the exact complexity of the pose. Grok Imagine Image Pro completely ignored the character reference, essentially reproducing the source image with a slightly different feminine face. Therefore, GPT Image 1 Mini is the winner for actually attempting and largely achieving the person-transfer edit.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1 Mini
- + High cinematic visual quality and atmosphere
- + Realistic lighting and texture on the space suit and horse
- + Excellent composition with the moon element
- − Failed the negative constraint: astronaut is riding the horse instead of horse riding the astronaut
- − Lacks the surreal interpretation requested in the prompt
Grok Imagine Image Pro
- + Successfully followed the specific instruction for the horse to be on top
- + Vibrant colors and a strong surreal feel
- + Intricate details in the nebula and background planet
- − The horse is floating above rather than 'riding' the astronaut
- − Slightly messy anatomical transition where the horse's back leg meets the astronaut's pack
Verdict: While GPT Image 1 Mini produced a much more photorealistic and aesthetically pleasing image, it completely failed to follow the unusual spatial instruction of having the horse on top. Grok Imagine Image Pro correctly interpreted the surreal 'horse on top' request, making it the clear winner for prompt adherence despite a slightly more cluttered composition.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully extracted and applied the exact outfit from Image 2, including the specific plaid pattern and coat style.
- + Maintains the skin vitiligo details on the hands and face.
- + The composition and perspective feel cohesive with the new body pose.
- − Changed the person's face and head shape significantly from Image 1.
- − The sand on the face and the specific hair patch from Image 1 were lost in the generation.
- − The transition between the neck and the clothing is slightly blurry.
Grok Imagine Image Pro
- + Perfectly preserved the original face and hair from Image 1.
- + High-quality rendering of fabric textures and lighting.
- + Maintains the exact wooden structure and background from the source image.
- − Completely ignored the instruction to use the outfit from Image 2, instead generating a generic royal costume.
- − Created white-skinned hands that do not match the person in Image 1.
Verdict: This is a case of two different failures: GPT Image 1 Mini captured the correct outfit from Image 2 but failed to preserve the person's identity from Image 1, while Grok Imagine Image Pro perfectly preserved the face/background but completely ignored the target outfit. Model A is slightly preferred because it at least attempted to combine elements from both images, whereas Model B hallucinated an entirely new set of clothing.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent close-up detail on the capybara's fur and expression.
- + Strong atmospheric lighting and authentic bokeh effect.
- − The passenger is very blurry and lacks detail.
- − Only one paw is clearly visible on the steering wheel.
Grok Imagine Image Pro
- + Perfect adherence to all prompt details including the jacket, two paws on the wheel, and the woman's expression.
- + Impressive text rendering on the hat with specific NYC medallion details.
- + Very clear composition that shows both subjects and the city environment equally well.
- − The capybara's paws look more like crab claws/talons than actual capybara paws.
- − The perspective from the dashboard looking back is slightly less intimate than Model A.
Verdict: While GPT Image 1 Mini takes a portrait-style approach with great lighting, Grok Imagine Image Pro provides a much more complete realization of the prompt. Grok successfully captures the capybara's full outfit, the specific seating arrangement, and legible text on the driver's cap, making it the more accurate and detailed interpretation.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with clean, legible text
- + Sturdy composition with a well-integrated border
- + Captures an authentic vintage parchment texture and moody lighting
- − Lacks the small scroll banner requested, instead using a simple ribbon
- − The background trees and bats are slightly muddy and less distinct
Grok Imagine Image Pro
- + Highly detailed environment with visible thorns, spiders, and a moonlit sky
- + Accurately includes the requested scroll banner with a wax seal
- + More literal interpretation of the border with webs and thorns
- − The main title text is cramped and the 'y' in 'Party' is partially cut off
- − The text on the scroll banner is awkwardly small and slightly off-center
Verdict: GPT Image 1 Mini produced a more professional-looking graphic design with superior typography and layout, whereas Grok Imagine Image Pro followed the specific physical requirements of the prompt more closely including the scroll and detailed border elements. However, GPT Image 1 Mini's clean execution and readable text make it a better overall invitation.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1 Mini
- + Fulfilled the request for thick, voluminous hair.
- + Successfully adjusted the lighting on the forehead to match the new hairline.
- − Significantly altered the facial structure, making the man look younger and like a different person.
- − The hair texture looks somewhat digital and painterly compared to the original photo's grain.
Grok Imagine Image Pro
- + Excellent source preservation, keeping the man's face and features exactly as they appeared in the original.
- + The hair texture and lighting are perfectly integrated with the original photographic style.
- + The hairline and blend into the temples are highly realistic.
- − The hairstyle is more conservative than 'full' or 'thick', though it fits the character well.
Verdict: Grok Imagine Image Pro performed much better as an image editor by perfectly preserving the subject's identity and facial features while adding realistic hair. GPT Image 1 Mini fundamentally changed the person's face, resulting in a different person entirely, which fails the primary goal of a portrait edit.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1 Mini
- + Perfectly clean and bold text rendering with no shadows or distortions.
- + Excellent miniature cartoon aesthetic with soft, refined 3D textures.
- + Clean composition with a well-defined isometric diorama base.
- − The sushi count is a bit minimal compared to typical 'signature dish' representations.
- − The flag icon is slightly merged with the text block rather than being a separate element.
Grok Imagine Image Pro
- + Greater variety of sushi types including rolls, which feels more comprehensive.
- + Stronger '3D render' lighting with subtle drop shadows on the text.
- + More detailed garnishes like ginger and wasabi.
- − Small artifacts in the text rendering, such as the slightly inconsistent 'S' in SUSHI.
- − The rice texture looks a bit like individual plastic beads rather than 'soft refined textures'.
- − The diorama base has a slight perspective warp on the right edge.
Verdict: Both models followed the prompt exceptionally well, but GPT Image 1 Mini takes the lead due to its superior cleanliness and professional graphic design feel. While Grok Imagine Image Pro offers more visual variety in the sushi itself, GPT Image 1 Mini's text is much sharper and the soft toy-like textures perfectly match the requested 'refined miniature' aesthetic.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent hand rendering on both the character and the dog hug.
- + Consistent colored pencil art style throughout the image.
- + Maintains character likeness while applying appropriate caricature exaggeration.
- − Composition is a bit simple compared to the 'exaggerated' request.
- − The hockey stick is slightly warped near the top.
Grok Imagine Image Pro
- + Successfully incorporates all elements with high energy and humor.
- + Clear, legible text rendering that enhances the 'TV anchor' theme.
- + Creative use of the 'Puppy of the Day' graphic and trophy.
- − Hand holding the hockey stick has anatomical issues (six fingers/severed appearance).
- − The dog in the helmet has awkward blending between its head and the gear.
- − Style is more like a digital sticker than a traditional caricature.
Verdict: Both models followed the prompt successfully, including the hockey and dog elements. GPT Image 1 Mini feels more like a cohesive work of art with better anatomical correctness, but Grok Imagine Image Pro captures the spirit of a 'humorous and exaggerated' caricature much better by adding several extra thematic elements like the puppy wall and the hockey trophy. Although Grok has noticeable AI artifacts in the hands and dog details, its creative interpretation of the profession makes it the more entertaining result.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent fur texture and lighting integration.
- + Natural, dynamic poses that feel like a real action snapshot.
- + Superior rendering of 'god rays' and soft, atmospheric depth.
- − The fox's anatomy is slightly wonky, particularly the mouth/teeth area.
- − The butterfly above the kitten lacks a shadow or realistic lighting interaction.
Grok Imagine Image Pro
- + Accurately included more varied butterflies and wildflowers.
- + Great interaction between animals, with the fox tumbling as requested.
- + Very sharp focus across all subjects.
- − The lighting feels more synthetic and 'over-processed' compared to Model A.
- − Added an extra kitten not explicitly requested, crowding the composition.
- − The fur texture looks slightly more like a digital painting than a photograph.
Verdict: GPT Image 1 Mini captures the requested 'hyper-photorealistic' look much better, with natural lighting and soft focus that creates a convincing atmospheric scene. Grok Imagine Image Pro follows the specific actions (tumbling) and flower types more closely, but the image feels less realistic and more like a high-quality digital illustration.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent hand-drawn texture with a soft colored pencil or pastel feel
- + Great warm and nostalgic color palette
- + Good capture of the characters' expressions in an illustrative style
- − Faces look a bit more generic compared to the source than Model B
- − Lines are somewhat fuzzy due to the heavy texture
Grok Imagine Image Pro
- + Perfectly captures the Studio Ghibli cel-shaded aesthetic
- + Excellent preservation of the source characters' likenesses and poses
- + Clean background that feels like a hand-painted Ghibli environment
- − The man's eyes are slightly misaligned in a way that feels a bit googly
- − Could have slightly softer pastel hues to match the prompt's specific request
Verdict: Both models did an excellent job translating the meme to an illustrative style. GPT Image 1 Mini leans more into the texture of a hand-drawn illustration with warm tones, while Grok Imagine Image Pro perfectly captures the specific 'Ghibli anime' aesthetic with clean lines and painted backgrounds, while preserving the source image's layout and characters more accurately.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1 Mini
- + Successfully added wind-blown hair effect
- + Includes falling leaves as requested
- + Maintains high resolution and color vibrancy
- − Significantly alters the woman's face and features
- − Modified the dog's appearance and the background bridge
- − The hair physics look a bit stiff
Grok Imagine Image Pro
- + Excellent source preservation, keeping the face and dog identical to the original
- + Superior hair motion rendering with natural strands
- + Very high density of leaves creates a strong sense of wind
- − Some leaves look slightly flat or lack motion blur
Verdict: Grok Imagine Image Pro is the clear winner because it successfully applied all requested edits—blowing hair and falling leaves—while perfectly preserving the woman's face and the dog from the source image. GPT Image 1 Mini failed at source preservation, regenerating the woman's face and changing various background details, which is undesirable in an image editing task.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent typography with correct accent usage on 'Caffè'
- + Strong vintage texture and aesthetic
- + Clear and readable 'Est. 1720' banner
- − Failed the 'light background' prompt instruction entirely
- − Cloche illustration is a bit heavy-handed and less refined
Grok Imagine Image Pro
- + Followed all color and background instructions perfectly
- + Clean vector emblem style with a professional layout
- + Elegant steam and cloche illustration
- − Text is slightly distorted and lacks the accent mark on 'Caffè'
- − The 'Est. 1720' text is slightly off-center within the banner
Verdict: Grok Imagine Image Pro is the winner for adhering to the requested color palette and 'light background' instruction, which GPT Image 1 Mini ignored by using a black background. While GPT Image 1 Mini had superior typography, Grok Imagine Image Pro delivered a more professional-looking vector logo that fits the minimalist vintage prompt more effectively.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1 Mini
- + Excellent text legibility and clean line art
- + Strong adherence to the requested NASA-inspired color palette
- + Clear, bold iconography that matches the flat-vector style
- − The 'Translunar' icon is a messy scribble rather than a clear trajectory arc
- − Includes a redundant, floating navy circle next to the moon
- − Layout feels cramped with inconsistent spacing between the numbered steps
Grok Imagine Image Pro
- + Professional, balanced infographic composition with a logical vertical flow
- + High level of detail in text rendering, including specific crew names and landing sites
- + Consistent use of circular framing for all icons which enhances the modern aesthetic
- − The icons are smaller and slightly less 'crisp' for a vector style compared to Model A
- − The horizontal red arc for 'Translunar' is a bit thin and abstract
Verdict: Grok Imagine 2.0 Pro is the clear winner for its superior layout and information hierarchy, successfully including all requested steps plus accurate auxiliary details like crew names. While GPT-4o-mini (via Image 1 Mini) has slightly cleaner individual line art, it fails significantly on the 'Translunar' step and has a much more cluttered, less professional composition.
Explore each model
xAI's premium image generation model offering higher fidelity output and stronger performance on single-image editing benchmarks compared to the standard Grok Imagine model