An image generation model by xAI designed to generate highly aesthetic images from text descriptions.
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
Grok Imagine Image
#27 of 62 in Text-to-Image
Wan 2.7
#39 of 62 in Text-to-Image
Where the votes landed
Grok Imagine Image
0%
win rate
Ties
0%
Wan 2.7
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
Grok Imagine Image
- + Excellent soft window lighting from the left that feels natural
- + Very high photorealism in the wood grain and glass texture
- + Accurately represents the transparency of the glass cube and the plant behind it
- − The blue sphere appears to be floating mid-air inside the cube rather than resting on the bottom
- − The glass cube has unusually thick edges that look more like a frame
Wan 2.7
- + The blue sphere realistically rests on the bottom surface of the cube with a clear reflection
- + Includes more intricate detail on the red book, such as text on the spine
- + Good interpretation of the plant being partially visible through the glass
- − The cube geometry is slightly warped on the right side
- − The reflections inside the glass are somewhat cluttered and confusing
Verdict: Both models followed the prompt instructions well, but Grok Imagine succeeds with superior lighting and texture realism. While Wan 2.7 placed the sphere more logically on the floor of the cube, Grok Imagine's overall photographic quality and better handling of soft window light make it the more visually appealing result.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
Grok Imagine Image
- + Excellent photographic quality and lighting
- + Maintains the specific Rolls Royce model aesthetic well
- + Correctly places a driver in the vehicle with motion blur on the road
- − Completely failed to use the specific man provided in the source image, replacing him with an older man with short gray hair
Wan 2.7
- + Successfully preserved the specific car model from the source image
- + Accurately captured the California coastline environment and road motion
- − The driver is a distorted, low-quality figure that does not resemble the man in the source image
- − Significant resolution artifacts around the driver's head
Verdict: Both models struggled to incorporate the specific person from the second source image into the car from the first. Grok Imagine Image produced a high-quality, professional-looking photo but ignored the person's likeness entirely, replacing him with a generic figure. Wan 2.7 attempted to place a figure in the car, but the result is a blurry, distorted mess that ruins the composition.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
Grok Imagine Image
- + Perfect adherence to technical photography prompts like motion blur and imperfect framing
- + Highly realistic skin texture and film-like color grading
- + Successful use of 50mm shallow depth of field
- − The man is wearing a face mask which obscures much of the 'elderly' facial detail requested
- − The bicycle geometry is slightly warped near the seat post
Wan 2.7
- + Excellent character portrait with clear elderly features and realistic damp clothing
- + High level of detail in the street background and reflections
- − Failed to include motion blur from passing cars
- − The bicycle has significant structural errors, including a disconnected fork and missing seat
- − Lack of shallow depth of field makes the image feel less like the 50mm lens prompt
Verdict: Grok Imagine Image follows the technical prompt much more effectively, accurately depicting the requested motion blur, 50mm focal length, and imperfect framing that gives it a true 'candid' feel. While Wan 2.7 has a strong character model, it fails on several technical instructions and produces a physically nonsensical bicycle with no seat and a floating front wheel.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
Grok Imagine Image
- + Exceptional intricate engraving on the plate armor.
- + Strong atmospheric lighting with warm torchlight reflecting beautifully on the metal surface.
- + Highly detailed cloth texture in the collar area.
- − The character looks somewhat too 'pristine' and youthful for a 'battle-worn' description despite the scars.
- − The background torches feel a bit synthetic compared to the character.
Wan 2.7
- + Excellent depiction of 'battle-worn' with realistic skin texture, deep scars, and a weary expression.
- + Superior leather strap textures and buckle detailing as requested.
- + The hair braiding with beads is very clearly defined and matches the prompt perfectly.
- − The armor engraving is much simpler and less 'ornate' than Model A.
- − Lighting is a bit flatter on the face compared to the dramatic lighting in Model A.
Verdict: Grok Imagine produces a more visually striking, 'heroic' image with incredible armor detail and atmospheric lighting. However, Wan 2.7 better captures the 'battle-worn' grit, realistic facial textures, and specific leather strap details requested in the prompt. Wan 2.7 is preferred for its technical adherence to the weathered character description.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
Grok Imagine Image
- + Follows the requested headers (Appetizers, Pizza, Mains) exactly.
- + Text rendering for main headers is very clean and bold.
- + Integrated food elements like scattered herbs and leaves add a professional look.
- − Sub-text is mostly gibberish or repeats the same items multiple times.
- − Layout is slightly cluttered with overlapping food images and text.
Wan 2.7
- + Excellent grid layout that feels like a real, ready-to-use menu.
- + High-quality food photography with consistent lighting and framing.
- + Includes realistic details such as pricing, address, social media icons, and a QR code.
- − Spelling errors in item names (e.g., 'Calanfri', 'Sannon', 'Bisge').
- − Missing a specific 'Mains' header in the body, using 'Bread & Bowl' as the primary title instead.
Verdict: Wan 2.7 produced a much more realistic and professional menu design that adheres to the 'grid' and 'vibrant accents' request more effectively than Grok Imagine. While Grok Imagine followed the specific section names more closely, Wan 2.7 felt like a complete brand identity with superior composition and high-quality food assets.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
Grok Imagine Image
- + Excellent text rendering with clean, vibrant fiery effects.
- + Highly photorealistic textures on the burger bun and patty.
- + Dynamic composition with a great sense of motion in the sauce splashes.
- − The starburst for the price is a bit simple and lacks the fiery integration of the main title.
Wan 2.7
- + Creative typography with flames emerging directly from the letterforms.
- + Good inclusion of all requested elements including the starburst price tag.
- + Balanced and clean layout suitable for a digital advertisement.
- − The burger ingredients look slightly more illustrative and less photorealistic than Model A.
- − Includes strange floating sesame seeds or nuts that look like artifacts in the upper left/right.
Verdict: Both models followed the prompt exceptionally well, but Grok Imagine Image edges out a victory due to superior photorealism in the food textures and a more natural sense of 'exploded' motion. While Wan 2.7 has creative fiery lettering, Grok Imagine Image's rendering of the patty and bun makes the advertisement more appetizing and professionally polished.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
Grok Imagine Image
- + Excellent chalk texture with realistic dusty strokes and smudges on the board.
- + Authentic handwritten variation with a natural slant and inconsistent baseline.
- + Perfectly rendered text that precisely matches the requested prompt, including the date.
- − The 'elegant cursive' request for the title was interpreted as uppercase print, though it still looks handwritten.
Wan 2.7
- + Perfect text accuracy for every word in the prompt.
- + Attractive composition with decorative dividers.
- + Clean and legible presentation.
- − Text looks like a digital font or a sticker rather than actual chalk on the surface.
- − Lacks the characteristic texture, powder residue, and stroke pressure of real chalk.
- − No 'elegant cursive' applied to the title as requested.
Verdict: Grok Imagine is the clear winner because it successfully captured the 'chalk texture' and 'handwritten style' requested in the prompt, looking like a real physical object. Wan 2.7 produced very accurate text, but the rendering is flat and looks like a digital overlay, failing to meet the requirement for a realistic chalk texture.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
Grok Imagine Image
- + Successfully replicates the lighting and color grade of Image 1.
- + Maintains high resolution and visual clarity.
- − Completely failed to use the character from Image 2.
- − Changed the pose from Image 1, significantly altering the leg and body arrangement.
- − Failed to include any character details (face, hair, clothing) from Image 2.
Wan 2.7
- + Maintains the majority of the environment and pose from Image 1.
- + Preserves the original source image texture and lighting.
- − Completely failed to use the character from Image 2.
- − Made no attempt to integrate the face, hair, or clothing of the second subject.
- − Anatomy on the hands and feet is slightly distorted compared to the original.
Verdict: Both Grok Imagine and Wan 2.7 failed the image editing task of character swapping, instead primarily returning variants of the first image. Grok Imagine went as far as to alter the pose from Image 1, while Wan 2.7 stayed closer to the original pose but ignored the character reference entirely. Since neither model followed the character swap instruction, this is a tie in terms of failure, though Wan 2.7 technically preserved more of the pose reference.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
Grok Imagine Image
- + Successfully followed the specific instruction for the horse to be on top
- + Stunning cinematic color palette and lighting
- + Dynamic and surreal composition
- − The horse's front leg is merging awkwardly with the astronaut's hand
Wan 2.7
- + High resolution and clean rendering
- + Good anatomical detail on both the astronaut and the horse
- − Failed the negative constraint/specific instruction by placing the astronaut on top
- − Very conventional interpretation instead of the requested surreal reversal
Verdict: Grok Imagine followed the difficult prompt logic of having the horse on top of the astronaut, resulting in a unique and surreal image. Wan 2.7 defaulted to a standard 'astronaut riding horse' trope, entirely ignoring the specific positional instruction.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
Grok Imagine Image
- + Excellent preservation of the source person face and upper hair detail.
- + The outfit is intricate and fits the royal theme.
- + Maintains the pose and background from Image 1 perfectly.
- − Completely failed to use the clothing from Image 2 (a modern pea coat and scarf).
- − The hand and rings look somewhat distorted and unnatural.
Wan 2.7
- + Attempts to follow the pose and maintain the background from Image 1.
- + Includes various jewelry and layers as requested.
- + Preserves the face and vitiligo markings accurately.
- − Failed to use the correct outfit from Image 2 (a modern pea coat and scarf).
- − The hands and fingers are significantly distorted with extra digits/anatomical errors.
Verdict: Both models failed the specific 'Image 2' constraint, completely ignoring the modern blue pea coat and plaid scarf in favor of generic 'elaborate' historical/fantasy costumes. Grok Imagine produced a much cleaner and higher quality image overall, whereas Wan 2.7 had severe anatomical issues with the hands and feet.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
Grok Imagine Image
- + Excellent adherence to the 'back seat' positioning for the passenger
- + Realistic lighting and photorealistic textures for both the capybara and the businesswoman
- + Accurate depiction of a bored, normal expression on the human
- − The passenger is technically in the passenger seat rather than the back seat, based on car proportions shown
Wan 2.7
- + High level of detail in the taxi driver cap and the capybara's fur
- + Good interpretation of the 'bored' expression on the passenger
- − The passenger is clearly in the front passenger seat which contradicts the 'back seat' prompt
- − The perspective and composition are slightly more cramped and less cinematic
- − The passenger's hands and phone interaction look less natural
Verdict: Grok Imagine followed the prompt more effectively by placing the passenger further back and maintaining a professional, cinematic composition that felt more like a New York scene. Wan 2.7 failed to put the passenger in the back seat, placing her directly next to the driver instead, and had less convincing human-to-object interaction.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
Grok Imagine Image
- + Excellent text rendering with no spelling errors.
- + The moody, cinematic lighting perfectly captures the requested spooky atmosphere.
- + Strong adherence to the 'dark parchment' and thorn border request.
- − The parchment edges are a bit generic compared to the intricate illustrations of the competitor.
Wan 2.7
- + Highly detailed, illustrative border with extra thematic elements like ravens and skulls.
- + Perfect text layout and accuracy across all requested fields.
- + Beautiful composition that balances the central art with the typography.
- − The overall aesthetic leans more towards a clean digital illustration than a 'vintage dark parchment.'
- − The lighting is less 'cinematic' and more evenly lit, reducing the mystery.
Verdict: Both models followed the prompt exceptionally well, producing accurate text and all requested design elements. Grok Imagine creates a more atmospheric and moody vintage poster that feels authentic to a gothic theme, while Wan 2.7 delivers a cleaner, highly detailed, and professional layout that looks like a high-end modern greeting card.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
Grok Imagine Image
- + Excellent preservation of the original image, changing only the hair area
- + The hair texture and color match the existing beard perfectly
- + Highly realistic and natural hairline
- − The hair is a bit flat and thin compared to the 'full and thick' request
Wan 2.7
- + Better adherence to the 'full, thick' part of the prompt
- + Keeps the facial features and clothing accurately preserved
- + Natural-looking hair flow and volume
- − Minor distortion where the hair meets the temple and glasses
- − The hair color is slightly darker than the existing salt-and-pepper beard
Verdict: Grok Imagine Image provides a more seamless and realistic technical edit, perfectly matching the hair to the beard, but it is a bit conservative regarding the 'thick' part of the prompt. Wan 2.7 provides more volume as requested, but has slight blending issues near the glasses. Grok Imagine Image is the preferred winner for its superior realism and flawless preservation of the original's lighting and texture.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
Grok Imagine Image
- + Excellent text rendering and alignment with the flag icon
- + Rich, saturated colors and appealing lighting
- + High-quality textures on the rice and wooden base
- − The sushi and bowl are slightly oversized for the base
Wan 2.7
- + Perfectly executed isometric perspective and diorama base
- + Clean, minimalist composition with realistic PBR material effects
- + Accurate inclusion of small garnishes like ginger and wasabi as requested
- − Text styling is slightly less centered and elegant than Image A
Verdict: Both models followed the prompt exceptionally well. Grok Imagine (Image A) produces a more vibrant, graphic-design-focused image with superior typography, while Wan 2.7 (Image B) captures the 'miniature diorama' and 'isometric' technical requirements more accurately with its scale and layout. Grok Imagine is slightly ahead due to the overall visual appeal and professional placement of the heading and flag icon.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
Grok Imagine Image
- + Excellent caricature style with a classic 'big head' exaggeration.
- + Clever integration of all elements including a dog in skates and a hockey-themed news studio.
- + Extremely high likeness preservation and very clean text rendering.
- − The hands on the desk are slightly small even for caricature proportions.
Wan 2.7
- + Creative use of speech bubbles and multiple dogs with diverse hockey accessories.
- + Maintains the 'selfie' pose from the original source image.
- + Good energy and chaotic humor that fits the prompt.
- − Likeness to the original woman is less accurate than the competitor.
- − Contains minor text errors like 'Dogs Rolee!' instead of 'Rule'.
- − The composition feels a bit cluttered with overlapping elements.
Verdict: Grok Imagine Image is the winner for its superior execution of the 'caricature' style, which perfectly balances the exaggerated features while keeping a clear likeness to the woman in the source image. Wan 2.7 has a more energetic and busy composition, but it suffers from a typo and loses more of the facial essence of the original person compared to Grok.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
Grok Imagine Image
- + Strong vibrant colors and joyful atmosphere
- + Includes all requested animal types in a dynamic composition
- + Excellent lighting effects with clear god rays and bokeh
- − Stylized, 'cute' appearance fails the 'hyper-photorealistic' requirement
- − Missing butterflies which were a key part of the prompt
- − Anatomical oddities like the kitten's glowing blue eyes and floating positions
Wan 2.7
- + Achieves a high level of photorealism in fur textures and environments
- + Accurately includes butterflies as requested
- + Consistent anatomical proportions and realistic interactions between subjects
- − The fox looks slightly older than a 'baby kit'
- − The kitten's pose is a bit stiff compared to the others
Verdict: Wan 2.7 is the clear winner as it successfully follows the 'photorealistic' instruction and includes all prompt elements like butterflies, whereas Grok Imagine produced a stylized, cartoonish illustration. Wan 2.7 also handles the complex requirement of four different animals with much better anatomical accuracy and realistic lighting.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
Grok Imagine Image
- + Captures the Studio Ghibli art style perfectly with clean line work and distinct character features.
- + Preserves the composition and poses of the original meme accurately.
- + Includes thematic background elements like fluffy clouds and a watercolor wash effect.
- − The woman in the red dress looks significantly different from the original person compared to Model B.
Wan 2.7
- + Maintains a high level of facial similarity to the people in the original source image.
- + Features beautiful watercolor textures that feel like hand-painted concept art.
- + Successfully applies a soft, nostalgic color palette.
- − The art style is more generic watercolor illustration and less specifically 'Studio Ghibli' than model A.
Verdict: Both models did an excellent job transforming the 'Distracted Boyfriend' meme into an illustration. Grok Imagine Image (Model A) is the winner for better adhering to the specific 'Studio Ghibli' aesthetic, featuring the iconic line art and eye styles associated with the studio, whereas Wan 2.7 (Model B) feels like a standard watercolor filter, even though it preserved faces more accurately.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
Grok Imagine Image
- + Excellent adherence to the 'windy' hair request.
- + Adds a large quantity of leaves providing a high-energy feel.
- + Updates the dog's ears to reflect the motion.
- − The orange leaves look like stock stickers and don't match the scene's lighting well.
- − Face details slightly changed/softened compared to the original.
- − The leaf distribution is a bit cluttered and overlaps the subject poorly.
Wan 2.7
- + Subtle and realistic blowing hair effect.
- + Leaves are integrated realistically with motion blur and better lighting.
- + Maintains very high fidelity to the original person's facial features.
- − Fewer leaves make the scene feel less 'energetic' than Image A.
- − The motion effect on the hair is more restrained.
Verdict: Grok Imagine Image followed the 'dynamic and lively' prompt more aggressively, adding significant motion to the hair and dog's ears, though the leaves appear artificial. Wan 2.1 provided a much more realistic and aesthetically pleasing edit that looks like a real photograph, though it is more conservative with the motion effects. Wan 2.1 is the winner for its superior blend of realistic editing and source preservation.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
Grok Imagine Image
- + Excellent typography including the requested accent on 'Caffè'
- + Strong minimalist vector aesthetic
- + Excellent inclusion of professional-looking steam and banner
- − Redundant 'Est. 1720' text appearing twice
- − Slightly confusing additional handle/spoon shapes coming out of the cloche
Wan 2.7
- + Clean circular emblem composition
- + Good use of the 'Est. 1720' banner
- + High contrast and symmetrical layout
- − Misspelled the name as 'Florion' instead of 'Florian'
- − The cloche is transparent rather than a solid retro style
- − Steam is very faint and thin
Verdict: Grok Imagine followed the typography requirements much better than Wan 2.7, correctly spelling the name and including the accent. While Grok Imagine included redundant text, its vector style and interpretation of a minimalist logo are more polished and professional than the output from Wan 2.7, which suffered from a spelling error and a glass-like cloche that didn't fit the 'retro' descriptor as well.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
Grok Imagine Image
- + Excellent adherence to the 'flat-vector' style with bold, clean icons
- + High readability for main step headers
- + Accurate representation of the 3-member crew and the NASA logo
- − Significant spelling errors in secondary text (e.g., '3rajcoory')
- − The layout feels a bit crowded with large, disjointed icons
Wan 2.7
- + Superb vertical infographic composition that feels like a professional poster
- + Includes many accurate technical details like mission dates and velocities
- + Consistent iconography that follows the requested NASA color palette perfectly
- − Step 5 header contains a typo ('DESCRIPT' instead of Descent)
- − Text is quite small, making it difficult to read without zooming
Verdict: Grok Imagine Image creates a visually striking set of icons with very clear main headers, but fails on internal spelling and logical flow. Wan 2.7 produces a much more sophisticated infographic layout that successfully incorporates technical details and a clean vertical flow, making it feel like a genuine educational poster. Despite a small typo in one header, Wan 2.7 is the superior choice for its professional composition and adherence to the 'modern infographic' prompt.
Explore each model
Alibaba's Wan 2.7 image generation and editing model for text-to-image, reference-guided generation, and instruction-based image edits