Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [max] Black Forest Labs Grok Imagine Image xAI

Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [max]

23.7 arena score

#23 of 62 in Text-to-Image

Skill signature · Text-to-Image

Grok Imagine Image

23.4 arena score

#26 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [max]

0%

win rate

Ties

0%

Grok Imagine Image

0%

win rate

Shared challenges 17

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent depiction of window light and realistic caustic reflections on the table.
  • + High level of photographic realism and texture detail on the wooden table and sphere.
  • + Accurate spatial relationship between the plant behind the cube and its visibility through the glass.
  • The glass container has an open top rather than being a solid cube.
  • Text on the book spine is slightly garbled/nonsensical.

Grok Imagine Image

  • + Successfully placed a plant behind the cube visible through the glass.
  • + The blue sphere has a smooth, polished texture that contrasts well with the glass.
  • + Good adherence to the basic prompt requirements including lighting direction.
  • The blue sphere is inexplicably floating in the center of the cube.
  • The perspective of the cube is slightly warped, appearing more like a rectangular prism than a perfect cube.

Verdict: FLUX.1 Kontext [max] produces a significantly more realistic image with superior lighting, shadows, and textures, though the 'cube' is an open-top glass container. Grok Imagine follows the prompt's object placement well, but the blue sphere appears to be levitating, which breaks the physical realism of the scene. FLUX.1 Kontext [max] is the preferred choice for its photographic quality and convincing material rendering.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent depiction of rain with falling droplets and realistic wet pavement reflections
  • + Highly detailed and realistic hand texture and bicycle components
  • + Effective shallow depth of field that maintains focus on the subject and the red bike
  • The man appears more Caucasian or generic than specifically Japanese
  • The rain effects are a bit heavy-handed, looking like a uniform overlay in some areas

Grok Imagine Image

  • + Successfully captures the 'motion blur from passing cars' request
  • + Subject feels more authentically seated in a Japanese urban environment
  • + Captures an 'imperfect framing' candid feel very well
  • The red bicycle construction is physically incoherent in several places
  • The man's hands are very poorly defined and blurry
  • Minimal visible rain compared to the prompt's focus

Verdict: FLUX.1 Kontext [max] produces a much higher quality image with superior anatomical and mechanical detail, though it misses the motion blur aspect. Grok Imagine Image captures the 'candid street photo' vibe and specific motion blur requested, but suffers from significant distortions in the bicycle's geometry and the man's hands. FLUX.1 Kontext [max] is preferred for its technical clarity and realistic textures.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent skin texture with realistic pores and sweat details
  • + Superior material rendering on the engraved plate armor
  • + Complex lighting that feels integrated into the environment
  • Missed the request for beads in the hair braids
  • The hair and beard look slightly too wiry or sharp

Grok Imagine Image

  • + Perfectly included the specific detail of beads in the braids
  • + Beautiful cinematic composition with actual torches visible
  • + Highly detailed engraving on the chest piece
  • Armor texture looks slightly 'CG' or flat compared to the realism of Model A
  • Scars look more like surface-level paint than actual skin damage

Verdict: While Grok Imagine Image captured specific details like the hair beads and torches more accurately, FLUX.1 Kontext [max] produced a more lifelike image with superior skin textures and realistic battle-worn details. Model A's metallic reflections and facial rendering feel more tangible, whereas Model B's armor has a slightly repetitive, digital pattern feel.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Clean layout with consistent image framing
  • + Good adherence to the white background and minimalist aesthetic
  • + Vibrant and professional food photography
  • Text is mostly gibberish despite the header being clear
  • Failed to properly implement the requested categories (appetizers/pizza/mains) in a distinct way

Grok Imagine Image

  • + Perfect adherence to specific sections for appetizers, pizza, and mains
  • + Readability of main headers and item names is high
  • + Dynamic layout with high-quality food photography
  • Repetitive menu items (multiple instances of Steak Frites and Grilled Salmon)
  • Small body text is illegible

Verdict: Grok Imagine Image followed the prompt requirements much more accurately by including the specific requested sections (Appetizers, Pizza, Mains). While FLUX.1 Kontext [max] produced a very clean and professional-looking aesthetic, it failed the structural requirements of the prompt and filled the menu almost exclusively with pizza.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealistic texture on the burger bun and patty.
  • + Strong, cinematic lighting that integrates well with the dark, fiery background.
  • + Clear and legible typography for all text elements.
  • The burger is not truly 'exploded' as requested, remaining largely intact as a stack.
  • Failed to include the required starburst element for the price.
  • The price uses a comma instead of a decimal point.

Grok Imagine Image

  • + Successfully captured the 'exploded' aspect with components floating separately.
  • + Included all requested elements, including the starburst and fiery glowing text effects.
  • + Great sense of motion with liquid splashes and flying ingredients.
  • The starburst graphic looks like a flat clip-art element that clashes with the photorealistic style.
  • The lettuce and sauce droplets look slightly more digital/artificial than Image A.

Verdict: While FLUX.1 Kontext [max] produces a more realistic and professional-looking food photograph, it fails several key prompt instructions regarding the 'exploded' layout and the starburst. Grok Imagine Image followed every part of the prompt, including the specific layout of ingredients and the starburst element, making it the better adherence to the creative brief.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent text legibility and spelling accuracy
  • + Realistic chalk smudges and texture on the board surface
  • + Accurately completes the truncated 'Brown But...' prompt to a logical conclusion
  • The handwriting style is more of a neat print than the requested 'elegant cursive' for the title
  • The letter 'A' in the title lacks consistency with the rest of the handwriting style

Grok Imagine Image

  • + Beautiful chalk texture with realistic grain and variation in pressure
  • + Captured a more cursive-leaning style for the title as requested
  • + Excellent composition with cozy café lighting and atmosphere
  • The board shows some digit-like anomalies in the '2026' date
  • Slightly less crisp text edges compared to Model A

Verdict: Both models performed exceptionally well on this complex text rendering task. FLUX.1 Kontext [max] produced very clean, readable text and handled the truncated prompt intelligently by finishing the cookie description, whereas Grok Imagine Image captured a more authentic chalk aesthetic and atmospheric lighting. Grok is the slight winner for better adhering to the 'cursive' instruction for the title while maintaining a very high level of realism.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photographic texture on the horse's coat and astronaut's suit.
  • + Realistic lighting and cinematic composition with subtle lens flares.
  • + Clean, high-fidelity details in the horse's mane and space background.
  • Failed the primary logic constraint by placing the astronaut on top of the horse.

Grok Imagine Image

  • + Successfully interpreted the difficult prompt logic by placing the horse on top of the astronaut.
  • + Vibrant and surreal color palette with a colorful nebula.
  • + Great dynamic posing of both the horse and the astronaut.
  • Anatomical issues where the horse's front leg appears to be fused with the astronaut's hand.
  • Lower overall realism compared to the competing model.

Verdict: While FLUX.1 Kontext [max] produced a much higher quality image in terms of texture and lighting, it completely ignored the specific instruction for the horse to be on top. Grok Imagine followed the complex spatial instruction perfectly, though it suffered from minor anatomical merging artifacts and a less realistic finish. Grok Imagine is the winner for successfully executing the core surreal concept that FLUX.1 failed to grasp.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealistic fur texture and lighting
  • + Captures the 'professional expression' of the capybara well
  • + More realistic depth of field with the background blur
  • The passenger is holding her phone to her ear like a call, rather than looking at it as requested
  • The woman is sitting in the front passenger seat instead of the back seat

Grok Imagine Image

  • + Perfectly follows the spatial instruction of the passenger in the back seat
  • + Accurately depicts the businesswoman looking at her phone with a bored expression
  • + The composition captures the full interior atmosphere requested
  • The capybara's paws are clipping through the steering wheel
  • The steering wheel is positioned on the wrong side for a New York taxi

Verdict: While FLUX.1 Kontext [max] has superior texture and lighting, it fails several specific prompt instructions regarding the position and action of the passenger. Grok Imagine Image followed the prompt much more accurately, correctly placing the businesswoman in the back seat and showing her looking at her phone, despite minor anatomical clipping on the steering wheel.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent integration of the text into the overall graphic design
  • + Highly polished cinematic lighting on the jack-o-lantern
  • + Beautifully intricate web and thorn border that fits the square format perfectly
  • Redundant text at the bottom with the location listed twice
  • The Date includes commas instead of dots as requested

Grok Imagine Image

  • + Captures all requested elements including the moon and bats with high clarity
  • + Perfect adherence to the requested event details and date format
  • + Very clear scroll banner for the secondary text
  • The 'parchment' effect feels a bit like a cheap clip-art overlay on a black background
  • The font for the bottom details is very plain and lacks the gothic elegance requested

Verdict: Both models followed the prompt closely, but FLUX.1 Kontext [max] produced a more cohesive and professional-looking graphic design with superior lighting. While Grok Imagine followed the specific text instructions and date formatting better, its overall composition feels less like a polished invitation and more like a collection of separate assets.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.1 Kontext [max]
Before After
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent addition of thick, dense hair.
  • + Strong adherence to the 'full head of hair' instruction.
  • + Maintains the texture and color of the original beard.
  • Significantly alters the shape and features of the face, making the person look like a different individual.
  • The hairline looks slightly artificial and overly straight.

Grok Imagine Image

  • + Perfectly preserves the subject's identity and facial features.
  • + The hair texture and hairline look extremely natural and integrated into the existing image.
  • + Maintains the original lighting and background exactly.
  • The hair is somewhat thin and receding, which slightly misses the 'full, thick head' instruction.

Verdict: FLUX.1 Kontext [max] succeeded in adding a very thick head of hair, but significantly altered the man's facial features and bone structure, losing the identity of the original subject. Grok Imagine Image provided a much more subtle and realistic edit that perfectly preserved the original person's face and the environmental context, though the hair density is not quite as 'thick' as requested.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with a playful, bubbly 3D font.
  • + High-quality material rendering on the salmon and rice textures.
  • + Clean and professional isometric composition.
  • Completely missed the requested flag icon.
  • The chopsticks are merged together into a single blocky shape.

Grok Imagine Image

  • + Successfully included all requested elements, including the flag icon.
  • + Higher variety of sushi types on the diorama base.
  • + Extremely sharp, clean rendering of the 3D shapes.
  • Text is plain 2D rather than matching the 3D aesthetic of the scene.
  • The perspective on the 'JAPAN' text is slightly flat compared to the isometric base.

Verdict: Grok Imagine Image followed the complex prompt more accurately by including the requested flag icon and a wider variety of sushi. While FLUX.1 Kontext [max] had superior typography and material textures for the salmon, its failure to include all requested elements and the poorly rendered chopsticks make it the runner-up.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Transferred the blue denim shirt from the original image well.
  • + Incorporated a microphone and a news desk effectively.
  • + Cartoon style captures a playful, caricature aesthetic.
  • The hockey stick is barely visible in the crop at the top right.
  • Added glasses which were not in the original photo.
  • The caricature face does not strongly resemble the source subject.

Grok Imagine Image

  • + Maintained a much stronger facial resemblance to the original woman.
  • + Incorporates the hockey theme extensively with a stadium background and skates.
  • + Professional presentation of a TV news set with high-quality text.
  • Changed the clothing from the source image to a formal blazer.
  • The hands are very small and lacking detail compared to the head.
  • The floating pucks in the background look slightly cluttered.

Verdict: Grok Imagine Image is the clear winner for its superior ability to maintain the subject's likeness while masterfully blending the news anchor profession with the requested hockey and dog themes. While FLUX.1 Kontext [max] correctly preserved the denim shirt, its caricature face lost the identity of the user and the hockey elements were poorly integrated into the composition.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent anatomical realism for all four animals.
  • + Beautiful lighting with naturalistic god rays and golden hour atmosphere.
  • + Precise adherence to the butterfly and wildflower meadow environment.
  • The animals are sitting rather than actively 'chasing and tumbling' as requested.
  • The rabbit's face is slightly stylized compared to the very realistic pup and fox.

Grok Imagine Image

  • + Dynamic composition that captures the 'chasing and tumbling' action well.
  • + Vibrant color palette with clear dew sparkles in the foreground.
  • The fur texture appears overly digital and 'brushed' rather than photorealistic.
  • Missing the butterflies mentioned in the prompt.
  • The fox kit has unnatural black paws that look more like a red panda.

Verdict: FLUX.1 Kontext [max] delivers a much more photorealistic and technically superior image with natural lighting and accurate animal anatomy. While Grok Imagine captures the requested motion better, it fails to include the butterflies and has a distinctly CG, 'plastic' look to the fur compared to FLUX.1's masterpiece quality.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent capture of Studio Ghibli character design, particularly in the eye shapes and soft shading.
  • + Maintains high fidelity to the source image's composition and poses.
  • + Hand-painted texture is very convincing with watercolor-like bleed effects.
  • The faces of the two women look very similar to each other, losing their distinct features from the source.

Grok Imagine Image

  • + Successfully translates the background into a dreamy, painterly town reminiscent of 'Kiki's Delivery Service'.
  • + Maintains better facial distinction between the three characters.
  • + Colors are vibrant and evoke a sunny, nostalgic mood.
  • The character art style is a bit more generic anime/webtoon than specifically Studio Ghibli style.
  • The man's stubble is depicted in a way that feels slightly out of place for the requested aesthetic.

Verdict: FLUX.1 Kontext [max] provides a much more authentic Studio Ghibli aesthetic, particularly in the line work, eye design, and soft watercolor textures requested in the prompt. While Grok Imagine Image creates a beautiful background and maintains the likeness of the original people slightly better, it misses the specific 'soul' of the Ghibli art style which the first model captures perfectly.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.1 Kontext [max]
Before After
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Successfully added wind effects to the hair while keeping it looking natural.
  • + Preserved the original posture and facial features of the woman almost perfectly.
  • + Included subtle leaves that feel integrated into the physics of the scene.
  • The leash loop in the hand is slightly distorted compared to the original.
  • The motion feel is relatively subtle, perhaps leaning less into 'energetic' than 'breezy'.

Grok Imagine Image

  • + Strongly emphasized the request for leaves, creating a very dynamic visual.
  • + The hair motion is very well executed and clearly matches the wind direction of the leaves.
  • + Preserved the dog's appearance and the girl's face very well.
  • The orange autumn leaves clash slightly with the vivid green summer trees in the background.
  • Some leaves appear to be pasted on top of the woman's clothes without natural shadows.

Verdict: Both models performed excellently at this edit task, maintaining the integrity of the subjects while adding the requested motion. FLUX.1 Kontext [max] is more subtle and realistic, while Grok Imagine interprets 'energetic and lively' more literally by flooding the scene with wind-blown leaves, which results in a more dramatic change despite the seasonal inconsistency of orange leaves against green trees.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography adherence including the specific grave accent on the letter È.
  • + Traditional woodblock/linocut texture that fits the vintage aesthetic perfectly.
  • + Clean, balanced composition with a classic banner element.
  • The steam element is very small and lacks prominence compared to the rest of the logo.

Grok Imagine Image

  • + Stronger visual representation of steam and a more modern vector finish.
  • + Good use of color blocking and shadows to define the cloche.
  • Redundant text featuring 'Est. 1720' twice in two different locations.
  • Incorrect accent mark on the 'è' in Caffè.
  • The addition of a spoon and cup handle makes the icon cluttered and stray from the minimalist request.

Verdict: FLUX.1 Kontext [max] produced a superior, professional-grade logo that perfectly captures the vintage texture and correct Italian orthography requested. While Grok Imagine Image created a clean vector, it suffered from repetitive text and incorrect accent marks, making it less effective as a brand mark.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [max]
Grok Imagine Image

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography rendering for main text and names.
  • + Professional color palette following the NASA-inspired theme.
  • + Clean, minimalist composition that feels modern.
  • Confused logical flow between steps, with labels not matching the icons properly.
  • Missing formal numbered steps 1-6 as requested in the prompt.
  • The 'Saturn V' icon is a generic stylized rocket rather than specific to the mission.

Grok Imagine Image

  • + Perfectly adhered to the 6-step structure requested in the prompt.
  • + Consistent flat-vector style with charming iconography.
  • + Correctly interpreted specific requests like the NASA logo and crew information.
  • Several spelling errors in the small supporting text (e.g., '3rajoory', 'Moom').
  • Text overlap and clutter in the step 3 area.
  • Saturn V icon looks somewhat toy-like rather than a professional infographic.

Verdict: FLUX.1 Kontext [max] produces a more aesthetically pleasing and 'adult' infographic with superior text rendering, but it fails to follow the 6-step structural layout requested. Grok Imagine succeeds in following the exact logical instructions for all 6 mission steps but suffers from common AI text artifacts and spelling errors. Grok Imagine is the likely winner for better prompt adherence regarding the specific informational structure.

Next steps

Explore each model