Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [flex] Black Forest Labs Grok Imagine Image xAI

Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.

FLUX.2 [flex]

24.8 arena score

#14 of 62 in Text-to-Image

Skill signature · Text-to-Image

Grok Imagine Image

23.4 arena score

#26 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [flex]

80.0%

win rate

Ties

0.0%

Grok Imagine Image

20.0%

win rate

80.0% 0.0% ties 20.0%
Shared challenges 19

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Perfectly adheres to all spatial prompts including object placement.
  • + Excellent rendering of light and soft shadows.
  • + Clean, modern aesthetic with high clarity.
  • The sphere is quite large, pushing the definition of 'small' in the prompt.

Grok Imagine Image

  • + Captures the 'small' aspect of the blue sphere more accurately.
  • + Highly realistic wood texture and light scattering in the glass.
  • + Natural-looking plant and depth of field.
  • The glass object is a rectangular prism rather than a cube.
  • The sphere appears to be floating unnaturally in the center without support.

Verdict: Both models followed the complex spatial instructions well. FLUX.2 [flex] produced a better 'cube' and superior lighting, while Grok Imagine Image provided a more realistic texture on the wooden table and better followed the 'small' descriptor for the sphere. FLUX.2 [flex] is the winner for better geometric accuracy regarding the cube and more cohesive composition.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [flex]
Grok Imagine Image
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of the specific man's identity, including his hairstyle and clothing (visible plaid scarf).
  • + High accuracy in preserving the specific car model (Rolls-Royce Phantom Drophead Coupé) from the source image.
  • + Realistic motion blur on the road and wheels that enhances the sense of driving.
  • The man's scale and positioning in the driver's seat feel slightly off, appearing a bit small for the car.
  • The steering wheel placement is slightly detached from the driver's grip.

Grok Imagine Image

  • + Beautifully rendered California coastline background with great atmospheric perspective.
  • + The composition of the car on the road feels very dynamic and professional.
  • Fails to use the man from the source image, replacing him with a generic older white man.
  • Changes the car model slightly, evolving the classic Phantom front into a more modern Rolls-Royce Dawn style.
  • Does not follow the multi-image prompt requirement to combine both source images.

Verdict: FLUX.2 [flex] is the clear winner because it successfully followed the core instruction: combining the specific man and the specific car from the source images into the requested new setting. While Grok Imagine Image produced a high-quality visual, it completely ignored the subject's identity, replacing him with a random character, which defeats the purpose of an image editing task.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to technical photography prompts like 50mm shallow depth of field.
  • + Detailed skin textures and realistic rain/wet pavement interaction.
  • + Dynamic composition with effective use of bokeh and light.
  • The bicycle frame geometry is slightly warped/nonsensical near the crank set.
  • The scale of the bicycle seems a bit small relative to the man.

Grok Imagine Image

  • + Achieves a highly authentic 'street photography' look with realistic motion blur from cars.
  • + The bicycle design is more structurally plausible for a real-world bike.
  • + Captures the 'imperfect framing' prompt well with its candid feel.
  • The subject's face is obscured and partially covered by a mask, losing the 'elderly Japanese man' facial detail requested.
  • Overall lighting is a bit flatter and less 'cinematic' than technically requested.

Verdict: FLUX.2 [flex] produced a more visually striking and detailed image that followed the technical lighting and texture prompts more closely, though the bicycle's anatomy is slightly glitched. Grok Imagine Image captured the candid 'street photo' vibe and the motion blur of passing cars much more realistically, but it failed to showcase the facial details of the subject. FLUX.2 [flex] is the winner for its superior clarity and beautiful rendering of light and rain.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent depiction of leather and textile textures
  • + High detail in the ornate engravings on the plate armor
  • + More distinct and colorful beads in the hair braids
  • The scars look a bit painted on rather than deep or realistic
  • Spatial relationship of the torch in the background is a bit flat

Grok Imagine Image

  • + Superior cinematic lighting and atmosphere
  • + More realistic skin texture and integration of scars
  • + Exceptional hair detail with realistic flyaways and lighting
  • Armor engravings are slightly more repetitive and less 'hand-crafted' in appearance
  • Textile and leather textures are slightly less defined than in Model A

Verdict: Both models followed the prompt exceptionally well, but Grok Imagine Image edges ahead due to its superior cinematic lighting and more believable rendering of the character's facial features and hair. FLUX.2 [flex] produced sharper textures on the armor and straps, but the overall composition feels slightly more digital and less natural than Grok's output.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [flex]
Grok Imagine Image
0% wins 0% ties 100% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Strict adherence to the 3x2 grid layout for food photos.
  • + Clean, professional typography that is highly legible.
  • + Excellent use of color-coded headers for organization.
  • Text content is mostly gibberish.
  • Limited variety in food images, with some repetition.

Grok Imagine Image

  • + High accuracy in text rendering, including recognizable dish names like 'Bruschetta' and 'Margherita'.
  • + Dynamic and visually appealing layout with organic placement of food photos.
  • + Impressive variety in the types of food depicted.
  • Failed to follow the 'grid' requirement for photo placement.
  • Contains several duplicate item entries (e.g., multiple 'Steak Frites' and 'Grilled Salmon').

Verdict: FLUX.2 [flex] adhered much better to the specific layout requirements, providing a clean grid and clear sectioning, though the text is nonsensical. Grok Imagine Image produced much more legible and accurate text for a menu, but failed to follow the grid layout instruction and had several repetitive list entries.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent photorealistic texture on the patty and bun.
  • + Clean, professional typography with a polished fiery effect.
  • + Realistic condensation and sauce drips enhance food appeal.
  • The 'exploded' effect is less dynamic than requested, with components mostly stacked.

Grok Imagine Image

  • + Highly dynamic 'exploded' composition with components flying outwards.
  • + Strong sense of motion with sauce splashes and scattered toppings.
  • + Vibrant colors and high-energy background flames.
  • The lettuce and tomato look slightly more illustrative/artificial compared to the burger patty.
  • The text effect is a bit flat compared to the lighting in Model A.

Verdict: Both models followed the prompt exceptionally well. FLUX.2 [flex] produced a more photorealistic, appetizing image with superior text rendering, while Grok Imagine Image captured the 'exploded' and 'dynamic' aspects of the prompt more effectively with its energetic composition. FLUX.2 [flex] is the likely winner for its professional advertising aesthetic and material realism.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent chalk texture with realistic smudging and dusting on the board.
  • + Flawless text rendering exactly matching the lengthy prompt requirements.
  • + Beautifully balanced composition with professional lighting.
  • The font style feels slightly more like a digital 'script' font than natural human handwriting.

Grok Imagine Image

  • + Authentic handworded feel with natural variations in letter size and baseline.
  • + Superior chalk texture on the individual letters, showing grainy strokes.
  • + Included the requested 'Brown But...' completion logically with 'Brown Butter Chocolate Chip Cookies'.
  • The perspective on the bottom line of text is slightly warped compared to the board angle.
  • Minor artifacting in the very small dots/specks of chalk on the board.

Verdict: Both models performed exceptionally well on this complex text rendering task. FLUX.2 [flex] produced a cleaner, more aesthetically polished image with impressive chalk smudges, whereas Grok Imagine Image captured a more authentic, slightly imperfect handwritten feel that better matches the 'chalkboard artist' look. Grok is the narrow winner for the more realistic texture of the chalk strokes on the wood-grain frame background.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Successfully integrated character features including sunglasses, scarf, and facial hair
  • + Accurately matched the yellow background and red ottoman from the pose reference
  • + Followed the general arm orientation and crouching nature of the source pose
  • The pose is a simplified version rather than an exact recreation of the complex leg positioning
  • Significant anatomical issues with the length of the right arm and shape of the hands
  • Lighting on the character is much harsher and more saturated than the source imagery

Grok Imagine Image

  • + Maintained the original image 1 almost perfectly
  • Completely failed to incorporate any elements from the character reference (Image 2)
  • Did not perform the requested edit, essentially returning the source image with minor variations
  • Failed to change the person's gender, face, or clothing

Verdict: FLUX.2 [flex] successfully attempted the complex task of merging the identity of the person in Image 2 with the scene in Image 1, resulting in a recognizable transformation despite some anatomical distortion. Grok Imagine Image failed the prompt entirely, simply reproducing the first image without making any of the requested character or clothing changes. FLUX.2 [flex] is the clear winner for actually following the editing instructions.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Perfectly adheres to the complex 'horse on top' request with the horse actively riding the astronaut's shoulders.
  • + Exceptional detail in the馬's musculature and the astronaut's suit textures.
  • + Strong cinematic composition with a clear focal point and a dynamic sense of depth through the asteroid field.
  • The astronaut's hands and the horse's front legs exhibit some spatial clipping/merging artifacts.

Grok Imagine Image

  • + Solid interpretation of the 'horse on top' instruction.
  • + Beautiful vibrant colors and nebula effects in the background.
  • + Good rendering of the horse's coat and lighting.
  • The composition is less grounded; the horse appears to be floating above rather than 'riding' the astronaut.
  • Anatomical anatomy of the horse's lower front legs is slightly warped and lacks clear contact points.

Verdict: FLUX.2 [flex] is the clear winner as it interpreted the 'riding' instruction literally and creatively by placing the horse on the astronaut's shoulders. Grok Imagine followed the spatial instruction of keeping the horse on top, but the result looks more like two separate objects floating near each other rather than a cohesive, surreal action scene.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the 'inside a yellow taxi' requirement while showing the capybara clearly
  • + Stronger photorealistic lighting and textures on the capybara's fur
  • + Accurate depiction of a professional-style NYC taxi driver hat
  • The capybara's hands look more like human hands with hair than capybara paws
  • Diagonal composition makes the car geometry feel slightly warped

Grok Imagine Image

  • + Perfect bored expression on the passenger which matches the prompt instructions
  • + Realistic capybara paws with visible claws correctly positioned on the steering wheel
  • + Very logical and symmetrical interior composition
  • The passenger is sitting in the front passenger seat instead of the required 'back seat'
  • The capybara's hat is a casual baseball cap rather than a typical taxi driver cap

Verdict: While FLUX.2 [flex] adhered better to the spatial instruction of having the passenger in the back seat, Grok Imagine Image captured the nuanced 'bored expression' and realistic capybara anatomy much better. FLUX.2 [flex] is more stylistically vibrant, but Grok Imagine Image feels more like a coherent, real-world photograph despite the seating error.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent layout with a clean, readable typographic hierarchy.
  • + The parchment texture and deckled edges look realistic.
  • + High-quality lighting and atmospheric fog effects.
  • The thorn border is less prominent than described in the prompt.

Grok Imagine Image

  • + The thorns and spiderwebs in the border are highly detailed and visually striking.
  • + Dynamic bat silhouettes and more intricate gnarly tree designs.
  • + Good adherence to the gothic aesthetic and banner request.
  • The text placement feels a bit crowded against the bottom border.
  • The lighting on the pumpkin and the moon is slightly less cinematic compared to Model A.

Verdict: Both models followed the prompt exceptionally well, producing high-quality invitations with perfect text rendering. FLUX.2 [flex] offers a more balanced and professional design layout, while Grok Imagine Image excels in the illustrative details of the thorny border and gnarled trees, giving it a more 'spooky' feel.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [flex]
Before After
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of the background and original lighting on the face.
  • + Matches the hair color and texture perfectly to the existing beard.
  • The hair length is a bit short, resembling a buzz cut rather than a full head of hair.
  • The hairline follows the original scalp contour very closely, making it look slightly receded.

Grok Imagine Image

  • + Successfully provides a longer, fuller hairstyle as requested.
  • + Seamlessly blends the new hair with the sideburns and existing facial features.
  • Slightly alters the shape of the forehead/upper skull compared to the original.
  • Minor loss of skin texture detail on the forehead near the new hairline.

Verdict: Both models performed excellently at this image editing task, preserving the person's identity and surrounding environment perfectly. Grok Imagine Image is the winner as it provided a much 'fuller' head of hair as requested, whereas FLUX.2 [flex] provided more of a stubble/buzz-cut texture.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [flex]
Grok Imagine Image
50% wins 0% ties 50% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Clean, professional typography and layout.
  • + Excellent miniature 3D aesthetic with soft, clay-like textures.
  • + Accurate isometric perspective and centered composition.
  • The flag is placed below the text, whereas the prompt suggested text at top-center and sushi below it (though the layout is still very pleasing).

Grok Imagine Image

  • + Good adherence to the isometric diorama style.
  • + Correct inclusion of all requested elements (flag, text, sushi).
  • + High visual clarity with sharp shadows.
  • The text rendering is slightly less refined than Model A.
  • Lighting is a bit harsh compared to the 'gentle lighting' requested.

Verdict: Both models followed the prompt exceptionally well, capturing the isometric miniature style. FLUX.2 [flex] produced a more aesthetically pleasing image with superior 'soft refined textures' and better typography, while Grok Imagine Image provided a more complex sushi plate that also accurately followed the design requirements. FLUX.2 [flex] is the winner for its more professional, clean, and cohesive 3D render look.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent caricature style with highly exaggerated, expressive features.
  • + Integrates multiple dogs wearing hockey jerseys to combine both hobbies.
  • + Strong storytelling through the background monitors Showing hockey and news content.
  • The facial likeness to the source image is slightly lost in the cartoonish style.
  • Minor garbled text on the news desk graphics.

Grok Imagine Image

  • + Maintains a much better facial likeness to the original woman while still being a caricature.
  • + Clear, legible text on the news desk.
  • + Cleverly includes a dog in hockey skates and helmet to hit all prompt requirements.
  • The hockey pucks in the background appear to be floating randomly rather than being integrated into the scene.
  • The hands are a bit small and awkward in proportion even for a caricature.

Verdict: Both models followed the instructions well, but Grok Imagine excels in maintaining the subject's facial identity within the caricature style. While FLUX.2 [flex] created a more vibrant 'cartoon' world with a higher volume of dogs, Grok Imagine's inclusion of a dog in skates was a more humorous interpretation of the prompt's specific hobby combination.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [flex]
Grok Imagine Image
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent adherence to the 'chasing butterflies' and 'tumbling' motion aspects of the prompt.
  • + Highly realistic fur textures and anatomical proportions for all four animals.
  • + Beautifully rendered god rays and dew sparkles that feel integrated into the scene.
  • The fox kit has slightly unusual dark legs that look more like black paws than a typical fox kit's markings.

Grok Imagine Image

  • + Warm, vibrant color palette with strong backlighting.
  • + Cute, stylized 'expressive eyes' as requested.
  • The animals are static and posing rather than 'playfully chasing and tumbling' as requested.
  • The butterfly rendering is poor, appearing as small white blobs rather than detailed butterflies.
  • The fur texture looks overly smooth and 'AI-processed' compared to the 8K masterpiece request.

Verdict: FLUX.2 [flex] is the clear winner as it successfully captures the dynamic action of the animals chasing butterflies in a realistic meadow, whereas Grok Imagine produced a static, posed shot. FLUX.2 [flex] also delivered much higher detail in the fur, background elements, and the butterflies themselves, whereas Grok Imagine struggled with the butterfly details and overall realism.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Expertly captures the soft, washed-out watercolor aesthetic characteristic of many Studio Ghibli backgrounds.
  • + Maintains the exact positioning and outfits of the original characters with high fidelity.
  • + Successfully uses soft pastel tones as requested in the prompt.
  • The character on the right has an unfocused, smiling expression that misses the 'jealous/angry' narrative of the original image.
  • The foreground character in the red dress is slightly less detailed compared to the background elements.

Grok Imagine Image

  • + Maintains the correct emotional narrative by keeping the disgruntled expression on the character to the right.
  • + Features a more vibrant and detailed 'Ghibli-esque' town background with clearer architectural elements.
  • + Strong preservation of character likeness and original composition.
  • The colors are a bit more saturated than the 'soft pastel' request suggested.
  • The watercolor texture is slightly less pronounced on the characters themselves compared to the background.

Verdict: Both models performed exceptionally well at translating a famous meme into a specific art style. FLUX.2 [flex] achieved a more accurate 'dreamy pastel' look with authentic paper textures, but Grok Imagine preserved the essential narrative context of the original photo (specifically the jealousy of the girlfriend) much better, making it a more successful overall edit.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [flex]
Before After
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent preservation of the subject's face and clothing from the source image.
  • + Subtle, realistic motion blur on the flying leaves.
  • + The hair blowing animation is natural and well-integrated.
  • The green leaves feel slightly incongruous with the established summer lighting in the background.
  • Failed to add motion to the dog's ears or fur.

Grok Imagine Image

  • + Strong 'energetic' feel with a large volume of flying leaves.
  • + Successfully added wind motion to the dog's ears to match the person's hair.
  • The leaves appear like a flat overlay rather than being integrated into the 3D space of the scene.
  • The subject's face has changed slightly compared to the source image, losing some likeness.
  • The hair edit has jittery edges and lacks realistic flow.

Verdict: FLUX.2 [flex] provides a much higher quality edit by preserving the original subject's appearance perfectly while adding realistic wind effects to her hair. While Grok Imagine Image included more leaves and added motion to the dog's ears, the leaves look like 2D stickers and the model did not maintain the source image's facial features as accurately.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [flex]
Grok Imagine Image

AI Judge Analysis

FLUX.2 [flex]

  • + Perfect adherence to the 'vintage minimalist' and 'vector emblem' descriptors.
  • + Clean, professional typography that captures a classic Italian cafe aesthetic.
  • + Excellent layout balance with the arched text and banner.
  • The texture on the background is very subtle, almost unnoticeable.
  • The steam icons are slightly thin compared to the rest of the stroke weights.

Grok Imagine Image

  • + Good use of color depth and shading within the cloche icon.
  • + Includes a nice paper-like texture on the light background.
  • + Clear, legible text rendering.
  • Redundant text repeating 'Est. 1720' twice in the layout.
  • The cloche icon has nonsensical additions that look like a spoon and a handle merging into the dome.
  • Less 'minimalist' than requested, feeling more like a modern mascot logo.

Verdict: FLUX.2 [flex] is the clear winner as it perfectly captures the 'minimalist' and 'vector emblem' style requested, producing a clean and professional logo. Grok Imagine Image fails on the minimalist aspect and introduces visual incoherence with strange artifacts protruding from the cloche, as well as repeating the establishment date unnecessarily.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [flex]
Grok Imagine Image
100% wins 0% ties 0% wins

AI Judge Analysis

FLUX.2 [flex]

  • + Excellent layout with clean, professional vector aesthetic
  • + High text legibility and mostly correct spelling
  • + Sophisticated iconography that feels like a real educational poster
  • Missed the final 'Landing' step from the specific prompt list
  • The order of steps (reading vertically vs horizontally) is slightly unconventional

Grok Imagine Image

  • + Successfully included all 6 numbered steps plus a crew section
  • + Very accurate adherence to the specific iconography requests for each step
  • + Creative composition that uses the bottom of the frame as the lunar surface
  • Multiple spelling errors in the text (e.g., '3rajoory', 'Transluiory', 'Moom')
  • Icon for the Saturn V looks less like the actual rocket compared to Model A

Verdict: FLUX.2 [flex] produced a much more professional and aesthetically pleasing infographic that looks like a finished product, though it missed the final step of the requested list. Grok Imagine followed the prompt's structural instructions more closely by including all six steps and specific icons, but it suffers from poor text rendering and slightly less refined vector art. FLUX.2 [flex] is the preferred choice for its superior visual quality and clean execution.

Next steps

Explore each model