Black Forest Labs' open-weights image generation model with frontier performance, available for non-commercial local deployment
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
FLUX.2 [dev]
#18 of 62 in Text-to-Image
Grok Imagine Image
#26 of 62 in Text-to-Image
Where the votes landed
FLUX.2 [dev]
71.4%
win rate
Ties
14.3%
Grok Imagine Image
14.3%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent representation of thick glass with realistic internal reflections and corner details.
- + Physical placement of the sphere at the bottom of the cube feels more grounded and realistic.
- + Soft window light is beautifully rendered, creating a natural atmosphere.
- − The plant in the background is very blurry, making it less distinct than requested.
- − The sphere has a slight transparency that might not be expected for a simple 'blue sphere'.
Grok Imagine Image
- + The plant is clearly 'behind the cube' and visible through the glass as requested.
- + The cube edges are sharp and clean, creating a very modern aesthetic.
- + The red book has a nice texture on the cover and realistic page edges.
- − The blue sphere is levitating in the center of the cube, which feels physically unnatural despite the prompt not specifying it should be floating.
- − The reflections on the table surface are slightly disconnected from the objects.
Verdict: FLUX.2 [dev] produces a more photorealistic scene with superior lighting and material physics, particularly in how it handles the weight and reflections of the glass. While Grok Imagine Image followed the spatial instruction for the plant more clearly, the floating blue sphere and slightly thinner glass rendering make it feel less grounded than FLUX.2 [dev].
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent preservation of the subject's clothing and hairstyle
- + Accurate interior details of a Rolls-Royce convertible
- + Dynamic sense of motion with background blur
- − Subject's facial features changed significantly from the source image
Grok Imagine Image
- + Successfully captures the requested California coastline scenery
- + Maintains the exterior appearance of the car well
- − Completely failed to use the man from the provided source image
- − Replaced the subject with a generic older white male
- − Car model changed slightly to a newer style (Wraith/Dawn headlights)
Verdict: FLUX.2 [dev] successfully followed the edit instructions by placing a character matching the source's unique clothing and hair into the specific car provided. While Grok Imagine captured the 'California coastline' aesthetic more broadly, it completely ignored the specific person provided in the source image, making it fail as an image editing task.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent depiction of motion blur in the background cars.
- + Highly detailed and realistic facial features and skin texture.
- + Strong composition that highlights the subject while maintaining the street atmosphere.
- − The red bicycle's frame and handlebars are structurally confusing and messy.
- − The rain effect is very subtle, making the pavement look more like it's already wet rather than currently raining.
Grok Imagine Image
- + Captured the 'imperfect framing' prompt well with a more candid, snapshot aesthetic.
- + The bicycle structure is much more coherent and realistic.
- + Good use of color and lighting for a cinematic feel.
- − The subject's face is obscured and lacks the 'natural skin texture' detail requested.
- − Motion blur on the background car is less convincing compared to Model A.
Verdict: FLUX.2 [dev] excels in photographic detail, particularly in the subject's face and the convincing motion blur of the passing traffic, though the physical structure of the bicycle is warped. Grok Imagine Image better captures the requested 'imperfect framing' and provides a more realistic bicycle, but it fails to deliver the high-quality skin textures and facial detail requested in the prompt. FLUX.2 [dev] is the winner for its superior rendering of the complex lighting and textures required for a realistic cinematic shot.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent execution of the leather straps and buckle textures
- + Accurate depiction of hair beads as requested in the prompt
- + More realistic skin texture with grime and faint scars
- − The scars look a bit like fresh cuts rather than faint, healed scars
- − Armor engraving is slightly less intricate than the competitor
Grok Imagine Image
- + Extremely intricate and beautiful armor engravings
- + Strong cinematic lighting with vibrant bokeh sparks
- + High contrast and sharp focus on the facial features
- − The skin looks a bit too smooth and airbrushed for a 'battle-worn' character
- − The 'hair beads' are less distinct, looking more like hair ties or wraps
- − Less visible texture on the leather straps compared to Model A
Verdict: Both models followed the prompt very well, but FLUX.2 [dev] stands out for its superior textural realism, particularly on the leather and skin, as well as the more literal interpretation of 'hair beads'. Grok Imagine Image produced a very aesthetically pleasing image with more ornate armor, but the face looks slightly too clean and stylized for the 'battle-worn' description.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent structure following a professional bi-fold or two-column layout.
- + High-quality, realistic food photography that looks appetizing and professional.
- + Clean use of vibrant accent colors to distinguish sections.
- − Text is mostly gibberish despite looking like words at a glance.
- − Small artifacts present in the pizza images where crusts blend slightly.
Grok Imagine Image
- + Legible text for item names like 'Appetizers' and 'Margherita'.
- + Creative layout with circular food crops integrated into the text flow.
- + Better adherence to the 'grid' request for the overall document structure.
- − Significant repetition of items (e.g., 'Steak Frites' and 'Grilled Salmon' appear multiple times).
- − Layout feels a bit cluttered compared to a professional print menu.
- − Some food items look flatter and less realistic than in the competing model.
Verdict: FLUX.2 [dev] produces an image that feels like a real, high-end professional menu design with superior photography, but it fails on text legibility. Grok Imagine Image succeeds in providing readable text and a creative layout, though it suffers from repetitive entries and slightly less polished food imagery. Overall, FLUX.2 [dev] is preferred for its superior aesthetic and professional design coherence.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent photorealistic texture on the patty and bun
- + Unified fiery lighting that affects all ingredients across the scene
- + Accurate rendering of the requested € symbol and starburst
- − The 'exploded' effect is a bit static with ingredients stacked almost perfectly
- − The fire in the background looks slightly like a low-resolution overlay
Grok Imagine Image
- + Strong sense of motion with ingredients scattered dynamically
- + Complex interaction between ingredients like melting cheese and sauce splashes
- + Vibrant colors and a very clean, professional layout for the text elements
- − The textures look slightly more 'digital' or illustrative than photorealistic
- − Lighting on the flying tomato and lettuce doesn't perfectly match the bottom fire sources
Verdict: Both models followed the prompt exceptionally well, particularly with the complex text and fire effects. FLUX.2 [dev] produced a more photorealistic image with superior realistic textures on the meat and bread, while Grok Imagine Image captured the 'exploded' and 'dynamic' aspects of the prompt more effectively with its energetic composition. FLUX.2 [dev] is the likely winner for its more cohesive lighting and more convincing food photography style.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent chalk texture on the board and lettering
- + Accurate spelling and completion of the menu items
- + Natural variation in handwriting slants and sizes
- − The word 'with' in the second item is smudged and partially repeated
- − Text layout feels slightly cramped toward the bottom
Grok Imagine Image
- + Clean and highly legible handwriting
- + Well-balanced composition with consistent spacing
- + Beautiful framing with the wooden border and cafe atmosphere
- − Lettering appears a bit too uniform, lacking raw chalk texture
- − The handwriting style is more of a clean font-like script than raw chalk
Verdict: Both models performed exceptionally well on complex text rendering, but FLUX.2 [dev] captures the authentic 'chalk' feel with better texture and realistic smudging on the board. Grok Imagine produced a cleaner, more aesthetically pleasing layout, though the text looks slightly more like a digital font than real hand-drawn chalk.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
FLUX.2 [dev]
- + Attempts to integrate the character and clothing from Image 2
- + Successfully captures the yellow background and red ottoman
- − Horrific anatomical errors, including a second head protruding from the torso
- − Fails to replicate the specific dynamic pose of the legs
- − Image exhibits severe artifacts and lack of physical coherence
Grok Imagine Image
- + High visual quality and resolution
- + Preserves the pose and background of Image 1 perfectly
- − Completely ignored the instruction to use Image 2 as the character reference
- − Failed the edit task entirely by just recreating Image 1
Verdict: FLUX.2 [dev] attempted the complex task but failed catastrophically, resulting in a disturbing image with multiple heads and broken anatomy. Grok Imagine Image completely ignored the prompt's instructions to swap characters and simply outputted a variation of Source Image 1, resulting in a total failure to edit the image.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent photographic realism and cinematic lighting
- + Balanced composition with detailed cosmic background
- − Failed the primary instruction to place the horse on top of the astronaut
- − Generates the cliché astronaut-riding-horse image instead of the requested surreal inversion
Grok Imagine Image
- + Followed the specific spatial instruction with the horse positioned on top
- + Captures the surreal nature of the prompt effectively
- + Vibrant colors and high-quality rendering of the nebula
- − The connection between the horse and astronaut is a bit ambiguous
- − The horse's anatomy is slightly elongated in the hindquarters
Verdict: FLUX.2 [dev] produced a high-quality but generic image of an astronaut riding a horse, completely ignoring the specific instruction 'horse on top, not vice versa'. Grok Imagine successfully interpreted the complex spatial request, placing the horse above the astronaut to create a truly surreal scene as requested.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
FLUX.2 [dev]
- + Successfully transferred the exact clothing items from Image 2 including the scarf and undergarment text.
- + Maintained the source background with high fidelity.
- − Failed to preserve the person from Image 1, instead creating a hybrid face that resembles the person in Image 2.
- − Added jewelry that was not present in either source image.
Grok Imagine Image
- + Perfectly preserved the character's face, hair, and specific skin details from Image 1.
- + Applied the clothing with a highly realistic fit to the character's pose and build.
- − Completely ignored Image 2 for the clothing source, providing a generic 'elaborate' royal outfit instead.
- − The added clothing has minor lighting inconsistencies with the background.
Verdict: Both models failed the specific multi-image task in different ways. FLUX.2 [dev] successfully identified the clothing from Image 2 but failed the instruction to keep the person from Image 1 unchanged, while Grok Imagine Image perfectly preserved the person but ignored the specific style of the second image. Grok Imagine Image is slightly preferred as its failure is a stylistic departure rather than a failure of identity preservation, though neither met all instructions.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent capybara face rendering with professional expression
- + Paws are placed more naturally on the steering wheel rim
- + Beautiful depth of field with the background bokeh
- − The passenger's hands and phone interaction are slightly blurry
- − The composition is a bit tight, losing some of the taxi exterior
Grok Imagine Image
- + Wide composition clearly shows more of the NYC street environment
- + Perfectly captures the 'bored' expression of the businesswoman
- + High detail in both the taxi interior and the outdoor lights
- − The capybara's left paw is positioned oddly, seemingly gripping the air behind the wheel
- − The taxi light on the roof appears inside the car's frame due to perspective distortion
Verdict: Both models successfully followed the complex prompt with high photorealism. FLUX.2 [dev] produced a more convincing capybara with better integration into the driver's seat, while Grok Imagine Image provided a more cinematic wide shot that better captured the requested atmosphere of the city streets and the passenger's expression.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent text legibility and alignment
- + Very clean and deliberate composition
- + Integrates the borders and scroll naturally into the frame
- − The parchment texture is less pronounced and feels a bit flatter
- − Lacks a visible moon in the night sky as compared to Model B
Grok Imagine Image
- + Rich parchment texture with more aged characteristics
- + Dynamic background with moon and detailed bats
- + More elaborate and decorative scroll banner
- − Text density is a bit cramped at the bottom
- − The border feels slightly detached from the parchment edges
- − The 'p' in 'frights' on the scroll is slightly garbled
Verdict: Both models followed the prompt exceptionally well, producing high-quality invitations with accurate text rendering. FLUX.2 [dev] is preferred for its superior layout balance and clean typography, whereas Grok Imagine offers a more textured, 'vintage' aesthetic but with slightly less polished text integration.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent preservation of the background and original lighting
- + High detail in the hair texture
- + Perfect preservation of the facial features and jacket details
- − The choice of an afro-textured hair style does not match the person's existing facial hair texture or likely ethnicity
- − The hairline integration looks slightly artificial at the temples
Grok Imagine Image
- + Natural-looking hair style that matches the length and texture of the person's beard
- + Seamless blend with the original facial structure
- + Excellent preservation of the source image's environment and accessories
- − None
Verdict: Grok Imagine Image provides a superior result because the chosen hairstyle is logically consistent with the subject's existing beard and features, creating a more realistic outcome. FLUX.2 [dev] successfully added hair while preserving the image perfectly, but the choice of hair texture feels mismatched to the subject.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent text rendering and placement according to instructions.
- + Very clean, realistic PBR materials with professional-looking lighting.
- + Perfectly follows the requested isometric 45-degree angle and diorama base.
- − The sushi rolls have slightly less variety in color than model b.
Grok Imagine Image
- + Bright, vibrant colors and appealing 3D cartoon style.
- + Good variety of sushi types shown in the scene.
- − Text is stacked differently than the prompt's request for 'JAPAN' then 'SUSHI' below it, with the flag at the very top.
- − The perspective on the 'JAPAN' text is slightly warped.
- − The rendering of the rice grains looks somewhat repetitive and artificial compared to model a.
Verdict: FLUX.2 [dev] followed the specific text layout and miniature diorama instructions much more accurately than Grok Imagine. While Grok Imagine produced a visually pleasing image, FLUX.2 [dev]'s rendering of materials and lighting felt more like a professional 3D isometric asset, which better matched the 'refined textures' and 'PBR materials' requested.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent caricature style with great facial resemblance and exaggerated features.
- + Includes many creative details like a cameraman and a hockey arena background.
- + Effectively integrates the dogs as jersey-wearing team members.
- − The text on the anchor desk is garbled and unintelligible.
- − The inclusion of the extra male character in the back is slightly confusing relative to the specific 'caricature of me' request.
Grok Imagine Image
- + Very clean and legible text on the news desk.
- + Good facial resemblance that translates the source image's features into cartoon form.
- + Creative use of hockey pucks and a dog in skates to satisfy the prompt requirements.
- − The hands on the desk are slightly mutated and lack proper finger definition.
- − The overall composition feels a bit more like a stickers-on-a-background layout rather than a cohesive caricature scene.
Verdict: Both models followed the complex instructions well, incorporating the profession, dogs, and hockey into a caricature. FLUX.2 [dev] felt more like a traditional editorial caricature with a cohesive hand-drawn art style and better environment integration, while Grok Imagine Image provided a more modern 'avatar' style with much clearer text but weaker anatomical details on the hands.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent fur texture and realistic lighting interactions.
- + Natural composition with varied, lifelike poses for the animals.
- + Very high prompt adherence including the specifically requested animals and butterflies.
- − Included two rabbits instead of the requested one baby bunny.
- − The animals are mostly sitting rather than 'tumbling together'.
Grok Imagine Image
- + Bold, vibrant colors with distinct god rays.
- + Cute, expressive faces that lean into the 'wholesome vibe'.
- − Anatomical issues with the kitten, particularly the paws and flat facial structure.
- − The 'butterflies' are rendered more like small insects or moths.
- − Less photorealistic and more stylized/CGI-like compared to the prompt's request.
Verdict: FLUX.2 [dev] produces a significantly more realistic image with superior fur detail and natural lighting, though it added an extra rabbit. Grok Imagine Image has a more saturated, 'cuter' aesthetic but suffers from anatomical distortions and fails to render convincing butterflies. FLUX.2 [dev] is the clear winner for its adherence to the 'hyper-photorealistic' part of the prompt.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent adherence to the 'soft pastel colors' and 'dreamy' prompt requirements
- + Captures the Ghibli aesthetic well through the use of watercolor-like textures and soft lighting
- + Maintains the composition and character poses of the original meme perfectly
- − Changed the background from a city street to a field of flowers, reducing source preservation
- − The faces look a bit generic and lose some of the specific character likeness from the source image
Grok Imagine Image
- + Successfully preserves the urban environment from the source image
- + Character faces are clearly recognizable as the individuals from the original meme
- + Good balance of Ghibli-style character lines with realistic environment details
- − The colors are slightly too saturated for the requested 'soft pastel' palette
- − The lighting is less 'gentle' or 'dreamy' compared to Model A
Verdict: Both models did an excellent job of translating a famous meme into the Ghibli art style. FLUX.2 [dev] achieved a more authentic Ghibli 'dreamy' atmosphere with its pastel palette and floral background, but Grok Imagine Image was superior at preserving the source material's urban setting and character likenesses. Overall, Grok Imagine Image is the winner for a more balanced edit that stays true to the original scene while applying the stylistic shift.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent rendition of hair blowing in the wind with natural physics.
- + Maintains high identity preservation and background consistency.
- + The leaves integrated realistically with the scene lighting and depth.
- − The leash handle has been distorted compared to the source image.
- − Some leaves appear slightly blurred in a way that looks like artifacting rather than motion.
Grok Imagine Image
- + Successfully adds a large quantity of leaves for a more 'energetic' feel.
- + Modifies the dog's ears to also show wind motion, enhancing the theme.
- + Overall layout of the scene is well-preserved.
- − The hair blowing effect is less voluminous and dynamic than Image A.
- − The orange leaves look like 2D stickers placed over the image rather than being part of the 3D environment.
- − The dog's face and eyes have lost some of the clarity seen in the source image.
Verdict: FLUX.2 [dev] followed the instructions with superior visual quality, particularly in how it rendered the hair blowing and integrated the leaves into the lighting of the scene. Grok Imagine Image added more 'energy' with a higher leaf count and moving the dog's ears, but the leaves themselves look poorly composited and the overall image quality took a slight hit in detail.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent typography capturing the requested 'classic' feel with perfect spelling.
- + Strongest adherence to the 'banner' restraint with a high-quality vector ribbon design.
- + Appropriate subtle texture that adds to the vintage aesthetic without being distracting.
- − The steam icon is very small and lacks the visual weight of the rest of the emblem.
Grok Imagine Image
- + Elegant cloche design with a more expressive and appealing steam illustration.
- + Good color palette using the requested warm brown and cream tones.
- + Clever integration of a coffee cup handle and spoon into the cloche shape.
- − Redundant text with 'Est. 1720' appearing twice in the composition.
- − Does not place 'Est. 1720' on a banner as specifically requested in the prompt.
- − The 'c' in 'Caffè' is lowercase and the accent mark on the 'è' is slightly disconnected.
Verdict: FLUX.2 [dev] followed the prompt more precisely by including the 'Est. 1720' text specifically on a banner and executing the typography with superior balance and spelling. While Grok Imagine produced a more creative cloche illustration, it failed on the specific banner instruction and included redundant text elements.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.2 [dev]
- + Excellent vector art style with professional gradients and crisp lines
- + Includes high-quality custom icons for the lunar module and landing site
- + Accurately represents the crew names with good legibility
- − Failed to follow the requested chronological order of steps
- − Text rendering is garbled for several labels (e.g., 'ULASIST ORDIT')
- − Included extra redundant icons that were not requested
Grok Imagine Image
- + Followed the numbered sequence of steps perfectly
- + Very clean alignment and layout fitting for an infographic
- + Correctly rendered the main titles and crew name labels
- − Icon design is slightly more simplistic than model_a
- − Includes some meta-text in the image like 'NASA inspired'
- − Minor spelling artifact on step 3 labels
Verdict: Grok Imagine Image is the winner because it successfully followed the logic of the prompt, providing the 6 requested steps in the correct numerical order with a clean, consistent layout. While FLUX.2 [dev] produced more sophisticated vector illustrations, its failure to organize the layout chronologically and the presence of severely garbled text labels made it less effective as an infographic.
Explore each model
An image generation model by xAI designed to generate highly aesthetic images from text descriptions.