Black Forest Labs' flagship image generation model delivering state-of-the-art quality with exceptional realism, precision, and consistency for both text-to-image and advanced image editing
Settled by community votes across 19 shared challenges, with an AI judge weighing in on each.
FLUX.2 [max]
#10 of 62 in Text-to-Image
Wan 2.5 (Preview)
#27 of 62 in Text-to-Image
Where the votes landed
FLUX.2 [max]
0.0%
win rate
Ties
100.0%
Wan 2.5 (Preview)
0.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent adherence to lighting instructions with realistic soft window light and shadows.
- + Highly realistic textures on the book cover and wood grain.
- + Clean, modern aesthetic with perfect geometric cube construction.
- − The plant is more behind the scene than specifically behind the cube to be seen through the glass clearly.
Wan 2.5 (Preview)
- + Successfully placed the plant directly behind the cube so it is visible through the glass.
- + Good tactile quality on the aged red book cover.
- + Dynamic lighting with visible dust motes in the air.
- − The cube is not a perfect cube; the bottom glass panel extends beyond the side walls.
- − The blue sphere has a slightly matte, less realistic interaction with the glass floor compared to Model A.
Verdict: Both models followed the prompt instructions very well. FLUX.2 [max] is the winner due to its superior geometric accuracy and more polished, realistic rendering of materials and lighting. While Wan 2.5 better handled the request to see the plant through the glass, it suffered from structural errors in the cube's construction.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
FLUX.2 [max]
- + Excellent preservation of the car's model and specific details from the source image
- + Accurately places the man from the second source image into the driver's seat including his distinct hairstyle and clothing
- + Realistic lighting and shadow integration for the outdoor setting
- − The scale of the man relative to the car seems slightly too large
Wan 2.5 (Preview)
- + Successfully captures a dynamic sense of motion with blurred wheels and background
- + Features iconic California elements like palm trees and coastal cliffs
- − Very poor preservation of the man's identity; the face and details are mostly lost
- − The car, though similar, has distorted proportions and is a left-hand drive version which places the driver on the wrong side for a standard PCH shot
Verdict: FLUX.2 [max] is the clear winner as it successfully merges both source images, preserving the specific car model and the man's unique appearance while placing them in the requested setting. Wan 2.5 (Preview) ignores most of the specific features of the man and has significant perspective distortions on the car.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent depiction of motion blur from passing cars as requested.
- + Realistic skin textures and weather effects with rain droplets visible on clothing and bike.
- + Strong adherence to the 'imperfect framing' cue, giving a candid street photography feel.
- − The bike basket has some structural inconsistencies where the metal wires meet.
- − The man's kneeling posture on the wet ground feels slightly unnatural for a quick repair.
Wan 2.5 (Preview)
- + Beautiful, clear reflections on the wet pavement.
- + High level of detail in the bicycle components and repair tools scattered on the ground.
- + Great shallow depth of field and bokeh execution.
- − Failed to include the requested motion blur from passing cars.
- − The scene feels a bit too staged and clean for a 'candid' street photo.
- − The lighting on the man's hair appears slightly artificial compared to the environment.
Verdict: FLUX.2 [max] followed the prompt more comprehensively, successfully incorporating the difficult motion blur and candid imperfect framing cues that give the image a realistic street-photo aesthetic. Wan 2.5 (Preview) produced a very sharp and aesthetically pleasing image with better reflections, but it missed the motion blur requirement and looks more like a professional portrait than a candid moment.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.2 [max]
- + Exquisite engraving details on the plate armor
- + Highly realistic skin texture including pores and fine scars
- + Subtle and sophisticated warm lighting integration
- − The braids are less distinct and partially obscured
- − The character looks slightly more like a general fantasy warrior than specifically a paladin
Wan 2.5 (Preview)
- + Excellent depiction of the braided hair and beads
- + Very clear and detailed texture on the leather straps and frayed cloth
- + Includes an actual torch in the background to justify the lighting
- − The engraving on the armor is a bit flat and less ornate than Model A
- − The face has a slightly smoother, more 'CG' look compared to the high-realism of the other image
Verdict: FLUX.2 [max] wins on raw photographic realism and the incredible intricate detail of the engraved armor. While Wan 2.5 (Preview) followed the hair braiding and clothing texture instructions more explicitly, FLUX.2 [max] produced a more cinematic and convincingly 'battle-worn' portrait with superior skin rendering.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent typography with clean, legible sans-serif fonts
- + Professional grid layout that logically separates text and images
- + Includes realistic price points and social media icons for a complete design feel
- − Internal logic errors where pizza images are listed under the Appetizers header
- − Text descriptions contain many gibberish characters
Wan 2.5 (Preview)
- + Strict adherence to the 3-section layout requested (Appetizers, Pizza, Mains)
- + High-quality, vibrant food photography with consistent styling
- + Good use of color-coded horizontal rules to separate sections
- − Text rendering is significantly worse with warped characters and illegible descriptions
- − Top header text is redundant and poorly integrated into the design
Verdict: FLUX.2 [max] creates a much more professional-looking menu with superior typography and a realistic commercial layout, though it struggles with the logic of which images belong in which section. Wan 2.5 (Preview) follows the section prompts more accurately and has very vibrant food images, but the text is nearly unreadable. FLUX.2 [max] is the winner for its overall aesthetic coherence and superior font rendering.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent text legibility and clean graphic design.
- + High-resolution texture on the burger bun and patty.
- + Effective use of the 'starburst' element for pricing as requested.
- − The 'exploded' effect is less dynamic, with many components still appearing stacked.
- − The bottom bun looks slightly flat and detached compared to the overall lighting.
Wan 2.5 (Preview)
- + Superb dynamic composition with a true 'exploded' feel and motion.
- + Creative typography with a melting/fiery effect that blends with the theme.
- + Vibrant lighting and more energetic background with glowing embers and smoke.
- − The pricing text is slightly less polished in its rendering.
- − The lettuce in the upper right appears a bit disconnected from the main burger assembly.
Verdict: Both models followed the prompt exceptionally well, but Wan 2.5 (Preview) captured the 'dynamic' and 'exploded' requirement much better than FLUX.2 [max], creating a sense of genuine motion. While FLUX.2 produced a cleaner, more standard advertisement look, Wan 2.5 achieved a more impressive artistic result with superior lighting and energetic composition.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.2 [max]
- + Exceptional chalk texture with realistic dusty smudges and varying opacity
- + Flawless text rendering including longer, complex words
- + Superior composition with realistic lighting from an overhead lamp
- − The handwriting style is slightly more uniform than a natural human slant might be
Wan 2.5 (Preview)
- + Good handwritten flow with natural slants and variations in letter size
- + Includes realistic chalk dust piles on the frame and board
- − Failed to include 'Herbs' in the second menu item
- − Repeated the price '$9' on two separate lines for the cookies
- − The letter 's' in 'Specials' has a slight digital font artifact look compared to the rougher chalk in Model A
Verdict: FLUX.2 [max] is the clear winner as it followed the complex menu text perfectly and produced a much more realistic chalk texture on the board. Wan 2.5 (Preview) struggled with the text content, omitting words and repeating the price for the final item, which made the menu nonsensical.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent high-detail rendering of the spacesuit and horse texture.
- + Strong composition with the asteroid providing a logical grounding for the surreal scene.
- + Good use of cinematic lighting and atmospheric depth in the background.
- − The horse's legs appear somewhat thin and elongated compared to the body.
- − Failed the specific spatial instruction; the astronaut is on top of the horse.
Wan 2.5 (Preview)
- + Dynamic sense of movement with the galloping pose and debris trails.
- + Vibrant color palette and clear background elements like the planet and galaxy.
- + Clean anatomical rendering of the horse's head.
- − Failed the specific spatial instruction; the astronaut is riding the horse instead of the reverse.
- − Minor artifacts in the rendering of the horse's rear hooves against the foreground blur.
Verdict: Both models failed the specific prompt constraint to have the horse on top of the astronaut, instead defaulting to the standard image of an astronaut riding a horse. FLUX.2 [max] is the preferred choice as it offers superior textural detail and a more balanced, cinematic composition compared to Wan 2.5 (Preview).
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
FLUX.2 [max]
- + Successfully preserved the exact facial features and skin condition (vitiligo) of the person in Image 1.
- + Accurately reproduced the outfit details, including the specific tartan pattern, coat texture, and jewelry.
- + Maintained the background and lighting of the source image while naturally integrating the clothing.
- − The person's hands are slightly distorted and don't match the original skin tone perfectly.
Wan 2.5 (Preview)
- + Captured the clothing and accessories (sunglasses) from Image 2 well.
- − Failed the primary instruction to keep the person from Image 1, replacing him with the person from Image 2.
- − Altered the face, hair, and skin characteristics entirely, ignoring the source preservation requirement.
Verdict: FLUX.2 [max] followed the complex instructions perfectly, successfully 'dressing' the specific individual from Image 1 in the clothing from Image 2 while preserving his unique physical characteristics. Wan 2.5 (Preview) failed the editing task by simply overlaying the person from Image 2 onto the background of Image 1, disregarding the instruction to keep the original person's face and hair.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent photographic lighting and skin textures
- + High level of interior detail such as the radio and seat textures
- − The driver has human hands coming out of the sleeves instead of capybara paws
- − The capybara's head is not naturally joined to the human-like body
Wan 2.5 (Preview)
- + Correctly depicts capybara paws on the steering wheel
- + Very cinematic composition with a clear view of both subjects and the Times Square-style background
- + Accurately captures the 'bored' expression of the passenger
- − The capybara's face is a bit flatly lit compared to the background
- − Slightly less realistic interior textures compared to the other model
Verdict: Wan 2.5 (Preview) provided the superior result because it correctly interpreted the anatomy of the driver, giving the capybara paws instead of the unsettling human hands found in the FLUX.2 [max] output. While FLUX.2 has slightly better photographic grain and lighting, the anatomical failure on the main subject makes Wan 2.5 the clear winner for prompt adherence.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent atmospheric lighting with a cinematic feel
- + Perfect text legibility and alignment
- + Superior border integration that feels cohesive with the gothic theme
- − The jack-o-lantern face is a bit standard compared to the more creative flame effect in the other image
Wan 2.5 (Preview)
- + Intricate flame details inside the jack-o-lantern eyes and mouth
- + Vibrant colors that pop against the parchment background
- + Good interpretation of the 'twisted trees' and gothic architecture elements
- − The central illustration being a circle within a square feels less like a poster and more like a sticker
- − The thorn border looks a bit repetitive and digitally pasted on
- − The text alignment is slightly cramped at the bottom
Verdict: FLUX.2 [max] creates a more cohesive and professional-looking invitation with a moody, cinematic atmosphere and perfectly integrated text. Wan 2.5 (Preview) has impressive details within the jack-o-lantern, but its composition feels more like a collection of clip-art elements rather than a unified parchment poster.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent preservation of original identity and facial features
- + Hair texture looks wind-swept and natural for the setting
- + Perfect lighting integration with the existing scene
- − One small stray hair cluster at the very top looks slightly detached
Wan 2.5 (Preview)
- + Successfully added thick hair with good volume
- + Consistent lighting on the hair
- − Significantly altered the facial structure, making the subject look like a different person
- − The hair style is a bit too manicured/styled for the rugged desert context
Verdict: FLUX.2 (max) followed the instructions perfectly, adding a realistic head of hair while keeping the man's identity completely intact. Wan 2.5 (Preview) failed at source preservation by changing the man's face and features to a point where he is no longer recognizable as the original subject.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.2 [max]
- + Perfect adherence to the 45° top-down isometric perspective.
- + Accurately represents a miniature 3D cartoon scene with a raised diorama base.
- + Excellent text rendering and layout balance.
- − The color of the text is slightly muted compared to the prompt's 'bold' request, though it is still very clear.
Wan 2.5 (Preview)
- + High-quality PBR materials with attractive reflections on the salmon.
- + Clear, bold text rendering with a well-placed flag icon.
- + Gentle, appealing lighting.
- − Completely failed the '45° top-down isometric' perspective request, opting for a low-angle eye-level shot instead.
- − The base is a simple cylinder rather than the requested 'diorama base'.
Verdict: FLUX.2 [max] followed every detail of the prompt perfectly, specifically the difficult isometric perspective and the miniature diorama aesthetics. While Wan 2.5 (Preview) produced high-quality materials, it failed on the core compositional requirement of an isometric top-down view.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent integration of all three themes (hockey, news anchor, dogs) within a cohesive scene.
- + Character design strongly resembles the original subject, especially the hair and smile.
- + High visual quality with detailed background elements like the scoreboard and studio lighting.
- − Text on the news desk and scoreboard is slightly distorted/garbled.
Wan 2.5 (Preview)
- + Features a fun detail with a dog wearing a hockey jersey.
- + Clear, clean vector-style caricature aesthetic.
- + Successfully captures the core elements of the prompt including the microphone and hockey stick.
- − The facial likeness is significantly weaker than the other model, losing the original subject's features.
- − The composition feels a bit more disjointed with floating elements against a simple background.
Verdict: FLUX.2 [max] is the winner because it maintains a high degree of facial resemblance to the source image while creatively blending the hockey rink and news desk environments. Wan 2.5 (Preview) produces a generic caricature face that lacks the identity of the original subject, although it does include a charming dog in a hockey jersey.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent balance of lighting with realistic god rays and soft morning mist.
- + Superior anatomical silhouettes and naturalistic fur textures for all animals.
- + Better sense of scale and depth within a cohesive environment.
- − The animals' eyes are slightly less 'big and expressive' than requested compared to Model B.
Wan 2.5 (Preview)
- + Strong adherence to the 'big expressive eyes' and 'joyful' descriptors.
- + Very vibrant colors that emphasize a whimsical, storybook atmosphere.
- + Captures the dew sparkles well with floating droplets.
- − The fox's eyes appear unnaturally glowing and slightly distorted.
- − Anatomical issues in the front paws and limbs of the kitten and puppy.
- − The butterfly on the left has anatomical inconsistencies and looks pasted on.
Verdict: FLUX.2 [max] is the winner for its superior realism, sophisticated lighting, and coherent composition that makes the scene feel grounded and high-quality. While Wan 2.5 (Preview) captures the requested 'expressive' emotion more overtly, it suffers from anatomical artifacts and less realistic textures that detract from the overall 8K masterpiece requirement.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
FLUX.2 [max]
- + Matches the 'hand-painted textures' and 'soft pastel' prompt perfectly with a watercolor paper aesthetic.
- + Captures the warm, nostalgic Ghibli mood through textured overlays and glowing lighting.
- + Preserves the specific color palette of the original clothing very well.
Wan 2.5 (Preview)
- + Very clean line art and character design that feels like modern digital anime.
- + Creative addition of falling leaves adds a whimsical, Ghibli-esque environmental detail.
- + Maintains clear distinction between the characters and the background.
- − Lacks the requested 'hand-painted texture', looking more like standard digital cel-shading.
- − The colors are a bit too vibrant and high-contrast compared to the requested 'soft pastels'.
Verdict: FLUX.2 [max] is the winner because it adhered much more closely to the specific textural requirements of the prompt, delivering a beautiful watercolor, hand-painted aesthetic that feels authentically like a Ghibli background. While Wan 2.5 (Preview) produced a high-quality anime illustration, it felt more like a modern digital render rather than the soft, nostalgic, and textured illustration requested.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
FLUX.2 [max]
- + Successfully added highly detailed leaves with realistic shading and depth.
- + Created a strong sense of motion in the hair that looks natural and dynamic.
- + Preserved the subject's face and the dog's appearance with high fidelity.
- − The leash has a slight continuity break where it meets the hand/collar compared to the source.
- − Some leaves in the foreground are slightly blurry, though this adds to the depth of field.
Wan 2.5 (Preview)
- + Nicely rendered hair movement that feels light and airy.
- + Preserved the overall composition and lighting of the source image well.
- − The flying leaves look like flat, bright green illustrations rather than realistic foliage.
- − The dog's tail has a motion blur artifact that looks slightly smudged rather than energetic.
- − The edit to the leaves feels low-effort compared to the rest of the scene.
Verdict: FLUX.2 [max] is the clear winner as it integrated the requested elements more realistically, especially the flying leaves which have varied textures and colors. Wan 2.5 (Preview) produced leaves that appeared like cartoonish overlays, detracting from the overall photo quality despite a decent attempt at the hair motion.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.2 [max]
- + Excellent typography including the grave accent in 'Caffè'
- + Clean vector emblem style with balanced proportions
- + Subtle and professional paper texture background
- − The steam lines are very thin and lack the weight of the rest of the illustration
Wan 2.5 (Preview)
- + Stronger visual contrast with bold brown tones
- + Accurate typography and branding placement
- + Dynamic steam illustration that matches the logo weight
- − Background texture is a bit heavy-handed with the artificial paper creases
- − Composition feels slightly more generic than the circular emblem
Verdict: Both models followed the prompt exceptionally well, producing high-quality logos with accurate text. FLUX.2 [max] creates a more sophisticated circular emblem that feels like a finished brand identity, whereas Wan 2.5 (Preview) provides a more illustrative and bold approach. FLUX.2 [max] is the winner for its cleaner vector aesthetic and more balanced use of the vintage background texture.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.2 [max]
- + Includes all 6 steps required in the prompt.
- + Follows the muted NASA-inspired color palette accurately.
- + Text rendering for the steps and astronaut names is highly legible.
- − Repeats the label 'Earth Orbit' three times incorrectly.
- − The rocket design is more generic and less resembling a Saturn V.
Wan 2.5 (Preview)
- + More accurate Saturn V rocket vector illustration.
- + Stronger 'modern vector' aesthetic with a more professional infographic layout.
- + Excellent use of the navy, white, and red color scheme.
- − Fails to include all 6 steps, missing visual representations for Descent and Landing specifically.
- − One of the astronaut icons appears to be wearing a helmet that resembles a soldier rather than an astronaut.
Verdict: FLUX.2 [max] followed the complex step-by-step instructions more literally and included all requested stages, though it suffered from repeated text errors. Wan 2.5 (Preview) produced a much more visually appealing and stylistically accurate 'NASA' vector poster, but it skipped several requested steps in the sequence.
Explore each model
Alibaba's text-to-image and image-to-image generation model from the Wan AI suite, offering high-quality visual generation capabilities