ByteDance's latest image generation model unifying text-to-image and image editing in a single architecture, with improved text rendering and 30-40% faster generation than v4.0
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
Seedream 4.5
#9 of 62 in Text-to-Image
Wan 2.7
#38 of 62 in Text-to-Image
Where the votes landed
Seedream 4.5
100.0%
win rate
Ties
0.0%
Wan 2.7
0.0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
Seedream 4.5
- + Excellent photographic realism and soft lighting.
- + Accurate reflection and refraction through the glass panels.
- + Clean, modern aesthetic with high focus on the central subject.
- − The glass cube seems to lack a front panel or the bottom joinery is slightly inconsistent.
Wan 2.7
- + Detailed textures on the book and the rustic wooden table.
- + Impressive interaction between the plants and the glass reflections.
- + Includes realistic text/embossing on the spine of the red book.
- − The glass geometry is slightly confusing with extra vertical lines in the corners.
- − The blue sphere has an odd textured/stony appearance rather than being a smooth ball.
Verdict: Both models followed the spatial prompt perfectly. Seedream 4.5 produced a cleaner, more professional-looking photograph with superior lighting, while Wan 2.7 excelled at fine surface details like the wood grain and book texture. Seedream 4.5 is the preferred choice for its more coherent depiction of glass and light.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
Seedream 4.5
- + Successfully preserved the subject's face, hair, and entire outfit including shoes.
- + High-fidelity preservation of the specific car model and interior details.
- + Excellent integration of the character into the driver's seat with realistic lighting.
- − The car door is wide open while driving, which is physically illogical.
- − The subject is sitting on the side sill/edge of the seat rather than fully inside.
Wan 2.7
- + The car is correctly positioned on a road moving along a coastline.
- + Maintains the exterior appearance of the car well.
- − Completely failed to preserve the subject's face and identity.
- − The driver appears as a distorted, low-detail figure that does not resemble the source man.
- − Lost all detail of the subject's specific outfit.
Verdict: Seedream 4.5 is the clear winner for its superior ability to preserve the subject's identity and specific clothing, making it a true image composite. While it has a logical flaw with the car door being open, Wan 2.7 fails the primary task by replacing the man with a generic, distorted figure that looks nothing like the source image.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
Seedream 4.5
- + Excellent adherence to the motion blur and low-light cinematic requirements.
- + Strong focus on micro-details like realistic skin texture and raindrops on the jacket.
- + Composition feels candid and tells a story with vibrant reflections on the pavement.
- − The transition from the car to the background shows some digital smearing.
- − The position of the man's hands on the spokes is a bit physically ambiguous.
Wan 2.7
- + Natural, everyday street photography aesthetic.
- + Good character design that fits the description of an elderly Japanese man.
- + Reflections on the wet pavement are very crisp.
- − Failed to include the requested motion blur from passing cars.
- − The depth of field is deeper than the 'shallow' request, making the background distracting.
- − The bicycle's structure has major geometric flaws, particularly near the front wheel and pedals.
Verdict: Seedream 4.5 followed the prompt much more accurately, successfully incorporating difficult elements like motion blur and a shallow depth of field to create a cinematic atmosphere. In contrast, Wan 2.7 missed several stylistic cues and produced a bicycle with noticeable structural errors, though it did capture a convincing candid street photography vibe.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
Seedream 4.5
- + Exceptional skin texture with realistic pores and fine hairs
- + Masterful use of warm, cinematic lighting and bokeh
- + Highly intricate engraving detail on the pauldrons
- − The 'close portrait' framing cuts off much of the requested armor and leather straps
Wan 2.7
- + Excellent adherence to the 'battle-worn' description with convincing dents and scratches on armor
- + Clearer representation of the braided hair and beads
- + Better framing to show the full range of textures including leather and cloth
- − Slightly less realistic skin rendering compared to the competitor
- − The torch lighting feels less integrated with the subject's face
Verdict: Seedream 4.5 produces a more photorealistic and stylistically beautiful 'close portrait' with incredible skin and lighting detail. However, Wan 2.7 better captures the 'battle-worn' aspect and provides a more complete view of the materials requested in the prompt, such as the leather straps and multiple layers of clothing.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
Seedream 4.5
- + Excellent adherence to high-level layout requirements including specific sections for appetizers, pizza, and mains.
- + High-quality, appetizing food photography with vibrant color-coded accents.
- − Text rendering is poor with several typos like 'Appetizters' and nonsensical item names.
- − Overall layout feels a bit sparse for a practical menu.
Wan 2.7
- + Highly sophisticated and professional graphic design layout that looks like a real-world product.
- + Impressive text rendering for a menu, including realistic pricing and detailed food descriptions.
- + Creative use of a lifestyle background with rosemary and a pen to ground the design.
- − The 'grid' requested is heavily populated, making individual photos smaller than might be ideal for some minimalist designs.
- − Minor spelling errors in smaller text blocks, though much better than the competitor.
Verdict: Wan 2.7 is the clear winner as it produces a fully realized, professional-grade menu design with complex typography, logos, and realistic item descriptions. While Seedream 4.5 followed the requested sections accurately, its text rendering and layout are too primitive compared to the high-fidelity graphic design output of Wan 2.7.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
Seedream 4.5
- + Excellent photorealistic texture on the meat and bun
- + Very natural-looking motion blur on flying ingredients
- + Seamless integration of the burger with the fiery environment
- − The burger is not fully 'exploded' as requested, with most layers still touching
- − Text is slightly less sharp than in the competing model
Wan 2.7
- + Perfect adherence to the 'exploded' aspect of the prompt with all layers separated
- + Extremely clean and vibrant graphic design and typography
- + Creative use of sauce splashes to enhance the sense of motion
- − Some elements like the sesame seeds and smoke look more 'digital' and less photorealistic
- − The lettuce and tomato cross-section look slightly flat compared to the lighting on the meat
Verdict: Both models followed the prompt details remarkably well, including all requested text and the fiery theme. Seedream 4.5 produces a more convincing photorealistic image with superior lighting and motion blur, while Wan 2.7 provides a much better 'exploded' composition and sharper, more professional-looking advertisement typography.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
Seedream 4.5
- + Features realistic chalk texture with varying opacity and dust particles.
- + Successfully renders handwritten text with natural variations in letter size and slant.
- + Accurately completes truncated items from the prompt like 'Brown Butter Chocolate Chip Cookies'.
- − Repeats the word 'Risotto' and the price '$24' unnecessarily on two lines.
- − Text slant and spacing are slightly messy, though this fits the 'handwritten' request.
Wan 2.7
- + Text is perfectly centered and highly legible.
- + Includes all requested menu items with no spelling errors.
- + Clean, professional composition with nice lighting and background elements.
- − Failed the negative constraint; the text looks like a clean digital font rather than authentic chalk handwriting.
- − The 'handwriting' is too uniform, lacking the requested slants and natural variations.
Verdict: Seedream 4.5 much better captures the intended 'chalkboard' aesthetic, featuring authentic textures, smudges, and imperfect handwriting that truly looks hand-drawn. While Wan 2.7 has cleaner layout and spelling, it relies on a digital-looking font that ignores the specific prompt instructions for natural variations and a non-digital appearance.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
Seedream 4.5
- + Successfully integrated character identity from Image 2 including face, sunglasses, scarf, and clothing details
- + Accurately recreates the dynamic pose from Image 1
- + Matches the studio lighting and yellow background of the reference perfectly
- − The right hand (upper) has slightly warped finger anatomy
- − The character's expression is more smiling/relaxed than the serious expression in Image 2
Wan 2.7
- + High resolution and clarity
- − Completely failed to incorporate the character reference from Image 2
- − Essentially just recreated Image 1 with minor color adjustments
- − Did not follow the instruction to change the person/clothing
Verdict: Seedream 4.5 successfully performed the complex task of merging the identity of the character in Image 2 with the precise pose and environment of Image 1, maintaining clothing details like the scarf and sunglasses. Wan 2.7 failed the prompt entirely, simply outputting a version of Image 1 and ignoring the character reference image.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
Seedream 4.5
- + Successfully followed the difficult positional constraint with the horse on top of the astronaut.
- + Dynamic and cinematic lighting with vibrant nebula colors.
- + Excellent anatomical detail on the horse's head and neck.
- − The astronaut's lower body merging into the horse's back is a bit messy.
- − The perspective of the astronaut's legs is slightly distorted.
Wan 2.7
- + Clean, high-resolution rendering of the astronaut and planets.
- + Stable composition with clear celestial objects.
- + Good lighting consistency between the subjects and the orbital background.
- − Completely failed the negative constraint, depicting the astronaut riding the horse.
- − Nonsensical repetition of Earth scattered in the background.
- − Lacks the 'surreal' quality requested, looking like a standard collage.
Verdict: Seedream 4.5 is the clear winner because it correctly interpreted the specific request for the horse to be on top of the astronaut, creating a truly surreal image. Wan 2.7 ignored the positional instruction and generated a standard astronaut riding a horse, while also featuring odd repeated Earths that diminished the visual quality.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
Seedream 4.5
- + Excellent adherence to the specific clothing items from Image 2
- + Maintains the exact face and hair of the person from Image 1
- + Includes the specific accessories like the ring and gold watch
- − Fails to realistically layer the shirt over the person's torso, resulting in a strange cutout of the skin
- − The pose is slightly altered from the original image
- − Poor proportions on the hands
Wan 2.7
- + Successfully adapts a new pose that feels natural in the environment
- + Maintains the likeness/skin patterns of the base person well
- + High visual quality and resolution
- − Completely ignored the clothing in Image 2, creating a generic 'elaborate' outfit instead
- − Altered the layout and composition of the original scene significantly
- − The hands and fingers have anatomical errors
Verdict: Seedream 4.5 followed the complex instructions much better by using the actual coat, scarf, and jeans from Image 2, even though the internal rendering of the shirt and skin is messy. Wan 2.7 failed the primary task of using the 'exact' outfit, instead generating a random decorative suit that was never present in the source images. Seedream 4.5 is the preferred choice for following the specific edit constraints despite its technical artifacts.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
Seedream 4.5
- + Excellent adherence to the 'inside the taxi' perspective.
- + High-quality fur texture and realistic passenger expression.
- + Clean, readable text on the taxi cap.
- − The perspective makes the capybara look giant compared to the passenger.
- − The lighting on the front of the capybara is slightly flat.
Wan 2.7
- + Natural composition from an outside-looking-in perspective.
- + Very detailed capybara hands/paws on the wheel.
- + Good lighting and integration with the city backdrop.
- − The capybara's head has a slightly 'pasted-on' look with some artifacts around the neck.
- − The steering wheel appears to be emerging from the car door frame incorrectly.
- − Garbled text on the taxi roof sign.
Verdict: Seedream 4.5 captures the requested 'inside' perspective much better than Wan 2.7, creating a more immersion scene that feels directly from the car's interior. While Wan 2.7 has impressive anatomical detail on the paws, the composition and overall coherence of Seedream 4.5 make it the superior choice for this prompt.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
Seedream 4.5
- + Excellent photographic realism in the jack-o-lantern and mood lighting
- + Perfectly accurate text rendering for all requested details
- + Effective use of negative space for readability
- − The border elements like the barbed wire feel like separate overlays rather than an integrated part of a 'vintage poster'
- − Composition is slightly sparse compared to the 'vintage gothic' aesthetic requested
Wan 2.7
- + Outstanding interpretation of the 'vintage gothic' style with an intricate, hand-illustrated feel
- + Highly detailed border containing all requested elements (webs, thorns, bats) plus extra thematic additions
- + Text is beautifully integrated into the layout with a more appropriate gothic font
- − The jack-o-lantern and lighting are more illustrative/cartoonish than cinematic
- − The image includes extra text not requested in the prompt, such as 'Est. 1847' and 'Dress code'
Verdict: Seedream 4.5 produces a high-fidelity, modern take on the invitation with superior text clarity and realistic lighting. However, Wan 2.7 far better captures the 'vintage gothic' aesthetic with its complex, thematic border and stylized illustration, making it look much more like a real invitation despite some minor creative liberties with the text.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
Seedream 4.5
- + Excellent preservation of the original image's identity and lighting.
- + The hair texture and color perfectly match the existing beard.
- + Maintains original head shape and facial features without distortion.
- − The hairline is a bit high and simplified in shape.
Wan 2.7
- + Natural, modern hair style with good volume and texture.
- + Realistic blending at the temples and forehead.
- − Slightly altered the person's facial features, making them look like a different individual.
- − The head shape and forehead dimensions feel slightly adjusted from the source.
Verdict: Seedream 4.5 is the winner because it successfully adds the requested hair while perfectly preserving the identity, lighting, and background of the source image. Wan 2.7 provides a more stylish haircut, but it subtly changes the subject's face, failing the 'preserve facial features' part of the instruction.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
Seedream 4.5
- + Excellent PBR textures with a realistic stone-like diorama base
- + High-quality soft lighting and depth of field
- + Clean, accurate text rendering
- − Low count of sushi pieces making the diorama look slightly empty
- − The flag icon is floating somewhat awkwardly compared to the text
Wan 2.7
- + More diverse variety of sushi items including nigiri and rolls
- + Clean, toy-like aesthetic that fits the 3D cartoon scene description well
- + Perfectly centered composition with tidy layout
- − The text has slight white fringing/outlining not requested in the 'large bold text' prompt
- − Textures are a bit flatter and less 'realistic PBR' compared to Model A
Verdict: Seedream 4.5 captures the requested PBR materials and realistic lighting much better, but Wan 2.7 provides a more complete and visually interesting miniature scene with a wider variety of sushi. While Seedream 4.5 has superior texture work on the base and salmon, Wan 2.7's composition feels more like a finished diorama asset and follows the spirit of a 'cartoon scene' more effectively.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
Seedream 4.5
- + Excellent preservation of the subject's facial features in a caricature style.
- + Clear and logical incorporation of all three hobby/job elements into a single scene.
- + High visual quality with realistic lighting and textures.
- − The caricature style is a bit subtle, mostly just enlarging the head rather than exaggerating features.
Wan 2.7
- + Strong 'caricature' art style with high energy and humor.
- + Creative integration of the hockey theme with the dogs.
- + Maintains the selfie perspective of the source image.
- − Facial likeness is significantly altered compared to the source image.
- − Text in speech bubbles contains typos ('Rolee').
- − The composition is a bit cluttered and chaotic.
Verdict: Seedream 4.5 is the winner because it successfully creates a professional-looking edit that maintains a clear likeness to the woman in the source image while neatly organizing the requested profession and hobbies. While Wan 2.7 has more 'humorous' energy and a more distinct caricature art style, it loses the subject's identity and includes several text errors.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
Seedream 4.5
- + Excellent dynamic composition with a sense of movement and 'tumbling'.
- + Strong cinematic lighting with god rays and atmospheric bokeh effects.
- + Vibrant color palette that enhances the joyful theme.
- − The fox has overly stylized, 'Disney-fied' eyes that lean toward illustration rather than hyper-photorealism.
- − The puppy's back-left leg transition is slightly blurry/muddled.
Wan 2.7
- + Detailed fur textures across all four animals.
- + Better representation of a 'lush wildflower meadow' with distinct plant types.
- + Well-defined god rays emanating from the top center.
- − Static, somewhat stiff posing for most of the animals.
- − The kitten's tail appears to be missing or merged into the rabbit.
- − Anatomical oddity with the kitten's front-right paw looking like it's attached to the rabbit.
Verdict: Seedream 4.5 captures the spirit of the prompt much better by depicting the animals 'tumbling' and actively chasing butterflies, whereas Wan 2.7 feels like a static lineup. While Wan 2.7 has slightly sharper fur details, it suffers from anatomical errors where the kitten and rabbit overlap, whereas Seedream 4.5 produces a more cohesive and emotionally resonant masterpiece.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
Seedream 4.5
- + Perfectly captures the Studio Ghibli anime aesthetic in character design and line work.
- + Preserves the original composition and pose while adapting to a new style.
- + Excellent use of watercolor textures and soft pastel coloring.
- − The facial features of the woman in red are simplified and differ significantly from the original source.
Wan 2.7
- + Maintains a much higher level of facial likeness to the original people in the photo.
- + Beautiful watercolor texture and soft, warm lighting as requested.
- + Preserves subtle details of the original clothing and background.
- − The style leans more toward a realistic watercolor illustration than the specific 'Studio Ghibli' anime aesthetic.
- − The woman in the background has a slightly distorted eye.
Verdict: Seedream 4.5 is the clear winner for style adherence, as it perfectly replicates the iconic Studio Ghibli character designs and line art while maintaining the meme's composition. Wan 2.7 creates a beautiful watercolor painting that preserves the subjects' faces much better, but it fails to capture the specific anime 'look' requested in the prompt.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
Seedream 4.5
- + Excellent depiction of wind-blown hair that looks natural and dynamic.
- + Strong adherence to the 'energetic and lively' instruction through pose adjustment.
- + Good integration of flying leaves with motion blur.
- − Significantly altered the woman's pose and camera angle from the source image.
- − The background lighting and bridge details were modified unnecessarily.
Wan 2.7
- + Expertly preserved the original pose, facial features, and background elements.
- + Added believable hair motion that fans out naturally.
- + Leaves are integrated subtly without overwhelming the composition.
- − The leaf placement looks slightly more static compared to Model A.
Verdict: Seedream 4.5 captures the most 'energetic' feel with a more dramatic hair toss and pose change, but it fails as a precise edit by fundamentally changing the subject's stance and the background geometry. Wan 2.7 is the superior editing model, successfully adding the requested motion effects (blowing hair and flying leaves) while near-perfectly preserving the identity and composition of the original source image.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
Seedream 4.5
- + Excellent typography with perfect spelling and accent marks.
- + Clean vector emblem style that adheres to the minimalist request.
- + Strong composition with a balanced arrangement of text, cloche, and banner.
- − The 'steam' element is a bit small relative to the rest of the logo.
Wan 2.7
- + Detailed vintage badge aesthetic with nice border ornaments.
- + Good color palette that matches the requested brown and cream tones.
- + Creative use of transparency on the cloche dome.
- − Spelling error in the main name ('Florion' instead of 'Florian').
- − The composition is quite busy, diverging from the 'minimalist' requirement.
- − Inconsistent line weights in the steam and floral elements.
Verdict: Seedream 4.5 is the clear winner as it followed all prompt instructions perfectly, including difficult spelling and the minimalist style. Wan 2.7 produced an attractive badge but failed on the spelling of the brand name and created a much more cluttered design than requested.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
Seedream 4.5
- + Excellent typography rendering with almost perfect spelling and legibility.
- + The illustration for the Saturn V and Lunar Module are detailed and high-quality.
- + Adheres strictly to the color palette and requested stop/start points.
- − The 'Descent' icon is a satellite rather than the requested lunar module.
- − The composition feels slightly bottom-heavy with large empty space above.
Wan 2.7
- + Superior infographic layout with logical flow and supporting data points.
- + Consistent icon style across all six stages of the mission.
- + Excellent use of the NASA-inspired color palette and modern vector aesthetic.
- − Minor spelling errors such as 'DESCRIPT' instead of 'DESCENT' and 'Tranquiliry'.
- − Text becomes slightly fuzzy and less legible at smaller sizes compared to Model A.
Verdict: Model B (Wan 2.7) is the superior infographic due to its sophisticated layout, use of supporting data, and consistent iconography that tells a complete story, whereas Model A (Seedream 4.5) feels more like a list of icons. While Seedream 4.5 has better text rendering, Wan 2.7 captures the 'modern vector infographic' prompt more effectively through its composition and data visualization style.
Explore each model
Alibaba's Wan 2.7 image generation and editing model for text-to-image, reference-guided generation, and instruction-based image edits