Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [max] Black Forest Labs Vidu Q2 ShengShu Technology

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

FLUX.2 [max]

25.8 arena score

#10 of 62 in Text-to-Image

Skill signature · Text-to-Image

Vidu Q2

19.8 arena score

#42 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [max]

0%

win rate

Ties

0%

Vidu Q2

0%

win rate

Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent photorealistic texture on the red book cover
  • + Sophisticated glass refractions and reflections on the sphere
  • + Successfully placed all elements as requested
  • The plant is very blurred due to shallow depth of field, making it less visible through the glass than requested

Vidu Q2

  • + Vibrant colors and clear visibility of the plant through the glass
  • + Good adherence to all spatial requirements
  • + Sharp details across the entire frame
  • The blue sphere has a matte, plastic appearance that looks less realistic than the glass context
  • The lighting on the book spine contains some minor artifacts/unnatural text distortion

Verdict: Both models followed the prompt perfectly in terms of object placement. FLUX.2 [max] achieved a much more realistic, high-end photographic look with superior light behavior on the glass surfaces, whereas Vidu Q2 provided clearer visibility of the background plant but with slightly less realistic textures for the sphere and book.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent preservation of the car's geometry and details from the source image.
  • + Accurately places the man in the driver's seat with consistent clothing and hair.
  • + Superior realism in lighting and integration into the coastal environment.
  • The man's skin tone appears slightly darker than the source photo.
  • The car has a UK-style right-hand drive configuration (though consistent with some Rolls Royce models).

Vidu Q2

  • + Successfully captures a dynamic sense of motion with the background blur.
  • + Good interpretation of the 'California coastline' with a winding road.
  • The man's likeness and hair style are less consistent with the source image.
  • Significant anatomy issues with the hands on the steering wheel.
  • Distorted car proportions and perspective compared to the source image.

Verdict: FLUX.2 [max] is the clear winner as it maintains the integrity of both source images, perfectly preserving the car's design and the man's distinct hairstyle and clothing. While Vidu Q2 creates a more dramatic action shot, it fails on detail preservation, with warped car features and botched hand anatomy on the steering wheel.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the motion blur and 50mm lens depth of field requests.
  • + Realistic skin texture and natural interaction with the bike tools.
  • + Cinematic lighting and high-quality detail in the rain and pavement reflections.
  • The bike frame geometry is slightly distorted where the front post meets the basket.
  • The car in the background is a bit too sharp to perfectly represent motion blur.

Vidu Q2

  • + Successfully captured the 'imperfect framing' request with a tight, candid crop.
  • + Good skin texture and weathered look on the bike frame.
  • Significant anatomical errors with the hands, including an impossible grip on the handlebars.
  • The car in the background lacks the requested motion blur.
  • The bike's mechanical structures (chain and pedals) are physically incoherent.

Verdict: FLUX.2 [max] is the clear winner as it produces a high-quality, believable image that hits almost every prompt requirement, including the complex request for motion blur. While Vidu Q2 captures the 'imperfect framing' well, it suffers from severe anatomical and mechanical distortions that break the realism required by the prompt.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Exceptional skin texture with realistic dirt and faint scarring.
  • + Intricate, highly detailed engraving on the armor with subtle wear and tear.
  • + Naturalistic bokeh and professional shallow depth of field.
  • The lighting is a bit flat compared to the dramatic highlights requested.
  • Beads in the hair are numerous but somewhat small and blend in.

Vidu Q2

  • + Strong prompt adherence regarding the warm torchlight reflections on the metal.
  • + Excellent representation of the leather straps and padded cloth underlayer.
  • + Dynamic lighting that creates high contrast and visual interest.
  • Skin texture appears slightly too smooth and 'digital' for the character's description.
  • The depth of field is less shallow, making the background slightly more distracting.

Verdict: FLUX.2 [max] excels in photorealistic skin textures and fine engraving details, capturing a more believable 'battle-worn' ruggedness. Vidu Q2 provides better lighting effects and clear depiction of secondary materials like leather and cloth, but the character's facial rendering feels more like a CGI render than a lifelike portrait.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent typography with legible bold sans-serif headers
  • + Clean and logical grid layout for food photos
  • + Perfectly followed section instructions for Appetizers, Pizza, and Mains
  • Nonsense filler text for descriptions
  • Pricing is unrealistically low (e.g., 2.90 for a pizza)

Vidu Q2

  • + Bright and vibrant food photography
  • + High-key, energetic color palette
  • Poor typography with significant character distortion and spelling errors
  • Chaotic layout that fails to clearly separate sections as requested
  • Aesthetics feel cluttered and less professional than the prompt asked for

Verdict: FLUX.2 [max] significantly outperforms Vidu Q2 by delivering a cohesive, professional design that looks like a real menu. FLUX.2 [max] maintained a clean grid and legible headers, whereas Vidu Q2 suffered from illegible text and a disorganized layout that did not clearly respect the requested categories.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent typography with clean, professional-looking effects
  • + High photorealistic detail in the food textures
  • + Perfect alignment with the requested starry burst and price format
  • The 'exploded' effect is a bit condensed compared to the prompt's potential

Vidu Q2

  • + Strong sense of heat and fire in the background
  • + Good dynamic energy with the flying sauce droplets
  • The currency symbol is incorrect, looking more like an 'E' or 'F' with extra bars than a Euro symbol
  • The burger ingredients look slightly more illustrative and less photorealistic than Image A

Verdict: FLUX.2 [max] produced a superior result that looks like a finished professional advertisement, following every instruction including the specific currency symbol and starburst design. Vidu Q2 captured a more intense fiery atmosphere but struggled with the technical execution of the text and the photorealism of the food.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent text rendering with perfect spelling across all lines.
  • + Highly realistic chalk texture with smudge marks and dusty details.
  • + Strong prompt adherence regarding the layout, price points, and specific menu items.
  • The 'natural variations' in handwriting are very subtle, appearing slightly uniform across words.

Vidu Q2

  • + Natural variation in lettering size and slant that feels authentic to a human hand.
  • + Good chalk lighting and layering effect.
  • Severe spelling errors and garbled text in every menu item.
  • Incorrect price for the first item ($34 instead of $24).
  • Text becomes illegible and repetitive at the bottom of the board.

Verdict: FLUX.2 [max] is the clear winner as it followed every instruction, including the specific date and price points, while maintaining perfect legibility. In contrast, Vidu Q2 suffered from significant spelling hallucinations and failed to render the requested text accurately despite capturing a nice chalky aesthetic.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Strong resemblance to the character face and hair from Image 2
  • + Excellent lighting match to the warm yellow environment of Image 1
  • + Clean rendering of feet and toes
  • Failed to replicate the complex leg-crossing pose from Image 1
  • The scarf pattern is simplified compared to the source

Vidu Q2

  • + Successfully replicated the exact crossed-leg pose from Image 1
  • + Accurately captured the specific scarf pattern and clothing details from Image 2
  • + Maintains high facial fidelity to the character in Image 2
  • The transition between the dark sweatshirt and the skin on the legs is a bit sharp/unnatural
  • The hands are slightly stylized with long fingers

Verdict: Vidu Q2 is the clear winner as it successfully combined the complex physical geometry of Image 1 (specifically the crossed legs and torso lean) with the character identity of Image 2. FLUX.2 [max] failed to replicate the specific pose, reverting to a more standard standing position that missed the dynamic 'exact pose' requirement.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent anatomical realism for both the horse and the astronaut.
  • + High cinematic quality with sophisticated lighting and background galaxy details.
  • + Sophisticated composition with the asteroid providing a grounding element in space.
  • The astronaut is notably small in scale compared to the horse.

Vidu Q2

  • + Creative use of color, incorporating nebulae patterns into the horse's coat.
  • + Good sense of motion and dynamic energy.
  • + Adheres to the prompt's request for a surreal atmosphere.
  • The horse has severe anatomical issues, including five legs or poorly rendered leg-like appendages.
  • The image style leans more toward digital illustration than a cinematic photograph.

Verdict: While both models successfully put the rider on top of the horse, FLUX.2 [max] is the superior image due to its incredible anatomical accuracy and cinematic lighting. Vidu Q2 has a more vibrant color palette but fails on basic visual coherence, most notably by giving the horse an extra leg and having muddier textures.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent preservation of the original person's face, hair, and vitiligo patterns
  • + High fidelity to the requested outfit, specifically the scarf pattern and coat style
  • + Natural-looking hands and realistic lighting and shadows
  • None notable

Vidu Q2

  • + Successfully included the accessories like the sunglasses and watch
  • + Good composition that expands the original frame slightly
  • Failed to preserve the original person's face, adding a mustache and changing facial features
  • Severe anatomical issues with the left hand (deformed fingers)
  • Poor lighting integration with the background

Verdict: FLUX.2 [max] followed the complex instructions perfectly, successfully dressing the original subject in the new outfit while keeping his face and identity completely intact. Vidu Q2 failed the most critical part of the prompt by changing the subject's face (adding facial hair) and producing significant anatomical errors in the hands and clothing texture.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent photographic lighting on the human subject's face from the phone screen
  • + High level of fur texture and realistic car interior details
  • Anatomical failure with the capybara having human hands
  • The capybara's head is not looking toward the road

Vidu Q2

  • + Successfully depicted capybara paws on the steering wheel instead of human hands
  • + Better overall composition and framing that captures the city atmosphere
  • + The capybara's face is oriented toward the glass/road, looking more professional
  • Slightly lower fidelity on the businessman's face compared to the foreground

Verdict: While FLUX.2 [max] has impressive textures and lighting, it fails a key anatomical instruction by giving the capybara human hands. Vidu Q2 followed the prompt more accurately by rendering paws on the steering wheel and created a more coherent scene that feels like a professional driver at work.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Perfect text rendering for all requested strings.
  • + High-quality, cinematic lighting and realistic textures.
  • + Excellent composition with a balance of thorns, webs, and background depth.
  • The parchment effect is stylized as outer edges rather than the primary background for all text.

Vidu Q2

  • + Stronger 'vintage parchment' aesthetic consistent across the entire background.
  • + Captures the requested border elements clearly.
  • Contains multiple significant spelling errors in the title and scroll.
  • Incorrect event details such as the date and time strings.
  • Lower overall visual realism compared to the other model.

Verdict: FLUX.2 [max] significantly outperforms Vidu Q2 by following all text instructions perfectly, whereas Vidu Q2 fails on almost every line of text (e.g., 'Intovztion', 'invieed', and incorrect dates). FLUX.2 [max] also provides a more polished, atmospheric, and professional-looking cinematic invitation.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [max]
Before After
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent source preservation of the background and clothing.
  • + The hair texture and volume feel natural for the subject's age and ethnicity.
  • The forehead seems slightly elongated to accommodate the new hairline.
  • The transition from the sideburns to the top hair is a bit abrupt.

Vidu Q2

  • + Natural integration of the hairline with existing facial features.
  • + Consistent lighting on the hair that matches the overall scene.
  • There is a noticeable change to the jacket's collar area compared to the source.
  • Hair texture at the very top appears slightly feathered/digital and less coarse than the beard.

Verdict: Both models performed remarkably well at adding hair while maintaining the identity of the person. FLUX.2 [max] did a better job of preserving the source image's background and clothing exactly, whereas Vidu Q2 slightly altered the jacket collar, though its hairline placement feels more anatomically grounded.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the isometric diorama request
  • + Clean and legible text rendering
  • + Highly accurate representation of different sushi types in a cohesive style
  • The salmon nigiri texture is slightly repetitive

Vidu Q2

  • + Vibrant colors and high-gloss PBR materials
  • + Good layout of the central sushi plate
  • Failed the isometric diorama perspective, providing a flat perspective instead
  • Text rendering is slightly inconsistent with the N in JAPAN appearing stylized incorrectly
  • The flag icon is clumsily attached to the letter N

Verdict: FLUX.2 [max] followed the prompt instructions perfectly, delivering a true 45° isometric diorama with clean text and a professional 3D miniature aesthetic. Vidu Q2 ignored the isometric layout and diorama base request, providing a standard frontal view with less refined text integration. FLUX.2 [max] is the clear winner for its structural accuracy and superior adherence to the 'clean' and 'isometric' constraints.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Successfully incorporates all elements: hockey rink, dogs, and a news anchor desk.
  • + High quality caricature style that maintains the subject's likeness.
  • + Excellent composition with symmetrical elements and creative branding like 'JEMER NEWS'.
  • The eyes look slightly more generic/cartoonish compared to the source's eyes.

Vidu Q2

  • + Extremely high fidelity to the original person's facial features and hair.
  • + Creatively integrates the small dog in a sports jersey and paw-print notes on the desk.
  • + Smart preservation of the original outfit (denim shirt and black t-shirt).
  • Physical anatomy issues, specifically the six-fingered hand on the left side of the image.
  • The hand on the right is detached from the arm and floating near the small dog.

Verdict: Both models did an excellent job of capturing the subject's likeness and incorporating the specific requests for hockey, dogs, and a TV anchor profession. FLUX.2 [max] produced a more polished, error-free composition with a great newsroom-at-the-rink atmosphere, while Vidu Q2 captured a more accurate facial likeness but suffered from significant anatomical errors like a six-fingered hand and a floating detached hand.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent adherence to the list of specific animals with high anatomical accuracy.
  • + Superior lighting effects including soft god rays and realistic dew sparkles on the grass.
  • + Cohesive composition with a clear focus on the interaction between the animals.
  • The cat's front right paw looks slightly disconnected from the body.

Vidu Q2

  • + Bright, vibrant colors and high energy in the character poses.
  • + Effective use of butterflies to fill the sky and add to the playful theme.
  • Included two golden retriever puppies instead of one, deviating from the prompt.
  • Noticeable anatomy issues, such as the fox kit's front leg being detached from the body.
  • The background lighting is quite blown out, losing detail in the sky.

Verdict: FLUX.2 [max] followed the prompt more accurately by including exactly one of each requested animal, whereas Vidu Q2 added an extra puppy. FLUX.2 [max] also achieved a more photorealistic look with sophisticated lighting and better structural integrity for the animals' anatomy.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Perfectly captures the soft, muted pastel palette typical of Studio Ghibli films
  • + Excellent hand-painted texture with a subtle paper grain effect
  • + Preserves the composition and poses of the source image with high fidelity
  • The facial expressions are slightly softened, losing some of the sharp 'distracted' humor

Vidu Q2

  • + Maintains clear, vibrant lines that resemble high-quality anime cel shading
  • + Preserves the exact emotional expressions and facial features of the original subjects more accurately
  • + Good integration of the background architectural details in an illustrative style
  • The texture feels more digital and 'clean' than the requested hand-painted Ghibli style
  • The lighting is less 'gentle' and 'dreamy' compared to Model A

Verdict: Both models do an excellent job of translating the 'Distracted Boyfriend' meme into an anime style while preserving the source composition. FLUX.2 (max) is the winner as it better adheres to the specific aesthetic prompts for Ghibli, including the paper texture and soft pastel color grading, whereas Vidu Q2 produces a more standard modern anime look.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [max]
Before After
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent preservation of the subject's face and original features
  • + Highly realistic wind-blown hair effect that looks naturally integrated
  • Several leaves have a blurry, ghost-like appearance that looks like a technical artifact
  • The leaves are mostly clustered on one side of the frame

Vidu Q2

  • + Strong 'energetic' feel with leaves distributed throughout the entire depth of the image
  • + Vibrant colors on the falling leaves create high visual interest
  • Noticeable alterations to the subject's face compared to the source image
  • Some leaves appear to be floating statically rather than moving briskly

Verdict: FLUX.2 [max] succeeded in providing a very realistic hair-blowing effect while perfectly preserving the woman's identity, but its leaf artifacts are distracting. Vidu Q2 created a more lively composition with leaves filling the frame, however, it changed the woman's facial features and some leaves lack the desired dynamic motion. FLUX.2 [max] is the preferred winner for its superior source preservation and more convincing hair physics.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Perfect text rendering for 'Caffè Florian' and 'Est. 1720'
  • + Clean, vintage vector aesthetic that matches the 'minimalist' prompt
  • + Balanced composition within a classic circular emblem
  • The steam lines are very simple/thin compared to the rest of the illustration

Vidu Q2

  • + Warm color palette and golden tones feel upscale
  • + Visualizing steam through negative space is an interesting creative choice
  • Significant spelling errors including 'Caffe Farmiin' and 'Esttt'
  • Messy, repetitive text at the bottom of the logo
  • Lacks the minimalist refined feel requested by the prompt

Verdict: FLUX.2 [max] is the clear winner as it perfectly follows the typographic requirements and minimalist aesthetic requested in the prompt. Vidu Q2 fails on basic text generation, producing several misspellings and redundant text elements that clutter the design.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [max]
Vidu Q2

AI Judge Analysis

FLUX.2 [max]

  • + Excellent typography with correct spelling of mission stages and crew names.
  • + Clean, professional vector aesthetic that perfectly matches the 'flat-vector' request.
  • + Logical flow of information from launch to landing using consistent iconography.
  • The Saturn V silhouette is slightly stylized towards a generic space shuttle look.
  • The Earth Orbit label is repeated unnecessarily at the top of the first and second panels.

Vidu Q2

  • + Successfully captured the requested color palette and flat icon style.
  • + Interesting layout that attempts to separate stages into a grid format.
  • Extremely poor text rendering with significant misspellings throughout the image.
  • Inconsistent and confusing logical flow that does not follow the requested 6-step sequence.
  • The icons are repetitive and do not clearly distinguish between descent and landing stages.

Verdict: FLUX.2 [max] produced a high-quality, usable infographic with near-perfect text accuracy and a professional layout that follows the requested narrative steps. In contrast, Vidu Q2 failed significantly on text rendering and logical coherence, resulting in a 'lorem ipsum' style output that does not clearly communicate the mission steps. FLUX.2 [max] is the clear winner for its clear communication and adherence to the technical requirements of the prompt.

Next steps

Explore each model