Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [dev] Black Forest Labs Vidu Q2 ShengShu Technology

Settled by community votes across 17 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [dev]

16.5 arena score

#58 of 62 in Text-to-Image

Skill signature · Text-to-Image

Vidu Q2

19.8 arena score

#42 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [dev]

0%

win rate

Ties

0%

Vidu Q2

0%

win rate

Shared challenges 17

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent adherence to the 'soft window light from the left' instruction.
  • + Clean, modern aesthetic with high photographic realism.
  • + Accurate reflections on the glass and the surface of the sphere.
  • The glass cube lacks a physical bottom face, looking more like a 5-sided cover.
  • The reflection of the sphere on the bottom looks like a mirror rather than wood through glass.

Vidu Q2

  • + Strong spatial composition and realistic wood grain texture.
  • + Good rendering of the glass cube's edges and physical thickness.
  • + Effective 'partially visible through glass' effect for the plant.
  • The lighting is harsh and dappled, contradicting the 'soft window light' request.
  • The light direction is inconsistent with the shadows on the table.

Verdict: Both models followed the complex spatial instructions perfectly. FLUX.1 Kontext [dev] is the winner because it captured the specific 'soft' quality of the window light requested in the prompt, whereas Vidu Q2 produced a much higher-contrast, direct sunlight effect with dappled shadows. FLUX.1 also featured more realistic physics in the sphere's surface reflections.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent full-body composition and environmental storytelling
  • + Very realistic skin textures and facial details
  • + Accurate representation of light rain and wet pavement reflections
  • The man is posing with the bike rather than actively repairing it
  • Missing the requested motion blur on the background cars

Vidu Q2

  • + Perfect adherence to the 'repairing' action and 'imperfect framing' prompt
  • + Exceptional detail in the weathered skin and hyper-realistic hands
  • + Naturally integrated motion blur and candid feel
  • The bicycle frame geometry is slightly distorted/nonsensical near the front wheel
  • Stronger 'digital' sharpness may feel less like a 50mm film shot

Verdict: While FLUX.1 Kontext [dev] creates a beautiful and clean image, it fails the primary action of 'repairing' the bicycle. Vidu Q2 captures the essence of a candid street photograph much better, showing intense focus on the repair task with authentic skin textures and the requested imperfect framing.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent warm lighting that creates a dramatic atmosphere.
  • + Strong rendering of the engraved plate armor texture.
  • + Effective use of bokeh and shallow depth of field.
  • Failed to include the hair braided with beads mentioned in the prompt.
  • The 'battle-worn' effect feels a bit clean despite the small scar.

Vidu Q2

  • + Perfectly captured the braided hair with small beads.
  • + Excellent detail on the leather straps and underlayer fabric.
  • + More convincing 'battle-worn' aesthetic with dirt and varied metallic textures.
  • The lighting is a bit busy compared to the focused torchlight of the other image.
  • Slightly less 'close' as a portrait compared to Model A.

Verdict: While FLUX.1 Kontext [dev] produced a more dramatic and cinematic lighting effect, Vidu Q2 followed the prompt much more accurately, specifically including the braided hair with beads which Model A missed entirely. Vidu Q2 also excelled at rendering the specific textures of leather straps and cloth requested in the prompt.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent grid-based composition that feels high-end and modern
  • + The food photography looks professional and consistent in lighting.
  • + Bold, clean sans-serif typography that is very legible.
  • The text is largely gibberish and does not follow the requested sections accurately.
  • The font rendering has slight artifacts on some characters.
  • The food photos, while high quality, don't clearly represent 'pizza' well.

Vidu Q2

  • + Successfully includes the specific sections for pizza, appetizers, and mains.
  • + Includes price indicators which makes it look more like a functional menu.
  • + Colorful accents add a playful casual dining feel.
  • The text is highly distorted and contains many nonsensical characters.
  • Visual layout is cluttered with inconsistent spacing.
  • Overall image quality is blurry compared to the alternative.

Verdict: FLUX.1 Kontext [dev] produces a much higher quality image with professional lighting and a sleek, high-end editorial layout that perfectly captures the 'modern minimalist' aesthetic. Vidu Q2 followed the prompt instructions for specific menu sections (Pizza, Mains) better, but the execution suffered from poor image clarity and significant text distortion.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent typography rendering for 'MAGIC BURGER'.
  • + Good photorealistic texture on the meat patty.
  • + Clean layout suitable for a social media ad.
  • Failed the main request for an 'exploded' burger with components suspended in mid-air.
  • Significant typo in 'ONLY' (rendered as 'LNHLY').
  • The starburst element is a flat graphic rather than a glowing, integrated effect.

Vidu Q2

  • + Successfully followed the instruction for an 'exploded' burger with suspended components.
  • + High sense of motion with sauce droplets and flying embers.
  • + Better integration of the fiery, glowing effect on the text and starburst.
  • The currency symbol is incorrect (rendered as a stylized hash instead of Euro).
  • The 'LIMITED TIME ONLY' text is slightly less legible due to the glow intensity.
  • The bottom bun orientation looks slightly awkward.

Verdict: Vidu Q2 is the clear winner as it successfully interpreted the 'exploded burger' requirement which defines the motion of the image, whereas FLUX.1 Kontext [dev] generated a static burger. While Vidu Q2 missed the specific Euro symbol, its overall composition, internal text rendering, and adherence to the dynamic prompt create a much more effective advertisement.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent chalk texture and realistic handwriting aesthetic
  • + Consistent formatting across all lines
  • + High clarity and contrast on the lettering
  • Numerous spelling errors including 'Mashroom', 'Risoktso', and 'Octpus'
  • Mangled the month in the date, rendering it as gibberish
  • The 'Today Specials' header is blocky rather than 'elegant cursive' as requested

Vidu Q2

  • + Successfully spelled 'APRIL 30, 2026' correctly
  • + Good cursive style on the header text
  • + Realistic chalk smudging and atmospheric café background
  • Significant spelling hallucinations in the menu items like 'Octopd wpila' and 'Browd Botter'
  • Inconsistent line spacing and messy composition
  • The bottom text rows become complete illegible gibberish

Verdict: FLUX.1 Kontext [dev] provides a much cleaner and more legible image with superior chalk textures, but struggles significantly with spelling several requested words. Vidu Q2 follows the date formatting more accurately and uses a cursive style as requested, but the overall coherence and spelling for the menu items are very poor. FLUX.1 Kontext [dev] is the preferred choice for its professional layout and aesthetic, despite the typos.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Successfully followed the difficult spatial instruction of placing the horse on top of the astronaut.
  • + High realism in textures and lighting on the spacesuit.
  • + Distinct surrealist composition that aligns with the prompt's intent.

Vidu Q2

  • + Very vibrant and colorful cosmic aesthetic.
  • + High artistic detail in the horse's nebula-like fur.
  • Failed the core prompt instruction of placing the horse on top of the astronaut.
  • Cliche interpretation of 'astronaut on horse' despite the specific reversal request.

Verdict: The main differentiator is adherence to the specific 'horse on top' instruction. FLUX.1 Kontext [dev] followed this surreal prompt perfectly, creating a unique and technically sound image, whereas Vidu Q2 ignored the instruction and produced a standard astronaut-riding-horse image.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent texture on the capybara's fur
  • + Subtle and professional expression on both characters
  • + Rich, atmospheric lighting consistent with a New York night
  • Only one paw is on the steering wheel while the other is on its lap
  • The yellow cap is a modern baseball style rather than a traditional driver cap

Vidu Q2

  • + Successfully placed both paws on the steering wheel as requested
  • + The driver cap is a more traditional and recognizable taxi driver style
  • + Clear and vibrant Manhattan backdrop with bokeh effects
  • The passenger's phone is awkwardly merged with a clipboard or folder
  • The capybara's hands look slightly more primate-like than authentic rodent paws
  • The overall lighting is a bit flat compared to Model A

Verdict: Both models followed the prompt well, but Vidu Q2 adhered more closely to the specific instruction of having both paws on the steering wheel and provided a more iconic driver's cap. However, FLUX.1 Kontext [dev] achieved a much higher level of photorealism and artistic atmosphere despite the limb placement error.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent typography for the main title and specific dates.
  • + Clean graphic design style with a focused central subject.
  • + Includes the requested thorn border and twisted tree silhouettes.
  • The text on the scroll banner is illegible and garbled.
  • The background is solid black rather than the dark parchment requested.
  • Misspelled the location 'The Arches' as 'The Argiiah's'.

Vidu Q2

  • + Successfully captures the dark parchment texture and moody night sky background.
  • + The gothic title font perfectly matches the requested aesthetic.
  • + Follows complex border instructions including webs, thorns, and twisted trees.
  • Multiple spelling errors in the title, banner, and date ('Intovztion', 'invieed', '30.70.2025').
  • The glow from the pumpkin feels slightly disconnected from the parchment surface.
  • Resolution and detail on the pumpkin are less polished than in Model A.

Verdict: Both models struggled with the complex task of rendering three separate sections of text accurately. FLUX.1 Kontext [dev] produced a cleaner, more legible main title and date, but failed to provide the requested parchment background, while Vidu Q2 captured the gothic atmosphere and parchment texture beautifully but failed significantly on text accuracy and spelling.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.1 Kontext [dev]
Before After
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Successfully added a dense head of hair.
  • + Maintained high image resolution and clarity.
  • Significantly altered the facial structure, making the man look younger and like a different person.
  • The hair texture looks slightly artificial and overly groomed for the context.
  • Failed to preserve original facial features.

Vidu Q2

  • + Excellent source preservation, keeping the man's face and features identical to the original.
  • + The hair texture and style match the rugged aesthetic of the scene.
  • + Seamless integration with the existing beard and glasses.
  • The hairline on the forehead is slightly soft/blurry upon close inspection.

Verdict: While both models followed the instruction to add hair, FLUX.1 Kontext [dev] fundamentally failed the image editing requirement by changing the person's face entirely. Vidu Q2 successfully added realistic, natural-looking hair while perfectly preserving the identity, facial features, and lighting of the man in the source image.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent 3D toy-like textures and clean PBR materials.
  • + Perfectly follows the minimal garnish and 'less is more' aesthetic.
  • + Very clean typography that is well-integrated with the layout.
  • The 'flag icon' is unrecognizable and abstract.
  • The sushi design is overly simplified, looking like a plastic toy rather than a miniature food model.

Vidu Q2

  • + Features a correct and recognizable Japanese flag icon.
  • + Higher level of detail in the food textures while maintaining the cartoon 3D style.
  • + Includes the requested 'plate' on top of the diorama base.
  • Lighting is a bit harsh on the white plate areas.
  • The '45° top-down' angle is slightly flatter than ideal isometric projection.

Verdict: While FLUX.1 Kontext [dev] has a very clean and professional graphic design feel, Vidu Q2 is the overall winner for its superior attention to the specific prompt details. Vidu Q2 successfully included the plate, a correct flag icon, and more appetizing food textures, whereas FLUX.1 Kontext [dev] failed the flag requirement and was perhaps too minimal in its representation of sushi.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Successfully translates the subject's features into a bold comic-book style caricature.
  • + Includes a TV screen and a cartoon dog as requested.
  • + Preserves the selfie-style composition of the original photo.
  • Completely misses the hockey requirement.
  • The text 'JOB' and the bottom banner are nonsensical or unnecessary.
  • The clothing details like the shirt buttons are slightly inconsistent.

Vidu Q2

  • + Excellent adherence to all prompt elements, including the hockey rink background, puck, and news microphone.
  • + High visual quality with a professional digital illustration style.
  • + Strong composition that balances the profession (news desk), hobby (hockey), and dogs effectively.
  • Anatomical issues with the hands, including an extra jointed finger and a tiny hand supporting the dog.
  • The caricature style is less exaggerated and more 'portrait illustration' than a traditional caricature.

Verdict: Vidu Q2 is the clear winner for its thorough adherence to the prompt, successfully integrating the TV anchor, hockey, and dog elements into a cohesive scene. While FLUX.1 Kontext [dev] captured the individual's likeness well in a comic style, it failed to include the hockey theme and had less imaginative composition.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent fur texture and lighting on the animals.
  • + Clean composition with a beautiful bokeh effect.
  • Failed to include the bunny and fox kit entirely.
  • Identified as a tabby kitten but rendered a white/pale orange kitten.

Vidu Q2

  • + Successfully included all four animal types requested: puppy, kitten, bunny, and fox.
  • + Beautifully rendered 'god rays' and dew sparkles that match the prompt perfectly.
  • + Better adherence to the 'lush wildflower meadow' description with varied flower types.
  • Anatomy on the puppy on the right is a bit stiff/awkward.
  • Some butterfly renderings are a bit repetitive in placement.

Verdict: Vidu Q2 is the clear winner for prompt adherence, successfully including the puppy, kitten, bunny, and fox kit while FLUX.1 Kontext [dev] missed half of the requested animals. Vidu Q2 also better captured the atmospheric elements like god rays and dew sparkles, creating a more complete interpretation of the scene.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent preservation of the original image's composition and structure
  • + Strong anime aesthetic that matches the characters' features well
  • + Clean line work and vibrant colors
  • Looks more like modern digital anime than the specific hand-painted Studio Ghibli style
  • The lighting is flat compared to the requested dreamy atmosphere

Vidu Q2

  • + Stronger adherence to the 'Studio Ghibli' request with visible hand-painted watercolor textures
  • + Uses the requested soft pastel color palette and gentle lighting
  • + Preserves the original composition while successfully stylized
  • The man's expression is slightly less faithful to the original 'distracted' look than Model A
  • Linework is a bit sketchier, though this fits the requested style

Verdict: While both models successfully converted the meme into an illustration, Vidu Q2 followed the stylistic instructions much more closely by incorporating hand-painted textures and soft pastel tones characteristic of Studio Ghibli. FLUX.1 Kontext [dev] produced a very clean image but it feels more like a standard modern digital anime than the specific nostalgic, painted look requested.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.1 Kontext [dev]
Before After
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent preservation of the original subjects' faces and environment.
  • + Subtle, realistic hair motion that integrates naturally with the scene.
  • Very few leaves added, failing to meet that part of the prompt effectively.
  • The leaves that are present appear as small, green specks that lack motion blur.

Vidu Q2

  • + Strong adherence to the 'leaves flying' prompt with high volume and color variety.
  • + Good motion effect in the hair that suggests a strong breeze.
  • Noticeable distortion in the facial features of the woman, particularly the eyes and mouth.
  • Significant artifacts around the dog's mouth and the flowers on the left.
  • The leaves lack motion blur, appearing as static overlays.

Verdict: FLUX.1 Kontext [dev] produced a much higher quality image that preserved the integrity of the original source, though it was too subtle with the requested environmental changes. Vidu Q2 succeeded in adding many leaves and dynamic hair, but it severely degraded the visual quality of the faces and introduced several artifacts. FLUX.1 Kontext [dev] is the winner for creating a usable, high-quality edit, despite the missed prompt details.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Excellent typography with perfect spelling of all requested text.
  • + Clean minimalist vector aesthetic that fits a modern-vintage brand Identity.
  • + Accurate interpretation of a cloche dome using simple, bold lines.
  • Does not strictly include a 'banner' for the 'Est. 1720' text.
  • The steam element is very abstract and slightly disconnected.

Vidu Q2

  • + Includes the requested banner for the established date.
  • + Features a more literal representation of a retro cloche and steam.
  • + Captures the subtle paper texture on the background well.
  • Severe spelling errors throughout the logo ('FARMIIN', 'Esttt', 'FOPLI20').
  • Busy composition with redundant and hallucinated text at the bottom.
  • Cluttered design that lacks the 'minimalist' quality requested in the prompt.

Verdict: FLUX.1 Kontext [dev] produced a professional, usable logo with perfect spelling and a clean minimalist aesthetic, though it skipped the banner element. Vidu Q2 followed the layout instructions more literally by including a banner, but failed significantly on professional quality due to numerous spelling errors and a cluttered composition. FLUX.1 Kontext [dev] is the clear winner for its clarity and accuracy.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [dev]
Vidu Q2

AI Judge Analysis

FLUX.1 Kontext [dev]

  • + Stronger adherence to the navy blue color palette.
  • + Achieves a more 'authentic' modern flat vector graphic design look.
  • Poor spelling throughout the poster, including 'APOLO'.
  • Icons are abstract and messy, failing to clearly represent the specific mission steps requested.

Vidu Q2

  • + Successfully follows the requested infographic steps from launch to landing.
  • + Crisp and clear iconography that matches the flat-vector style requirements.
  • + Contains a more logical layout and clearer supporting elements like astronauts.
  • Gibberish text in place of labels, though it manages to capture the essence of words like 'Tranquility'.
  • Mismatched numbering for the steps which contradicts the visual progression.

Verdict: Vidu Q2 is the clear winner for its superior ability to visualize the requested infographic steps with recognizable icons that actually tell a story, whereas FLUX.1 Kontext [dev] produced abstract shapes that barely relate to the prompt. While both models struggled significantly with text legibility, Vidu Q2 followed the logical progression from rocket to lunar landing much more effectively.

Next steps

Explore each model