Head to head
Esc

Models · slot A

to navigate to pick

FLUX.2 [klein] 9B Black Forest Labs Vidu Q2 ShengShu Technology

Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.

FLUX.2 [klein] 9B

20.6 arena score

#13 of 32 in Image Editing

Skill signature · Image Editing

Vidu Q2

19.8 arena score

#42 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.2 [klein] 9B

0%

win rate

Ties

0%

Vidu Q2

0%

win rate

Shared challenges 20

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent adherence to lighting instructions with a soft, natural window light appearance.
  • + High photographic realism in the wooden table texture and glass reflections.
  • + Accurate spatial placement of all objects including the internal sphere.
  • The glass cube has an unusual double-walled or rounded internal design rather than sharp edges.

Vidu Q2

  • + The sphere is distinctly blue and centered well.
  • + The plant detail is very sharp and vibrant.
  • + The book has realistic gold-leaf detailing on the spine.
  • The lighting creates harsh shadows rather than the 'soft window light' requested.
  • Perspective issues where the book seems to float slightly off the back edge of the cube.
  • Strong teal tint on the glass edges feels less natural.

Verdict: Both models followed the complex spatial prompt perfectly. FLUX.2 [klein] 9B is the winner due to its superior handling of the 'soft window light' and its more realistic integration of the glass cube into the environment, whereas Vidu Q2 produced harsh, high-contrast shadows and slightly disconnected object physics.

Man and Car in California

Editing
Edit instruction

“Make a photo of the man driving the car down the California coastline”

Source
FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the original car's proportions and details
  • + Realistic lighting and motion blur on the road
  • + Perfect adherence to the California coastline setting with palm trees and beach
  • The man is barely visible, with only his hair popping over the driver's seat

Vidu Q2

  • + Successfully places the specific man from the source image in the driver's seat
  • + Captures an iconic highway 1 coastline aesthetic perfectly
  • + Maintains the car's identity and color
  • The car's front-end geometry is slightly warped compared to the source
  • The man's scale and positioning in the seat feel slightly unnatural

Verdict: Both models followed the instructions well, but Vidu Q2 is the more successful editor as it clearly shows the man from the source image driving, whereas FLUX.2 only shows the back of his head. While FLUX.2 preserved the car's technical details slightly better, Vidu Q2 delivered on the narrative requirement of the prompt much more effectively.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent skin texture and realistic age-related details on the face and hands.
  • + Strong adherence to the 'candid' and 'imperfect framing' prompt with a balanced cinematic composition.
  • + Accurate rendering of light rain and atmospheric depth.
  • The cars in the background lack motion blur, appearing relatively sharp.
  • The bicycle's brake cables are slightly disconnected or floating near the handlebars.

Vidu Q2

  • + Successfully captures the requested 'imperfect framing' with a tight, off-set crop.
  • + Includes the requested motion blur on the passing vehicle in the background.
  • Severe anatomical errors with the subject's hands and fingers.
  • The bicycle geometry is nonsensical, particularly the relationship between the handlebars and the frame.
  • The image has a more 'AI-generated' plastic look compared to the requested realism.

Verdict: FLUX.2 [klein] 9B produces a significantly more high-quality and realistic image with believable textures and a coherent scene, though it misses the motion blur requirement. Vidu Q2 attempts a more dynamic composition with motion blur but suffers from major anatomical and structural failures in the hands and bicycle.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Exquisite detail on the engraved armor and weathered leather straps.
  • + Captures the 'battle-worn' aesthetic with realistic scars, dirt, and facial hair.
  • + Perfect execution of multiple braids with beads as specified.
  • The torch in the background looks slightly artificial with a simplified flame shape.

Vidu Q2

  • + Dynamic cinematic lighting with strong highlights on the metal.
  • + Good garment texture on the cloth underlayer.
  • + Includes the requested bokeh sparks in the background.
  • The character looks too pristine and young to fully embody the 'battle-worn' description.
  • Lower detail in the armor engravings compared to Model A.
  • The hair contains only one or two braids rather than being 'braided with small beads' throughout.

Verdict: FLUX.2 [klein] 9B followed the prompt more accurately, particularly regarding the braids, beads, and the 'battle-worn' texture of the skin and armor. While Vidu Q2 has nice lighting, it produces a much cleaner, more 'polished' character that misses the gritty, weathered details requested in the prompt.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Stronger adherence to the clean, professional grid layout requested.
  • + Better font clarity for the main headings compared to the competing model.
  • + High visual quality in the food photography section.
  • Shows significant repetition in the food photos, displaying pizza for every category.
  • Text suffers from character overlapping and artifacts in the title and body text.

Vidu Q2

  • + Includes a much better variety of food types representing the requested sections.
  • + More creative use of vibrant accents and graphic elements while maintaining white space.
  • + The multi-column layout feels more authentic to a casual dining menu.
  • Severe spelling errors and non-existent words in almost every text field.
  • The section headings are inconsistent and some are poorly rendered.

Verdict: Model B (Vidu Q2) is the preferred choice because it captures the 'vibrant accents' and 'colorful food photos' requirements much more effectively with a variety of dishes, whereas Model A (FLUX.2 [klein] 9B) simply repeated pizza images for every category. Although both models struggle with legible text, Model B's layout and asset variety feel much more like a functional menu design for a casual restaurant.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent typography rendering with clean, glowing edges.
  • + Highly professional food photography aesthetic with realistic textures.
  • + Perfect adherence to the starburst and price request.
  • The burger is less 'exploded' than Model B, appearing more like a tall stack than suspended components.
  • Lacks the specific tomato slices requested, opting for a thicker burger stack.

Vidu Q2

  • + Stronger 'exploded' effect with clear separation between all requested ingredients.
  • + Dynamic background with intense, fiery atmosphere that matches the energy of the prompt.
  • + Accurately includes sliced tomatoes suspended in the air.
  • The currency symbol rendered as a distorted hybrid rather than a clear Euro sign.
  • The text 'LIMITED TIME ONLY' is slightly uneven and less integrated than Model A.
  • Some sauce droplets look a bit like plastic rather than liquid.

Verdict: FLUX.2 [klein] 9B produces a much more polished and professional-looking advertisement with superior text rendering and lighting. However, Vidu Q2 followed the 'exploded' instruction more literally, showing distinct layers for all ingredients, though it failed to properly render the Euro currency symbol. FLUX.2 is the winner for its commercial quality and precise adherence to the text requirements.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent text accuracy with perfect spelling for all items.
  • + Highly realistic chalk texture with convincing dust smudges on the board.
  • + Consistently elegant and legible cursive style that matches the prompt.
  • One minor typo in the footnote ('fress' instead of 'fresh').

Vidu Q2

  • + Natural variation in the chalk texture, showing different pressure levels.
  • + Captures the 'cozy café' background requested in the prompt.
  • Numerous spelling errors including 'Truffe Musshoom' and 'Octopd'.
  • The text becomes garbled and illegible toward the bottom of the board.
  • Inconsistent pricing characters with nonsensical symbols at the end.

Verdict: FLUX.2 [klein] 9B is the clear winner as it successfully rendered almost every word perfectly, maintaining a high-quality chalk texture and elegant layout. Vidu Q2 struggled significantly with text generation, resulting in several misspelled words and total incoherence in the bottom half of the image.

Pose & Character Mashup

Editing
Edit instruction

“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”

Source
FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent adherence to the complex leg and foot positioning from the pose reference.
  • + Very high character fidelity including sunglasses, scarf pattern, and facial features.
  • + Accurate preservation of the background lighting and red ottoman.
  • The left hand is rendered as a closed fist instead of the open gesture seen in Image 1.
  • Skin tone on the feet is significantly darker than the face and source image.

Vidu Q2

  • + Successfully captures the character's face, hair, and clothing from the reference.
  • + Accurately recreates the open hand gesture from the pose reference.
  • + Sharp image quality and vibrant colors.
  • Fails significantly on the body position, placing the character in shorts instead of pants and missing the leg cross.
  • The scarf pattern is slightly simplified compared to the source.

Verdict: FLUX.2 [klein] 9B is the clear winner as it successfully replicates the extremely difficult body contortion and leg-crossing pose from Image 1 while maintaining the character identity from Image 2. Vidu Q2 fails to maintain the correct clothing (switching to shorts) and does not correctly crossover the legs, which was the core requirement of the pose reference.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + High visual clarity and realistic textures on the astronaut suit and horse.
  • + Excellent cinematic lighting and planetary details in the background.
  • Completely failed the prompt instruction for a surreal 'horse on top' (horse riding the astronaut) arrangement.
  • Contains nonsensical AI-generated text at the bottom of the image.

Vidu Q2

  • + Dreamy, nebula-like style for the horse that fits the 'surreal' descriptor.
  • + Vibrant color palette and good use of light to create a celestial atmosphere.
  • Completely failed the prompt instruction for a surreal 'horse on top' (horse riding the astronaut) arrangement.
  • Anatomical issues with the horse's legs and the astronaut's seated position.

Verdict: Both FLUX.2 [klein] 9B and Vidu Q2 failed the core 'trick' of the prompt, which requested the horse to be on top of the astronaut rather than the standard astronaut-riding-a-horse trope. FLUX.2 [klein] 9B produced a much more realistic and cinematic image with better textures, whereas Vidu Q2 opted for a more stylized, ethereal aesthetic that suffers from slight anatomical distortions.

Outfit Transfer Challenge

Editing
Edit instruction

“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”

Source
FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the subject's face, skin patterns, and expression from Image 1.
  • + Accurately recreates the pea coat, plaid scarf, and jeans from Image 2.
  • + Maintains the wooden structure and beach background with high fidelity.
  • Added extra jewelry (gold necklaces and bracelets) that were not present in Image 2.
  • Face is slightly smoother than the source image.

Vidu Q2

  • + Successfully captures the outfit elements including the sunglasses from Image 2.
  • + Good integration of the textures from the background onto the coat (sand splashes).
  • Failed to preserve the subject's identity, completely changing the facial features and skin patterns.
  • The pose was altered significantly from Image 1.
  • The coat has strange cut-out sleeves that allow skin to show through incorrectly.

Verdict: FLUX.2 [klein] 9B followed the complex instructions much more effectively, managing to dress the original person in the requested outfit while keeping their identity perfectly intact. Vidu Q2 failed the primary goal of keeping the person's face and hair unchanged, effectively generating a new character that only loosely resembles the original. FLUX.2 also showed better technical execution in how the clothing fits the subject's pose and body shape.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent texture on the capybara's fur and the leather jacket
  • + Stronger sense of 'night city' atmosphere within the car lighting
  • + Highly detailed phone screen glow on the woman's face
  • The woman is clearly sitting in the front passenger seat, not the back seat as requested
  • The text on the driver's cap is nonsensical ('NEW TALA')

Vidu Q2

  • + Correctly placed the passenger in the back seat as per the prompt
  • + Superior composition with a wider view showing more of the taxi and city
  • + The capybara's expression and hand placement are very professional and well-rendered
  • The car interior has some structural inconsistencies, such as the roofline and rearview mirror attachment
  • The lighting on the woman in the back is a bit flat compared to the driver

Verdict: While FLUX.2 [klein] 9B offers higher texture detail and more realistic lighting, it failed a key spatial requirement by placing the woman in the front seat. Vidu Q2 followed the prompt's logical layout much better by putting the businesswoman in the back seat, creating a more convincing narrative scene despite slightly lower textural fidelity.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Perfect text rendering for all lines including dates and locations
  • + Excellent atmospheric lighting and cohesive gothic art style
  • + Professional composition that looks like a finished product
  • The jack-o-lantern is slightly more cartoonish than 'cinematic' style

Vidu Q2

  • + Good interpretation of the 'parchment' texture
  • + Strong color contrast between the background and the paper
  • Major spelling errors in the title and banner text
  • Incorrect factual details like the date and month (30.70.2025)
  • Graphic design elements are cluttered and less polished

Verdict: FLUX.2 [klein] 9B followed every instruction perfectly, producing a professional-grade invitation with flawless text rendering and a beautiful atmosphere. In contrast, Vidu Q2 failed significantly on text accuracy, with nonsense words and incorrect dates, and the overall composition felt much less cohesive.

Bald man challenge

Image Editing
Edit instruction

“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”

Before After
FLUX.2 [klein] 9B
Before After
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the source image's background, clothing, and facial features.
  • + The curly hair texture feels organic and consistent with the beard texture.
  • + Natural flyaway hairs add realism to the edge of the silhouette.
  • The hairline transition on the forehead is a bit sharp in some areas.
  • Slight alteration of the glasses' temple where they meet the ear.

Vidu Q2

  • + Natural-looking hairline with a soft, realistic transition from the forehead.
  • + Good preservation of the original jacket, lighting, and environmental details.
  • + Texture of the hair matches the rugged aesthetic of the character.
  • Slightly less 'full and thick' than Image A, looking more like a standard haircut.
  • Some slight smudging artifacts at the top center of the hair.

Verdict: Both models performed excellently, perfectly preserving the source image's identity, lighting, and background while adding realistic hair. FLUX.2 [klein] 9B followed the prompt for 'thick/full' hair more literally with a voluminous curly style, while Vidu Q2 provided a more subtle, natural hairline integration.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Clean, professional typography that perfectly follows the layout instructions.
  • + Superior rendering of wood textures and lighting for a high-quality 3D look.
  • + Consistent isometric perspective and minimal, balanced composition.
  • The flag icon is incorrect, showing a tricolor flag instead of the Japanese flag.

Vidu Q2

  • + Accurately includes the correct Japanese flag icon.
  • + Bright, vibrant colors that fit the 'cartoon' aspect of the prompt.
  • + Good placement of the requested text elements.
  • The sushi rendering is messy with blob-like textures and nonsensical ingredients.
  • The base platform is less defined and lacks the realistic PBR material quality requested.
  • The 45-degree isometric perspective is less geometrically precise than the competition.

Verdict: FLUX.2 [klein] 9B is the superior image in terms of rendering quality, text clarity, and overall aesthetic appeal, even though it failed to produce the correct Japanese flag. Vidu Q2 captured the flag correctly and followed the general layout, but the visual quality of the sushi and the textures was significantly lower and less defined.

Over-the-top cartoon caricature

Editing
Edit instruction

“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”

Source
FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent caricature style with facial features that strongly resemble the subject.
  • + Successfully incorporates all elements: TV anchor desk, hockey arena background, and many dogs.
  • + High-quality comic book aesthetic with great attention to background detail.
  • The hands, while stylized, have slightly messy finger counts on the left hand.
  • Television screens contain nonsensical text.

Vidu Q2

  • + Includes clear hockey elements like a goal, rink markings, and a hockey puck.
  • + Preserves the subject's denim shirt clothing from the original photo.
  • + Good facial likeness within the caricature art style.
  • Significant anatomical errors in the hands, particularly the left hand with six fingers.
  • The microphone looks more like a classic crooner mic than a TV anchor's setup.
  • The composition feels a bit cluttered with overlapping scales.

Verdict: FLUX.2 [klein] 9B followed the prompt more effectively by creating a complete scene that feels like a professional TV broadcast from a hockey arena, full of humorous dog fans. Vidu Q2 succeeded in preserving the subject's clothing and included discrete hockey items like a puck and goal, but was held back by significant anatomical errors in the hands. FLUX.2 is the winner for its superior composition, creativity, and cleaner execution of the caricature concept.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent fur texture and lighting on the animals.
  • + High resolution with beautiful god rays and bokeh.
  • + Character consistency in terms of 'expressive eyes' is very strong.
  • Failed to include the baby bunny entirely.
  • Anatomy of the kitten is slightly awkward with the overlapping limbs.

Vidu Q2

  • + Successfully included all requested animals: retriever, kitten, fox, and bunny.
  • + Dynamic composition that suggests movement and chasing better than Model A.
  • + Great adherence to the 'lush wildflower meadow' and 'dew sparkles' descriptors.
  • Anatomy issues, notably the fox kit having a cat-like face and some paws blending into the grass.
  • Texture is slightly more 'ai-painterly' and less photorealistic than Image A.

Verdict: While FLUX.2 [klein] 9B produced a more polished and photorealistic image with superior lighting, it failed a significant part of the prompt by omitting the rabbit. Vidu Q2 succeeded in including every requested element and animal, capturing a more active 'chasing' scene, making it the better choice for prompt adherence despite slightly lower technical photorealism.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent application of Ghibli-style watercolor textures and pastel palette
  • + Strong interpretation of a 'dreamy background' with rolling hills and soft light
  • + Maintains the composition and character poses while fully transforming the medium
  • The facial expressions are softened too much, losing the specific tension of the original meme
  • Changed the background from a city to a countryside, which deviates from the source image's setting

Vidu Q2

  • + Perfectly preserves the source image's background and character and facial expressions
  • + Accurately captures the 'distracted' and 'angry' emotions from the original meme in an anime style
  • + Good line work that stays true to Ghibli-style character design
  • The textures feel more like a flat digital cell-shading than the requested hand-painted texture
  • Background remains a bit too literal to the photo rather than being 'dreamy'

Verdict: Both models successfully translated the image into an illustration, but they chose different paths. FLUX.2 [klein] 9B followed the stylistic prompts for 'dreamy backgrounds' and 'soft watercolor' more closely, though it lost the narrative tension of the original meme, whereas Vidu Q2 perfectly preserved the core emotional content and urban setting of the source image while applying a cleaner anime aesthetic. Vidu Q2 is the winner for better preserving the identity and context of the source image while still achieving a quality Ghibli-esque look.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.2 [klein] 9B
Before After
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Excellent preservation of the subject's facial features and the dog's appearance
  • + Highly creative and dynamic 'wind-blown' hair effect that radiates outward
  • + Large number of falling leaves adds significantly to the energetic feel
  • The hair has some visible digital artifacts and unnatural strand separation
  • A few leaves appear to be 'floating' statically rather than moving through space

Vidu Q2

  • + Natural-looking hair movement with realistic physics for a lateral breeze
  • + Clean integration of the falling leaves with motion blur for a sense of speed
  • + Strong preservation of the source image's lighting and color palette
  • The hair effect is slightly less 'energetic' compared to the other model

Verdict: Both models successfully interpreted the prompt, but Vidu Q2 is the winner due to the more realistic and natural rendering of the wind-blown hair and motion-blurred leaves. While FLUX.2 [klein] 9B provided a more dramatic hair effect, it introduced some stringy digital artifacts that look less photographic than Vidu Q2's output.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Perfect text rendering of the restaurant name and establishment date.
  • + Clean, professional vector lines and effective use of negative space.
  • + Highly accurate adherence to the vintage minimalist color palette and style.

Vidu Q2

  • + Captures the warm brown and cream tones well.
  • + Good use of subtle texture on the background.
  • Serious spelling errors in both the brand name and 'Est.'
  • The steam effect is rendered inside the cloche dome incorrectly.
  • Chaotic composition with repetitive and misspelled text.

Verdict: FLUX.2 [klein] 9B is the clear winner as it perfectly renders all requested text and adheres to a professional graphic design standard. Vidu Q2 fails on basic literacy and logic, placing the steam inside the dome and misspelling nearly every word in the prompt.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.2 [klein] 9B
Vidu Q2

AI Judge Analysis

FLUX.2 [klein] 9B

  • + Stronger adherence to the requested NASA-inspired color palette.
  • + Superior text rendering for most labels compared to the competitor.
  • + Clever inclusion of three astronaut silhouettes as supporting details.
  • The layout flow is somewhat confusing, with steps jumping around the canvas.
  • Contains significant spelling errors like 'AOILLO' and 'EARDHT'.

Vidu Q2

  • + Layout is extremely clean and follows a logical linear progression.
  • + The flat vector style is more consistent across all icons.
  • + Iconography for the lunar module and planets is more detailed and balanced.
  • Fails significantly on text rendering, with 'ALFONCH' instead of 'Apollo'.
  • Does not follow the 6-step sequence as requested, repeating steps and icons.
  • Colors are somewhat washed out compared to the 'navy and red' requested.

Verdict: Both models struggled with text and specific step-by-step logic, but FLUX.2 [klein] 9B followed the requested color palette and mission details (like naming Armstrong) much more effectively. Vidu Q2 produced a more organized visual grid, but the text was mostly gibberish and the content of the steps did not match the prompt's instructions as well as FLUX.2 [klein] 9B.

Next steps

Explore each model