Fast distilled version of Black Forest Labs' FLUX.2 [dev] optimized for speed and cost efficiency.
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
FLUX.2 [dev] Flash
#7 of 62 in Text-to-Image
Vidu Q2
#42 of 62 in Text-to-Image
Where the votes landed
FLUX.2 [dev] Flash
0%
win rate
Ties
0%
Vidu Q2
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent adherence to the 'soft window light' lighting instruction.
- + Highly realistic glass textures and reflections.
- + Strong photographic composition with natural depth of field.
- − The plant is slightly more to the side than 'behind' compared to the other model.
Vidu Q2
- + Perfect object placement according to the prompt.
- + Crisp details on the book and plant leaves.
- + Strong spatial arrangement of all requested elements.
- − The lighting is harsh and direct sunlight rather than the requested 'soft window light'.
- − The shadows on the table are very sharp and clash with the soft light prompt.
Verdict: Both models followed the complex spatial instructions perfectly. FLUX.2 [dev] Flash is the winner because it successfully captured the 'soft window light' mood of the prompt, whereas Vidu Q2 generated harsh, high-contrast shadows that contradicted the requested lighting style.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent preservation of the man's facial features and specific hairstyle
- + Maintains the exact clothing (plaid coat and black scarf) from the source image
- + High-quality, realistic interior detailing of the car consistent with a luxury convertible
- − The driver is sitting on the right side, which is incorrect for a car driving in California
Vidu Q2
- + Better captures the overall scale of the car and the coastline environment
- + Positions the driver on the correct side for US driving (left-hand drive)
- − Significant loss of detail and accuracy regarding the man's outfit
- − The man's facial features and hair are simplified and less recognizable compared to the source
- − Lighting on the car looks slightly synthetic/HDR-heavy
Verdict: FLUX.2 [dev] Flash is the clear winner for its superior ability to preserve the subject's identity, hair, and clothing down to the specific plaid pattern of the coat. While Vidu Q2 does a better job with the driving orientation and scale, it loses the specific characteristics of the person being 'edited' into the scene, which is the primary goal of the prompt.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent full-body composition that shows the entire scene as requested.
- + Effective use of shallow depth of field and motion blur to create depth.
- + Realistic skin textures and clothing details on the subject.
- − The structural logic of the bicycle frame is slightly warped near the seat and pedals.
- − Several small artifacts on the ground resembling scattered tools or debris look messy.
Vidu Q2
- + Strong execution of the 'imperfect framing' prompt with a close, candid crop.
- + Highly detailed texture on the hands and forearms.
- + Good atmosphere with wet pavement reflections and rain.
- − Anatomical issues with the hands, including an extra finger or merged digits.
- − The bicycle's geometry is nonsensical, with the seat and handlebars arranged incorrectly.
- − The motion blur on the background car is less convincing than Model A.
Verdict: FLUX.2 [dev] Flash is the clear winner as it provides a coherent, storytelling image that adheres to all parts of the prompt, including the motion blur and framing. While Vidu Q2 captures the 'candid' feel well, it suffers from significant anatomical errors in the hands and structural failures in the bicycle rendering.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent texture on leather straps and cloth underlayer as requested
- + Highly detailed engraving on the plate armor
- + Successfully incorporates the beaded braids throughout the hair
- − The facial wounds appear more like fresh bloody marks than 'faint scars'
- − Symmetry in the facial lighting is a bit flat compared to the dramatic torchlight suggestion
Vidu Q2
- + Dynamic lighting with strong warm reflections on the metal highlights
- + Good facial expression and clear, lifelike eyes
- + Accurately represents the depth of field with visible bokeh
- − Lacks the small beads on the braids as requested in the hair
- − The engraving on the armor is less intricate than the competing model
- − Image composition is slightly less 'close portrait' than the requested framing
Verdict: FLUX.2 [dev] Flash captured nearly every specific detail of the prompt, particularly the beads in the hair and the high-resolution texture of the leather and cloth. While Vidu Q2 had more dramatic lighting and better skin tone rendering, it missed the specific detail of the hair beads and offered less detailed engraving, making FLUX.2 [dev] Flash the winner for prompt adherence.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent grid layout that looks like a real menu.
- + Professional use of bold, sans-serif fonts for section headers.
- + High-quality, appetizing food photography that is consistent in style.
- − Several spelling errors in headings like 'CASSLAL', 'R!ZLA', and 'PRESTU RVANT'.
- − The food images are almost exclusively pizza, even under the 'Appetizers' and 'Mains' labels.
Vidu Q2
- + Better variety of food photography including salads and drinks.
- + Includes more realistic 'filler' text for dish descriptions.
- − The layout is cluttered and confusing with overlapping text and images.
- − Text rendering is highly distorted and contains many nonsensical characters.
- − Lacks the clean, 'bold' modern aesthetic requested in the prompt.
Verdict: FLUX.2 [dev] Flash produced a much more professional and aesthetically pleasing design that looks like a real modern menu, despite some spelling errors. Vidu Q2 followed the prompt's request for food variety better but failed significantly on the 'minimalist' and 'clean' layout requirements, resulting in a cluttered and illegible output. FLUX.2 is the clear winner for its superior composition and font rendering.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent typography with a consistent fiery texture.
- + Photorealistic texture on the meat patty and vegetables.
- + Clean composition with a sophisticated color palette.
- − The 'exploded' effect is slightly more subtle than expected.
- − The price starburst is slightly less prominent than some advertising styles might prefer.
Vidu Q2
- + High-energy 'exploded' action with more dynamic sauce splashes.
- + Vivid, intense fire background that matches the prompt's theme.
- + Strong, high-contrast colors that grab attention.
- − The pricing text has a malformed currency symbol.
- − The lighting on the burger components is slightly oversaturated, making it look less photorealistic.
- − The text rendering on 'MAGIC BURGER' is a bit traditional and flat compared to the fiery effect requested.
Verdict: FLUX.2 [dev] Flash produces a more professional and realistic advertisement with perfect text rendering and sophisticated lighting. While Vidu Q2 captures the 'exploded' motion more aggressively, it fails on technical details like the currency symbol and has a less realistic finish.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Perfect spelling of all menu items including complex terms.
- + Highly realistic chalk texture with natural smudging and authentic handwriting style.
- + Correct completion of the truncated prompt for the final menu item.
- − Repeats the price '$9' on two lines for the same item.
- − Title handwriting is less 'elegant cursive' and more block-style than requested.
Vidu Q2
- + Good use of 'elegant cursive' for the top title line.
- + Vibrant chalk-like contrast and variation in letter size.
- − Severe spelling errors throughout the menu items such as 'Truffe Musshoom'.
- − Text becomes unintelligible gibberish toward the bottom of the board.
- − Incorrect pricing for the first item ($34 instead of $24).
Verdict: FLUX.2 [dev] Flash is the clear winner as it produced perfectly legible and correctly spelled text for the entire menu. While Vidu Q2 captured the 'elegant cursive' style for the header, it failed significantly on text coherence, resulting in numerous spelling errors and garbled text at the bottom.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Successfully replicates the clothing and face of the character reference.
- + Maintains the red ottoman and yellow backdrop perfectly.
- − Fails completely on the pose instruction, leaving parts of the original woman's head behind the man's torso.
- − Anatomical failure with a floating leg and an awkward upright stance that ignores the dynamic lean.
Vidu Q2
- + Accurately replicates the complex leg-crossing pose and leaning angle from Image 1.
- + Successfully translates the character's clothing, face, and accessories onto the new pose.
- + Good integration of the scarf and accessories in the dynamic position.
- − Minor distortion in the hand on the left side of the frame.
- − The character is wearing shorts instead of the full trousers from Image 2.
Verdict: Vidu Q2 is the clear winner as it successfully combined the dynamic, difficult pose from Image 1 with the character from Image 2. In contrast, FLUX.2 [dev] Flash failed the primary task, resulting in a disturbing image where the new character is simply placed on top of the original woman, leaving her hair and torso visible behind him.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent anatomical details on the horse especially the fur texture and muscles
- + High-quality cinematic lighting with a realistic space background
- + Clean rendering of the astronaut suit.
- − Failed the specific spatial instruction for the horse to be on top of the astronaut.
Vidu Q2
- + Creative use of nebulae and space colors on the horse's body
- + Vibrant and surreal aesthetic that matches part of the prompt.
- − Failed the specific spatial instruction for the horse to be on top of the astronaut
- − Distorted rendering of the astronaut's hands and the horse's legs
- − Over-saturation leads to loss of fine detail.
Verdict: Both models failed the complex spatial constraint of placing the horse on top of the astronaut, instead opting for the standard 'astronaut riding horse' trope. FLUX.2 [dev] Flash is the superior image due to its exceptional technical execution, realistic textures, and cinematic composition, whereas Vidu Q2 suffers from anatomical distortions and lower-quality rendering.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Successfully captured the coat and scarf textures.
- − Catastrophic failure in facial preservation, resulting in a distorted, multi-face composite.
- − Added excessive jewelry not found in either source image.
- − Fails to maintain the background or the identity of the person.
Vidu Q2
- + Excellently preserved the person's identity and unique skin patterns.
- + Accurately transferred the coat, scarf, sunglasses, and watch from Image 2.
- + Maintained the beach background and pose of the original subject.
- − Integrated the skin texture onto the fabric of the coat sleeves.
- − The hair was slightly altered from Image 1.
Verdict: Vidu Q2 is the clear winner as it successfully performed the complex task of transferring an outfit while preserving the subject's identity and unique features. FLUX.2 [dev] Flash suffered a total breakdown in image coherence, creating a nightmarish facial distortion and ignoring the instruction to keep the person unchanged.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent adherence to the 'inside' perspective requested in the prompt
- + Superior rendering of capybara textures and facial features
- + More realistic interior lighting and bokeh effects
- − The capybara's hands look slightly more like bird talons than capybara paws
- − The taxi sign on top is visible through the roof, which is physically impossible
Vidu Q2
- + Provides a full body view showing the capybara sitting in the driver's seat
- + Good adherence to the clothing and accessory requirements
- + Clear distinctness between the front and back seat compartments
- − The perspective is 'outside-in' rather than 'inside' as requested
- − The capybara's head is disproportionately small compared to its body and the human's head
- − Significant lighting inconsistencies on the passenger's face
Verdict: FLUX.2 [dev] Flash followed the perspective prompt much better, creating an immersive 'inside' feel with high-quality photorealistic textures. Vidu Q2 provided a wider shot that technically included all elements, but the scale of the capybara and the external camera angle made it less effective than the intimate, more realistic composition of FLUX.2 [dev] Flash.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent text rendering with almost no spelling errors
- + Cohesive gothic composition with atmospheric lighting
- + High visual quality and professional-looking layout
- − Includes some gibberish text in the middle line of the event details
Vidu Q2
- + Strongly features the requested parchment texture
- + Good use of traditional gothic foliage and thorns
- − Numerous spelling errors in the title, banner, and date
- − The date included a 70th month (30.70.2025)
- − The composition feels a bit cluttered compared to Model A
Verdict: FLUX.2 [dev] Flash is the clear winner as it correctly followed the complex text instructions with high legibility and an elegant design. Vidu Q2 struggled significantly with text rendering, resulting in misspelled words and an impossible date, although its parchment texture was quite nice.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent preservation of the original facial features and environment
- + Good texture and volume in the added hair
- + Natural light interaction between the hair and the background
- − The hair type (tight curls/afro) significantly changes the person's established character from the source image
- − The hairline looks slightly artificial where it meets the forehead
Vidu Q2
- + Natural and realistic hair texture and style that matches the existing beard
- + Seamless blending of the hairline with the forehead
- + Perfect preservation of the original image's lighting and background
- − Minor artifacting near the top right of the hair where it meets the sky
Verdict: Both models do an exceptional job of preserving the source image's facial features and background. Vidu Q2 is the winner because it provides a much more natural-looking hair style that realistically matches the person's existing facial hair, whereas FLUX.2 [dev] Flash chooses a very specific hair texture that feels less cohesive with the original subject.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent PBR materials and soft textures that look professional.
- + Clean and accurate typography with a professional flag icon design.
- + Perfect composition and lighting for a miniature diorama style.
- − The sushi roll/nigiri hybrid is a slightly strange interpretation of sushi anatomy.
Vidu Q2
- + Successfully captured the 3D cartoon style mentioned in the prompt.
- + Included a wider variety of sushi types on the plate.
- + Good isometric perspective and clean colors.
- − Text rendering is slightly wobbly and less 'bold' than requested.
- − Materials feel more like plastic than realistic PBR textures.
- − The flag placement feels slightly less integrated than the icon in Model A.
Verdict: FLUX.2 [dev] Flash produced a significantly more polished result with superior texture quality and professional graphic design elements. While Vidu Q2 followed the cartoon aesthetic well, it lacked the high-fidelity rendering and crisp typography provided by FLUX.2 [dev] Flash.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Successfully incorporates all prompt elements (TV anchor, dogs, hockey) in a single cohesive scene.
- + Captures the user's likeness exceptionally well within a caricature style.
- + High level of detail with humorous touches like dogs playing hockey in the background.
- − The hands on the anchor are slightly malformed/small.
- − The 'TV SHOW' text is a bit generic.
Vidu Q2
- + Strong caricature art style with bold lines and vibrant colors.
- + Preserves the denim shirt from the original source image.
- + Includes creative details like paw prints on the news scripts.
- − The hand holding the microphone has six fingers.
- − The composition feels a bit cluttered with the microphone and background hockey rink lines.
- − The face likeness is slightly less accurate than Model A.
Verdict: FLUX.2 [dev] Flash is the clear winner as it perfectly balances the person's likeness with the caricature style while seamlessly blending the three distinct themes of TV anchoring, dogs, and hockey into a fun, professional-looking illustration. Vidu Q2 follows the instructions well but suffers from significant anatomical errors, specifically a wide hand with six fingers, and the likeness is not quite as sharp.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent fur texture rendering and hyper-photorealistic detail.
- + Beautifully executed golden hour lighting with clear god rays.
- + Includes all requested animals with expressive eyes and a cohesive composition.
- − The animals are largely static and posed rather than 'playfully chasing' and 'tumbling'.
- − Duplicate animals (two bunnies and two fox kits) were included when the prompt suggested a singular group.
Vidu Q2
- + Captures a much better sense of motion and 'playfully chasing' as requested.
- + Vibrant and dynamic composition with many butterflies spread throughout the scene.
- + Excellent usage of the entire frame for the meadow environment.
- − Anatomical issues on the puppy's paws and the fox's legs.
- − The 'tabby kitten' looks more like a calico/patterned kitten and lacks the 'ultra-detailed' fur of the competitor.
- − Significant artifacts and blurring around the moving limbs of the animals.
Verdict: FLUX.2 [dev] Flash produces a much higher quality, sharper image with superior lighting and texture, although it fails to capture the requested action and instead poses the animals. Vidu Q2 better interprets the 'chasing' and 'tumbling' aspect of the prompt but suffers from poor anatomical accuracy and significant visual artifacts. FLUX.2 is the preferred choice for its technical excellence and photographic realism.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent adherence to the 'Studio Ghibli' aesthetic with soft, hand-painted textures.
- + Completely transforms the background into a dreamy, nostalgic landscape while maintaining core composition.
- + Beautiful use of soft pastel colors and atmospheric lighting.
- − Faces are a bit more generic anime rather than specific Ghibli character designs.
- − Replaced the city setting entirely with a meadow, which may be more than the user wanted.
Vidu Q2
- + Preserved the original city background accurately while applying an illustrative filter.
- + Character expressions feel more emotive and closer to the original meme's intent.
- + Good preservation of the source image's specific details like the checkered shirt pattern.
- − Texture feels more like a digital cel-shaded look than 'hand-painted' textures.
- − The lighting is flat compared to the requested 'gentle' and 'warm' mood.
Verdict: FLUX.2 [dev] Flash followed the stylistic prompts more effectively, creating a truly nostalgic, hand-painted scene that feels like a frame from a Ghibli film, though it traded away the original urban setting. Vidu Q2 preserved the original image's context and city setting much better, but the visual style feels more like a standard modern anime filter rather than the specific soft-textured look requested. FLUX.2 is the winner for its superior artistic execution of the difficult 'hand-painted' and 'dreamy' instructions.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent preservation of the subject's face and original features.
- + Effective hair motion that feels symmetrical and energetic.
- + Leaves are well-integrated with the background depth.
- − The hair edit looks slightly digital and stencil-like at the edges.
- − Overall color saturation of the leaves is a bit duller than model B.
Vidu Q2
- + Natural and aesthetically pleasing hair movement.
- + Vibrant, high-contrast leaves create a strong sense of autumn dynamic.
- + Excellent source preservation of the background and the dog.
- − Some leaves appear to overlap the subject's body in a slightly flat way.
- − The hair volume increased significantly compared to the original.
Verdict: Both models succeeded in following the edit instructions while maintaining high consistency with the source image. Vidu Q2 is preferred for its more natural-looking hair movement and the vibrant, high-energy feel of the colorful leaves, whereas FLUX.2 [dev] Flash produced a slightly more rigid hair effect.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Perfect text rendering including the grave accent in 'Caffè'.
- + Clean vector-style emblem with professional balance.
- + Excellent texture work on the background and ribbon.
- − The steam is a bit simplistic compared to the detail of the cloche.
Vidu Q2
- + Good color palette adherence to 'warm brown and cream'.
- + Interesting steam illustration placed both inside and outside the dome.
- − Major spelling errors ('FARMIIN', 'Esttt', 'Fopli20').
- − Poor composition with redundant and cluttered text elements at the bottom.
- − The cloche handle and dome base show line inconsistencies.
Verdict: FLUX.2 [dev] Flash delivered a near-perfect logo that is ready for use, featuring precise typography, clean lines, and a cohesive design. Vidu Q2 failed significantly on typography and composition, producing nonsensical text and a cluttered layout that lacks the requested minimalist aesthetic.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.2 [dev] Flash
- + Excellent typography rendering with mostly accurate names and labels
- + Comprehensive adherence to all six requested steps in a logical layout
- + Features a detailed Saturn V rocket and high-quality vector illustrations of the lunar module
- − Layout is a bit cluttered with redundant labels for descent and landing
- − The space shuttle-style boosters on the Saturn V are historically inaccurate
Vidu Q2
- + Clean, minimalist vector aesthetic that matches the 'flat' style requested
- + Good use of the muted NASA color palette
- + Consistent iconography across all elements
- − Severe text rendering issues with nonsense words like 'ALFONCH' and 'Lunan Orutt'
- − Missing specific steps like the trajectory arc and Saturn V icon
- − Small supporting details like the astronaut silhouettes are generic and not mission-specific
Verdict: FLUX.2 [dev] Flash is the clear winner as it successfully followed the complex instructional sequence and rendered legible, accurate text for the mission steps and crew members. Vidu Q2 failed significantly on the text rendering and missed several of the specific infographic requirements, resulting in a generic and confusing poster.
Explore each model
ShengShu Technology's text-to-image and reference-to-image model with support for character consistency and multi-reference image processing