Head to head
Esc

Models · slot A

to navigate to pick

FLUX.1 Kontext [max] Black Forest Labs OmniGen v2 VectorSpaceLab

Settled by community votes across 15 shared challenges, with an AI judge weighing in on each.

FLUX.1 Kontext [max]

23.7 arena score

#23 of 62 in Text-to-Image

Skill signature · Text-to-Image

OmniGen v2

16.9 arena score

#57 of 62 in Text-to-Image

Vote tally

Where the votes landed

FLUX.1 Kontext [max]

0%

win rate

Ties

0%

OmniGen v2

0%

win rate

Shared challenges 15

Challenge by challenge

The strongest take from each model on every shared challenge, with the AI judge's read.

Geometric Composition

Text-to-Image

“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent handling of light and shadows
  • + Highly realistic glass refraction and reflections
  • + Solid composition that creates a professional photograph feel
  • The plant is more alongside the cube than directly behind it
  • The blue sphere has a slightly rough texture rather than being a smooth sphere

OmniGen v2

  • + Perfect adherence to object placement
  • + Very clean, vibrant colors
  • + The plant is clearly visible through the glass as requested
  • The lighting is a bit flat and clinical
  • The blue sphere appears to be floating unnaturally without shadows on the cube's base

Verdict: FLUX.1 Kontext [max] produces a much more realistic and aesthetically pleasing image with sophisticated lighting and material physics, though it slightly misses the 'behind' placement of the plant. OmniGen v2 follows the spatial instructions perfectly but lacks the photographic quality and realistic shadow grounding seen in the first image.

Candid Street Photography

Text-to-Image

“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent adherence to the 'imperfect framing' and 'candid' aspects of the prompt.
  • + Superior textures on the skin, clothing, and wet pavement.
  • + Realistic bicycle mechanics and complex rain interaction with the scene.
  • The man's ethnicity appears somewhat ambiguous compared to the specific request.

OmniGen v2

  • + Accurately depicts an elderly Japanese man as requested.
  • + Clean composition with clear reflections.
  • + Follows the red bicycle requirement well.
  • The subject is just standing near the bike rather than 'repairing' it.
  • The image lacks the requested 'motion blur' and 'imperfect framing' for a candid look.
  • Lighting and textures feel slightly more artificial and less cinematic than the competitor.

Verdict: FLUX.1 Kontext [max] captured the specific mood, texture, and technical requirements of the prompt far better, successfully delivering the motion blur and candid framing requested. While OmniGen v2 adhered better to the specific ethnicity of the man, it failed to depict the action of repairing and felt significantly more like a staged studio shot than a candid street photo.

Fantasy Warrior

Text-to-Image

“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Incredible skin texture and realism, showing fine pores, sweat, and subtle dirt marks.
  • + Excellent adherence to the 'battle-worn' aesthetic with realistic scars and weathered armor.
  • + Superb rendering of the engraved plate armor with complex torchlight reflections.
  • The braids are a bit messy and lack the 'small beads' requested in the prompt.
  • High skin saturation makes the character look somewhat sunburnt.

OmniGen v2

  • + Successfully included the small beads within the braided hair as requested.
  • + Strong bokeh effect and clear torchlight lighting in the background.
  • + Clean composition with a focused, sharp gaze.
  • Does not look 'battle-worn'; the skin is too smooth and the dirt looks like makeup.
  • Armor texture is somewhat flat and lacks the metallic realism found in the first image.
  • The character's face looks too pristine and youthful for the requested veteran paladin archetype.

Verdict: FLUX.1 Kontext [max] delivered a much more convincing interpretation of a battle-worn paladin, with exceptional detail in the armor engravings and lifelike skin textures. While OmniGen v2 followed the specific request for hair beads more accurately, it failed to capture the gritty, realistic atmosphere and seasoned appearance required by the prompt, resulting in a character that looks more like a cosplayer than a veteran soldier.

Modern Clean Menu

Text-to-Image

“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Clean, professional aesthetic typical of a real menu.
  • + Clear food photography that looks appetizing and consistent.
  • + Good use of white space and hierarchy.
  • Failed to include prominent sections for appetizers and mains as requested.
  • The body text and section headers are mostly gibberish.
  • The layout is a bit repetitive with mostly pizza images.

OmniGen v2

  • + Successfully included specific sections for appetizers, pizza, and mains (though misspelled).
  • + Strong use of vibrant color accents as requested.
  • + Better variety in the food photos shown.
  • Multiple severe spelling errors in titles (e.g., 'RESTAURATED MENTS', 'PIZZZZAN').
  • The text blocks are illegible smudges rather than font characters.
  • Uneven grid alignment on the right-hand page.

Verdict: FLUX.1 Kontext [max] produces a much more believable and professional-looking menu with high-quality photography, whereas OmniGen v2 struggles with alignment and has comical spelling errors. While OmniGen v2 adhered closer to the prompt's request for specific categories (appetizers/mains), the overall visual quality of FLUX.1 Kontext [max] makes it the superior design choice.

Magic Burger Explosion: Fiery Photorealism Challenge

Text-to-Image

“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with clean, readable text that matches the glowing embers theme.
  • + Superior photorealistic textures on the lettuce, meat, and bun.
  • + Dynamic composition with many suspended ingredients creating a true 'exploded' effect.
  • The price is missing the starburst element requested.
  • The burger in the center remains largely intact rather than being fully vertically exploded.

OmniGen v2

  • + Successfully included the price in a starburst graphic.
  • + Strong graphic design style that looks like a polished digital advertisement.
  • Fails the 'exploded' requirement as the burger is fully assembled.
  • Lacks the requested 'photorealistic' detail, appearing more like a 3D render or vector illustration.
  • Text is cut off on the left side of the image.

Verdict: FLUX.1 Kontext [max] far exceeds OmniGen v2 in terms of photorealism and dynamic motion, capturing the 'exploded' burger concept much more effectively. While OmniGen v2 included the starburst element, it failed to separate the burger components and suffered from a cut-off text layout.

Chalkboard Menu

Text-to-Image

“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent text rendering with no spelling errors.
  • + Authentic chalk texture and smear marks on the board.
  • + Successfully completed the truncated 'Brown But...' prompt as 'Brown Butter Chocolate Chip Cookies'.
  • The title is not in 'elegant cursive' as requested, but rather a blocky print.

OmniGen v2

  • + Accurately captured the 'Specials' and date header with clean lines.
  • + Good contrast between the board and the frame.
  • Significant spelling errors throughout the menu items.
  • The handwriting looks digital and lacks natural chalk texture.
  • Mangled layout with prices scattered randomly.

Verdict: FLUX.1 Kontext [max] significantly outperformed OmniGen v2 by providing legible, accurate text and capturing the realistic smear and texture effects of a physical chalkboard. While FLUX.1 Kontext [max] missed the 'cursive' instruction for the title, it successfully inferred the full name of the final menu item and maintained a professional layout, whereas OmniGen v2 suffered from severe spelling and formatting issues.

The Reversed Rodeo

Text-to-Image

“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent cinematic lighting and textured detail on the horse and suit.
  • + High resolution with realistic star fields and planet rendering.
  • + More natural integration of the subject into the environment.
  • Failed the specific spatial instruction for the horse to be on top of the astronaut.

OmniGen v2

  • + Includes multiple moons to add to the space theme.
  • + Clean, bright colors for a graphic look.
  • Failed the specific spatial instruction for the horse to be on top of the astronaut.
  • Lower visual quality with a slightly cartoonish aesthetic.
  • Anatomical issues where the horse's back legs disappear into the background.

Verdict: Both FLUX.1 Kontext [max] and OmniGen v2 completely failed to follow the specific 'horse on top' instruction, which was intended to subvert the common 'astronaut on a horse' trope. FLUX.1 Kontext [max] is the preferred image because its visual execution is significantly more cinematic and detailed than the flatter, more artificial look of OmniGen v2.

The Capybara Taxi Driver

Text-to-Image

“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent photorealism with convincing textures on the capybara's fur and the leather jacket.
  • + High prompt adherence for the capybara's anatomy, correctly showing its paws on the steering wheel.
  • + Sophisticated lighting and bokeh that accurately mimic a New York night scene.
  • The passenger appears to be making a phone call rather than just looking at her phone as requested.

OmniGen v2

  • + The passenger's expression and action perfectly match the 'bored looking at phone' prompt.
  • + The composition clearly shows all elements of the scene, including the yellow taxi exterior.
  • Major anatomical failure, showing human hands coming out of the capybara's sleeves to steer.
  • The capybara's head is poorly integrated into the character's body, looking like a mask or a sticker.
  • Lower overall visual quality with more plastic-looking textures compared to the competitor.

Verdict: FLUX.1 Kontext [max] produced a high-quality, believable image that captures the atmosphere of a night taxi ride, even though it interpreted the passenger's action slightly differently. OmniGen v2 failed significantly on the central prompt requirement by giving the capybara human hands, which breaks the logic and immersion of the scene. Overall, FLUX.1 Kontext [max] is the superior choice for its realistic textures and correct animal anatomy.

The Halloween Invitation

Text-to-Image

“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with clear, legible text that matches the requested details exactly.
  • + Stunning atmospheric composition with a moody, cinematic gothic aesthetic.
  • + Intricate border details featuring webs and thorns as requested.
  • Repeats the location 'The Arches, NYC' twice at the bottom.
  • The commas in the date '30,10,2026' technically should be dots per the prompt.

OmniGen v2

  • + Good use of the parchment paper asset for a vintage poster feel.
  • + Follows the layout instructions including the scroll banner and central jack-o-lantern.
  • Significant text rendering issues and typos like 'FRIGTS' and 'ARCAS'.
  • Lower visual fidelity with a flat, more cartoonish style compared to the cinematic request.
  • Messy overlap of text elements at the bottom making it hard to read.

Verdict: FLUX.1 Kontext [max] is the clear winner, producing a highly polished and professional-looking invitation with impressive gothic detail and nearly perfect text. OmniGen v2 struggles with text legibility and generates a flatter, less sophisticated illustration that misses the 'cinematic' and 'polished' aesthetic requested.

Isometric Miniature Diorama Scenes

Text-to-Image

“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with a clean, rounded cartoon aesthetic.
  • + High-quality soft lighting and refined textures on the salmon.
  • + Great composition with inclusion of soy sauce and wasabi.
  • Missed the small flag icon requested in the prompt.
  • Rice texture is slightly lumpy and more stylized than realistic.

OmniGen v2

  • + Includes a flag icon as requested, although it is not a Japanese flag.
  • + Stronger 3D isometric feel for the diorama base.
  • + Extremely crisp text rendering with a 3D shadow effect.
  • The sushi design is nonsensical, featuring a fish tail on a roll and strange internal squares.
  • The flag is color-incorrect for Japan (red/yellow).

Verdict: Both models followed the layout and isometric instructions well. FLUX.1 Kontext [max] produced much more appetizing and logical sushi, whereas OmniGen v2 failed on the anatomical details of the food and the specific colors of the Japanese flag, despite following more of the small object instructions. FLUX.1's superior texture and coherence make it the stronger image.

Adorable Baby Animals in Sunny Meadow

Text-to-Image

“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent adherence to the 'hyper-photorealistic' part of the prompt with realistic fur textures and anatomy.
  • + Features all four requested animals: golden retriever, kitten, bunny, and fox kit.
  • + Natural-looking lighting, depth of field, and god rays that enhance the 'masterpiece' quality.
  • The bunny's anatomy is slightly strange around the paws.
  • The number of butterflies is high, which slightly distracts from the animal interactions.

OmniGen v2

  • + Bright, vibrant colors that fit a 'joyful' vibe.
  • + Clear composition and expressive eyes as requested.
  • Completely failed the 'hyper-photorealistic' requirement, delivering a 3D cartoon/CGI style instead.
  • Failed to include all requested animals, missing the baby bunny.
  • Butterflies and foliage look like clip-art and lack integration with the scene.

Verdict: FLUX.1 Kontext [max] far exceeded the quality of the other model by providing a truly photorealistic image that included all four requested animals with beautiful lighting and textures. OmniGen v2 failed the primary stylistic prompt for photorealism, missing an entire animal species from the request and producing a result that looks like a low-budget animated film.

Studio Ghibli Anime Style

Editing
Edit instruction

“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”

Source
FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent preservation of the specific plaid pattern and clothing details from the original image.
  • + Captures a high-quality hand-painted watercolor texture reminiscent of Studio Ghibli backgrounds.
  • + Maintains the correct facial expressions, specifically the girlfriend's annoyed look.
  • The woman in the foreground has a slightly blurry face compared to the characters in the middle ground.

OmniGen v2

  • + Successfully translates the scene into a clean anime aesthetic.
  • + Good preservation of the overall composition and character poses.
  • Fails to capture the 'gentle' or 'warm' nostalgic mood, appearing more like modern digital vector art.
  • Loses the specific character expressions, particularly making the girlfriend look happy or neutral instead of annoyed.
  • The lighting is flat and lacks the 'dreamy' quality requested in the prompt.

Verdict: FLUX.1 Kontext [max] is the clear winner as it perfectly captures the requested Studio Ghibli aesthetic with its soft, hand-painted watercolor textures and warm lighting. It also does a much better job of preserving the source image's character details (like the specific plaid shirt) and the original emotional context of the scene compared to OmniGen v2, which produced a more generic, flat anime style with incorrect facial expressions.

Golden Hour Stroll

Image Editing
Edit instruction

“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”

Before After
FLUX.1 Kontext [max]
Before After
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent preservation of the original subjects' facial features and textures.
  • + Realistic hair physics that look like actual wind rather than a static shape.
  • + Perfect source preservation of the background elements like the bridge and foliage.
  • The flying leaves are somewhat small and sparse.
  • The left hand holding the leash has become slightly distorted compared to the original.

OmniGen v2

  • + Successfully added colorful falling leaves that enhance the energetic feel.
  • + Clearly shows hair blowing in the wind as requested.
  • Changed the facial features of both the woman and the dog, losing the likeness of the source image.
  • The overall image quality looks more 'plastic' and AI-generated compared to the source.
  • The hair looks like a frozen, stylized shape rather than a dynamic motion effect.

Verdict: FLUX.1 Kontext [max] is the winner because it successfully applied the movement edits while maintaining the identity and photographic quality of the source image. OmniGen v2 failed to preserve the source subjects, significantly altering the faces of both the woman and the dog, and producing a less realistic result.

Vintage Cafe Logo

Text-to-Image

“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent typography with perfect spelling and correct accent marks.
  • + Successful implementation of the subtle vintage texture and woodblock aesthetic.
  • + Strong adherence to all prompt elements including the cloche and banner layout.
  • The steam icon is a bit simple compared to the detailed cloche.

OmniGen v2

  • + Clean vector-style lines suitable for a minimalist logo.
  • + Good use of the warm cream and brown color palette.
  • Failed spelling significantly, rendering 'CAFFFLORIN' instead of 'Caffè Florian'.
  • Lacks the requested 'subtle texture' on the background.
  • The composition feels a bit cramped with the text overlapping parts of the design.

Verdict: FLUX.1 Kontext [max] produced a superior, professional-grade logo that perfectly followed the typography and stylistic instructions. In contrast, OmniGen v2 failed on text accuracy and did not capture the vintage textured feel requested in the prompt.

Apollo 11: Journey to Tranquility

Text-to-Image

“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”

FLUX.1 Kontext [max]
OmniGen v2

AI Judge Analysis

FLUX.1 Kontext [max]

  • + Excellent vector illustration style that feels professional and cohesive.
  • + Follows the NASA-inspired color palette perfectly.
  • + Includes complex visual elements like the lunar module and landing site detail.
  • Incorrectly labels the Earth as 'Moon'.
  • Includes a random planet with rings (Saturn/Jupiter) which wasn't part of the journey.
  • Text labels are cluttered and contain several misspellings like 'LUNEAR MODULLE'.

OmniGen v2

  • + Matches the clean 'flat-vector' and 'infographic' layout request more accurately.
  • + Excellent use of negative space and crisp iconography.
  • + Perfectly adheres to the color palette constraints.
  • Misidentifies the mission as 'APOLO 17' instead of Apollo 11.
  • Text consists almost entirely of gibberish or severe misspellings.
  • Doesn't actually visualize the 6 requested steps, showing only a few icons.

Verdict: FLUX.1 Kontext [max] creates a beautiful, detailed illustration that follows the narrative of the prompt, but it fails significantly on factual accuracy by mislabeling Earth as the Moon and adding unnecessary planets. OmniGen v2 captures the modern infographic aesthetic and layout much better, but it fails on content by providing the wrong mission number and completely garbled text. FLUX.1 Kontext [max] is the preferred choice for its higher complex detail and adherence to the various mission stages.

Next steps

Explore each model