Black Forest Labs' open-weights multimodal flow transformer for in-context image generation and editing, available for non-commercial use with character consistency and style transfer capabilities
Settled by community votes across 7 shared challenges, with an AI judge weighing in on each.
FLUX.1 Kontext [dev]
#54 of 62 in Text-to-Image
ImagineArt 1.5 (Preview)
#6 of 62 in Text-to-Image
Where the votes landed
FLUX.1 Kontext [dev]
0%
win rate
Ties
0%
ImagineArt 1.5 (Preview)
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Excellent adherence to lighting instructions with a clear source from the left.
- + Precise and clean geometry of the glass cube and blue sphere.
- + Modern, high-resolution aesthetic with realistic textures.
- − The glass cube appears to have a mirrored base which wasn't specifically requested, though it adds to the visual appeal.
ImagineArt 1.5 (Preview)
- + Realistic vintage texture on the red book.
- + Good placement of the plant behind the cube.
- + Accurate physical interaction between the objects.
- − The glass cube is excessively thick and distorts the internal sphere significantly.
- − The lighting is somewhat flat and lacks the 'soft window light' definition seen in the other model.
- − There is a strange artifact/repetition of the sphere at the top inner edge of the cube.
Verdict: FLUX.1 Kontext [dev] followed all spatial and lighting instructions perfectly, producing a clean and aesthetically pleasing image. ImagineArt 1.5 (Preview) struggled with the transparency of the glass, creating a blocky, distorted look that made the sphere appear trapped in a solid block rather than a hollow cube.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Excellent handling of street reflections and atmospheric lighting
- + Detailed red bicycle model and rain streaks
- + Captures an 'imperfect framing' look by placing the subject in the middle of a road
- − The subject is posing with or sitting on the bike rather than repairing it
- − Cars in the background are static rather than having the requested motion blur
- − The man's suit is an unlikely choice for roadside bicycle repair
ImagineArt 1.5 (Preview)
- + Accurately depicts the action of repairing a bicycle
- + Outstanding natural skin texture and realistic age details on hands and face
- + Effective use of 'imperfect framing' for a candid street photography feel
- − Lacks visible motion blur from passing cars
- − The rain effect is very subtle compared to the wet pavement
- − Minor anatomical artifacts on the fingers of the right hand
Verdict: While FLUX.1 Kontext [dev] creates a more cinematic atmosphere with beautiful lighting and rain, it fails the primary prompt instruction of showing the man 'repairing' the bike. ImagineArt 1.5 (Preview) ignores most of the cinematic lighting but follows the core prompt much better, capturing a realistic, candid moment of repair with superior skin textures.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Exceptional sharpness and skin texture quality
- + Highly detailed and consistent engraving on the plate armor
- + Clean, artistic lighting with a dramatic cinematic feel
- − Missed the request for braided hair with beads
- − The scars appear more like fresh paint or light surface scratches than battle-worn marks
ImagineArt 1.5 (Preview)
- + Strict adherence to all prompt elements, including braided hair and beads
- + Excellent texture detailing on the leather straps, chainmail, and cloth layers
- + Strong interpretation of 'battle-worn' with realistic dirt and skin imperfections
- − Lighting is a bit harsh on the face, causing some blown-out highlights
- − Overall image has a slightly more 'digital' feel compared to the photographic quality of Model A
Verdict: Model B (ImagineArt 1.5) is the winner because it successfully followed every specific detail of the prompt, including the braided hair and the complex layering of leather and cloth. While Model A (FLUX.1 Kontext) produced a more polished and aesthetically pleasing cinematic portrait, it ignored the specific request for braids and beads.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Excellent image clarity and high-resolution food photography.
- + Strong adherence to the minimalist, bold sans-serif font request.
- + Balanced and creative grid layout that fits the modern aesthetic.
- − Text is mostly gibberish despite looking visually correct.
- − Some food items look slightly messy or unidentifiable upon close inspection.
ImagineArt 1.5 (Preview)
- + Successfully mimics the structure of an actual multi-section menu with prices.
- + Good use of white space and section breaks for a professional look.
- + Images of food, specifically the pizza, look very appetizing and realistic.
- − Text is extremely blurry and illegible.
- − Visual quality is lower with softer focus compared to the other model.
- − The 3D mock-up perspective makes the actual design harder to evaluate than a flat layout.
Verdict: FLUX.1 Kontext [dev] provides a much sharper, high-contrast design that captures the 'bold' and 'minimalist' requirements perfectly, even if the text doesn't make sense. ImagineArt 1.5 (Preview) creates a more realistic menu structure with prices and sections, but it suffers from low resolution and a presentation style that obscures the design details. FLUX.1 Kontext [dev] is the preferred choice for its vibrant, professional graphic quality.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Excellent adherence to the 'cartoon' and 'soft texture' style requirements.
- + The text layout is bold, clean, and perfectly centered as requested.
- + Very high clarity and consistent 3D rendering style.
- − The flag icon is abstract and does not represent the Japanese flag.
- − The sushi design is overly simplified, bordering on toy-like rather than a 'miniature 3D scene'.
ImagineArt 1.5 (Preview)
- + Features a much higher level of detail in the food items, including realistic PBR textures.
- + Includes a correct Japanese flag icon.
- + Better use of the 'diorama base' and 'isometric' perspective requested.
- − Text is top-right instead of top-center.
- − The scene contains a large amount of sushi, ignoring the 'minimal garnish and plate' instruction.
- − The text font is stylized with lines, making it less readable than Model A's bold text.
Verdict: FLUX.1 Kontext [dev] followed the layout and aesthetic instructions more strictly, particularly regarding text placement and the clean cartoon style, but failed on the flag icon. ImagineArt 1.5 (Preview) provided much better visual detail and a correct flag, but missed the central text alignment and the minimalism requested in the prompt. FLUX.1 Kontext [dev] is the likely winner for its superior cleanliness and adherence to the specific 'top-center' layout and 'solid background' framing.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Excellent text legibility and accuracy
- + True minimalist vector aesthetic
- + Clean and professional alignment
- − Missed the requested banner element
- − Missing the requested subtle texture on the background
ImagineArt 1.5 (Preview)
- + Stronger vintage/retro feel with gradients and texture
- + Better composition with the emblem style
- + Includes the steam and circular emblem layout
- − Text contains a small artifact/error on the letter 'i' in Florian
- − Cloche illustration is slightly less 'minimalist' than requested
Verdict: FLUX.1 Kontext [dev] delivers a much cleaner, modern minimalist logo with perfect typography, though it ignores some stylistic prompts like the banner and texture. ImagineArt 1.5 (Preview) captures the 'vintage' and 'emblem' atmosphere more effectively with a better use of tones and shapes, despite a slight rendering artifacts in the script font. ImagineArt is the winner for better adherence to the specific vintage aesthetic and requested layout elements.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.1 Kontext [dev]
- + Features a consistent, minimalist icon set
- + Accurately follows the requested navy blue background dominance
- − Contains significant spelling errors in the main title ('Apolo 11')
- − Iconography is abstract and difficult to interpret as specific mission stages
- − Text blocks are largely illegible gibberish
ImagineArt 1.5 (Preview)
- + Excellent adherence to the chronological flow and specific step descriptions
- + Successfully captures the NASA-inspired color palette and modern vector aesthetic
- + High-quality text rendering and clear, recognizable icons for the Saturn V and Lunar Module
- − Includes some minor spelling errors in supporting text like 'Landar' and 'Mafker'
- − Layout is slightly cluttered with overlapping elements in the center
Verdict: ImagineArt 1.5 (Preview) is the clear winner as it successfully interprets the educational intent of the prompt, providing a clear 6-step flow with recognizable icons for the rocket and landing gear. FLUX.1 Kontext [dev] fails significantly in basic text rendering (misspelling 'Apollo') and uses icons that are too abstract to communicate the mission steps effectively.
Explore each model
Vyro AI's professional-grade text-to-image model delivering photorealistic output with accurate text rendering and typography precision for commercial workflows