Black Forest Labs' aesthetically-tuned 12-billion parameter flow transformer optimized for high-quality images with incredible aesthetics, suitable for personal and commercial use
Settled by community votes across 13 shared challenges, with an AI judge weighing in on each.
FLUX.1 Krea [dev]
#47 of 62 in Text-to-Image
FLUX.1 [schnell]
#48 of 62 in Text-to-Image
Where the votes landed
FLUX.1 Krea [dev]
0%
win rate
Ties
0%
FLUX.1 [schnell]
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Perfect adherence to spatial instructions with the sphere inside the cube and book on top.
- + Superior photorealism with convincing glass reflections and natural lighting.
- + Excellent composition with a professional, moody atmosphere.
- − The plant in the background is quite dark and blurry, making it less distinct.
FLUX.1 [schnell]
- + Bright, clear colors and sharp detail on the plant leaves.
- + Very clean rendering with high resolution and no visible artifacts.
- − Logic error: includes a second blue sphere on top of the book and the internal sphere is floating.
- − The lighting feels a bit more synthetic compared to the realistic soft window light in the other model.
Verdict: FLUX.1 Krea [dev] followed the complex spatial prompt perfectly, placing the sphere inside the cube and the book atop it with high photographic realism. FLUX.1 [schnell] failed the logic of the prompt by adding an extra sphere on top of the book and rendering the internal sphere levitating without support. FLUX.1 Krea [dev] is the clear winner for its superior prompt adherence and natural lighting.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent depiction of motion blur on background vehicles.
- + Lighting and reflections on wet tiles feel highly realistic.
- + The imperfect framing enhances the candid street photography aesthetic.
- − The subject appears to be holding the bike rather than actively repairing it.
- − Anatomical issues with the subject's left hand and glove-like texture.
FLUX.1 [schnell]
- + Strong interaction between the subject and the bicycle handlebars.
- + Vibrant color palette and clear environmental details like signage.
- + Effective use of shallow depth of field for subject isolation.
- − The car in the background lacks the requested motion blur.
- − The overall image looks slightly more stylized and less like a candid snapshot.
Verdict: FLUX.1 Krea [dev] captured the 'candid' and 'motion blur' aspects of the prompt more effectively, resulting in a more authentic street photography feel. FLUX.1 [schnell] produced a cleaner image with better subject interaction, but it missed the specific kinetic energy requested in the background.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent depiction of ornate engraved plate armor with realistic lighting reflections.
- + Captures the atmosphere of a paladin with a stoic, battle-worn expression and appropriate scarring.
- + Subtle and professional execution of bokeh sparks in the background.
- − The beads in the hair are more like studs or clips rather than being woven into the braids.
FLUX.1 [schnell]
- + Very high level of detail on facial skin texture and individual beard hairs.
- + Striking, lifelike eye rendering that draws the viewer in immediately.
- − The composition is a bit tight, missing the opportunity to show the 'ornate plate armor' and 'leather straps' requested in the prompt.
- − The 'beads' in the hair look more like metallic tubes or hair cuffs, and are less frequent than implied.
Verdict: FLUX.1 Krea [dev] is the clear winner as it adheres to every element of the prompt, including the specific textures of the armor and leather straps which are mostly cut out of the frame in the other version. While FLUX.1 [schnell] has incredible facial detail, it fails to deliver the full 'portrait of a paladin' look by focusing too closely on the face and obscuring the equipment requested.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent grid layout for food photography that feels dynamic and professional.
- + Vibrant orange accents create high visual appeal and consistency.
- + Includes all requested sections (Appetizers, Mains) with a logical structure.
- − Text rendering is quite messy with significant gibberish and artifacts.
- − Repeats the 'Appetizers' heading twice instead of using the 'Pizza' category requested.
FLUX.1 [schnell]
- + Features much cleaner typography that is easier to read at a distance.
- + Accurately follows the category requirements including Pizza and Mains.
- + Effective use of white space to achieve the requested minimalist aesthetic.
- − The grid of photos feels a bit small and less integrated compared to Model A.
- − Food photography quality appears slightly lower resolution than Model A's food images.
Verdict: FLUX.1 Krea [dev] produces a more visually striking design with excellent color accents, but fails to include all specific categories correctly. FLUX.1 [schnell] adheres better to the prompt's structural requirements (Pizza/Mains) and offers a cleaner, more readable minimalist aesthetic, making it the more functional design choice.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent photorealistic texture on the bun and meat patty.
- + Perfect text rendering for all requested phrases including the price.
- + Creative and high-quality background composition with embers and fire.
- − The burger is not fully 'exploded' as requested, with many layers still touching.
FLUX.1 [schnell]
- + Highly dynamic 'exploded' effect with many flying ingredients.
- + Good lighting on the food items making them pop against the background.
- − Significant spelling error in the main title ('AGIC BURGER').
- − The price is rendered incorrectly as '€699' and '6.99' in a messy layout.
- − Ingredients are less photorealistic and appear slightly plastic or stylized.
Verdict: FLUX.1 Krea [dev] is the clear winner because it followed all text instructions perfectly, whereas FLUX.1 [schnell] failed to spell the product name correctly and botched the price. While FLUX.1 [schnell] had more motion in the burger pieces, the overall visual polish and accuracy of FLUX.1 Krea [dev] make it a much better advertisement.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent chalk texture with realistic smudges and dust on the board.
- + Strong adherence to the requested date and main title spelling.
- + Elegant, consistent cursive handwriting that maintains a cohesive aesthetic.
- − Hallucinates extra menu items and includes spelling errors like 'Gruffle' and 'Loman'.
- − Text begins to overlap and become cluttered toward the bottom of the board.
FLUX.1 [schnell]
- + Clearer spacing between lines of text.
- + Captures the cafe background environment more effectively.
- − Significant spelling failures including 'Pril' for April and 'Mushmnctionm'.
- − Handwriting looks more like a digital marker than actual chalk texture.
- − Doubles up on words and prices at the bottom, creating a nonsensical layout.
Verdict: FLUX.1 Krea [dev] is the clear winner as it successfully renders the requested chalk texture and most of the date correctly, despite some spelling errors in the menu items. FLUX.1 [schnell] fails significantly on prompt adherence, misspelling the month and generating garbled text for every menu item while lacking the authentic chalk-on-blackboard feel.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent photographic quality and lighting.
- + Highly detailed mechanical suit for the horse.
- + Very cinematic composition with the Earth backdrop.
- − Failed the primary prompt instruction; the astronaut is riding the horse, not vice-versa.
FLUX.1 [schnell]
- + Followed the specific instruction for the horse to be on top of the astronaut.
- + Effective surreal atmosphere with warm lighting.
- + Creative interpretation of a horse saddle on an astronaut.
- − Anatomical glitch with a small extra horse head appearing from the main horse's neck.
- − The astronaut's posture is somewhat disjointed and confusing.
Verdict: While FLUX.1 Krea (dev) produced a much more realistic and cinematic image, it completely failed the core logical requirement of the prompt (horse on top). FLUX.1 [schnell] successfully followed the 'horse riding astronaut' instruction despite some minor anatomical artifacts like the secondary head. Therefore, FLUX.1 [schnell] is the winner for prompt adherence.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent realization of the capybara's professional expression and clothing.
- + Clearly legible 'TAXI' text on the hat.
- + High-quality rendering of the city lights and taxi exterior.
- − Includes an extra human in the passenger seat not requested by the prompt.
- − The 'both paws on the wheel' instruction is not followed as one paw is resting.
FLUX.1 [schnell]
- + Successfully captures the requested 'bored' expression of the passenger.
- + The capybara's fur texture and lighting are very realistic.
- + Stronger 'inside the taxi' perspective as requested.
- − Only one paw is on the steering wheel.
- − The capybara's eyes are a bit misaligned/asymmetrical.
- − The 'TAXI' text on the hat is slightly warped.
Verdict: FLUX.1 Krea [dev] produces a more polished and cinematic image with better lighting, though it fails on the prompt constraint by adding an extra human in the front seat. FLUX.1 [schnell] adheres better to the requested composition and captures the 'bored' expression perfectly, though the technical execution of the capybara's face is slightly weaker.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Elegant border with high-quality cobweb details
- + Excellent lighting on the jack-o-lantern and surrounding environment
- + Crisp, beautiful gothic typography for the header
- − Several typos in the small text including 'Pasty', 'You are are', and 'Trights'
- − The banner text is crowded and contains repetitive elements
FLUX.1 [schnell]
- + Great atmospheric foggy background with a moodier sky
- + Included all required date and location details clearly at the bottom
- + Higher contrast in the color palette makes the text more readable
- − The text contains several hallucinations and repetitive lines of nonsensical characters
- − The border is very thin and lacks the thorns mentioned in the prompt
- − Composition feels a bit bottom-heavy due to text spacing
Verdict: FLUX.1 Krea [dev] produces a much more polished and artistic visual with superior lighting and a beautiful border, though it fails significantly on the actual text content ('Pasty Halloween'). FLUX.1 [schnell] manages to include more of the requested event details, but the image is hampered by numerous text hallucinations and a less intricate design. FLUX.1 Krea [dev] is the winner for its superior visual quality and adherence to the 'vintage gothic' aesthetic, despite the typos.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Features both 'JAPAN' and 'SUSHI' text clearly as requested.
- + The isometric diorama base and lighting are well-executed.
- + Strong material textures on the salmon and rice.
- − The flag icon is interpreted as a physical prop rather than a graphic element.
FLUX.1 [schnell]
- + Includes the flag as a clean graphic icon at the top.
- + Soft, pleasing 3D cartoon aesthetics match the requested style.
- + The food model includes realistic details like cucumber and seasoning.
- − Missing the word 'SUSHI' from the typography requirements.
- − The text 'JAPAN' is somewhat faint against the light background.
Verdict: FLUX.1 Krea [dev] followed the text prompt more accurately by including both required words, whereas FLUX.1 [schnell] omitted the word 'SUSHI'. While FLUX.1 [schnell] captures the 'miniature' aesthetic and graphic layout slightly better, FLUX.1 Krea [dev] is the more complete adherence to the specific prompt instructions.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Successfully included all requested animals including a fox and a kitten.
- + Excellent application of god rays and warm sunrise lighting.
- + High level of detail on the fur texture and the butterflies.
- − Failed to include the 'baby bunny' specifically, providing two kittens instead.
- − The anatomy of the jumping kitten in the background is slightly awkward.
FLUX.1 [schnell]
- + Captures a very emotive, wide-eyed look that fits the 'joyful' vibe.
- + Good depth of field with realistic foreground wildflowers.
- + Fur texture is soft and well-rendered.
- − Missing the 'baby bunny' from the prompt.
- − The central animals look like generic cat-fox hybrids rather than distinct species.
- − Anatomy of the kitten's limb holding the other kitten is poorly defined.
Verdict: FLUX.1 Krea [dev] is the winner as it adhered closer to the specific animal types requested, although both models failed to generate the bunny. FLUX.1 Krea [dev] also produced much better atmospheric effects with the requested god rays, whereas FLUX.1 [schnell] had more anatomy issues and generic-looking creatures.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Excellent engraving-style detail on the cloche and banner.
- + Accurately rendered the 'Est. 1720' text.
- + Beautiful vintage texture and atmospheric steam.
- − Misspelled 'Florian' as 'Flanelin'.
- − The steam plume is a bit thick, resembling a chimney.
FLUX.1 [schnell]
- + Clean, minimalist vector aesthetic.
- + Good layout and balance in a circular emblem style.
- − Failed the date text, rendering 'Est. 7720' instead of 1720.
- − Major misspelling of 'Florian' as 'Framilan'.
- − Omitted the requested steam element.
Verdict: FLUX.1 Krea [dev] is the clear winner for its artistic execution and higher adherence to the specific prompt details, despite a spelling error. FLUX.1 [schnell] failed to include the steam, had more significant typos in the brand name, and completely botched the date requested.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.1 Krea [dev]
- + Strong text legibility for the main title
- + Clean flat-vector aesthetic with a clear color palette
- − Icons do not match the specific mission steps requested
- − The flow of information is disorganized and nonsensical
FLUX.1 [schnell]
- + Better overall infographic composition and symmetry
- + Icons more closely represent astronomical concepts requested
- − Text consists primarily of illegible gibberish
- − The 'Saturn V' rocket icon is poorly formed and lacks crispness
Verdict: FLUX.1 [schnell] follows the infographic layout requirements more effectively, creating a logical visual flow despite the garbled text. FLUX.1 Krea [dev] produces much clearer text, but the individual elements fail to follow the requested sequential mission steps, resulting in a confusing collection of icons.
Explore each model
Black Forest Labs' 12 billion parameter distilled image generation model optimized for speed, capable of generating high-quality images in just 4 inference steps