Black Forest Labs' premium multimodal flow transformer with greatly improved prompt adherence and typography generation for in-context image generation and editing without compromise on speed
Settled by community votes across 16 shared challenges, with an AI judge weighing in on each.
FLUX.1 Kontext [max]
#23 of 62 in Text-to-Image
HiDream I1 Full
#60 of 62 in Text-to-Image
Where the votes landed
FLUX.1 Kontext [max]
0%
win rate
Ties
0%
HiDream I1 Full
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent adherence to lighting instructions with realistic caustics and window shadows.
- + Very sharp textures on the blue sphere and wooden table.
- + Perfect logical placement of objects within the scene.
- − The plant is more overhead than strictly behind the cube.
- − The blue sphere has a slightly rough, crystalline texture rather than being a smooth ball.
HiDream I1 Full
- + Successfully positions the plant directly behind the glass cube.
- + Clean, minimalist composition with smooth object surfaces.
- − The blue sphere is levitating inside the cube, which feels physically unnatural.
- − The glass thickness and refractions are less realistic compared to Model A.
- − Lighting feels a bit more generic despite the visible window.
Verdict: FLUX.1 Kontext [max] is the winner due to its superior rendering of physical properties, specifically the way the light interacts with the glass and wood and the realistic weight of the sphere sitting on the bottom of the cube. While HiDream I1 Full followed the spatial instruction for the plant slightly better, its sphere is floating unnaturally and the overall image quality lacks the photorealistic depth of its competitor.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent depiction of falling rain and realistic wet textures.
- + The bicycle mechanics and the man's interaction with the chain are physically plausible.
- + Successfully captures the 'imperfect framing' and cinematic bokeh of a 50mm lens.
- − The man's heritage appears more ambiguous than specifically Japanese.
- − The scale of the bicycle seems slightly oversized compared to the crouching man.
HiDream I1 Full
- + Stronger adherence to the requested Japanese ethnicity for the subject.
- + Effective shallow depth of field and beautiful ground reflections.
- + Good composition with the city background providing depth.
- − The bicycle geometry is broken, with the frame passing through the man's leg.
- − The man is sitting on the frame or a phantom seat rather than 'repairing' it actively.
- − Does not show visible falling rain, only wet surfaces.
Verdict: FLUX.1 Kontext [max] is the superior image because it depicts the requested action (repairing) with high physical accuracy and creates a convincing rainy atmosphere. HiDream I1 Full captures the ethnicity better but fails significantly on the technical details, with the bicycle frame clipping through the man's body and a lack of actual rain particles.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Naturalistic lighting and highly realistic skin texture
- + Excellent engraving details on the plate armor
- + Subtle and organic integration of braids and hair fibers
- − The sparks look slightly like flat streaks rather than glowing embers
- − Misses the specific request for beads in the hair
HiDream I1 Full
- + Strong adherence to the 'beads in hair' and 'faint scars' prompt elements
- + Good composition with a clear torch visible in the background
- + Distinct texture on the leather straps and metal buckles
- − Lighting feels overly artificial and high-contrast
- − Armor engraving looks a bit generic and repetitive compared to Model A
- − Facial scars look like surface-level paint rather than integrated skin texture
Verdict: FLUX.1 Kontext [max] produces a much more lifelike and cinematic image with superior texture work and lighting, though it fails to include the requested beads. HiDream I1 Full adheres more closely to the specific details of the prompt like the beads and scars, but the overall image quality feels more like a digital illustration than a realistic photograph. FLUX.1 Kontext [max] is the winner for its professional-grade visual fidelity and realistic depth of field.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent professional layout with a clean graphic design feel
- + Strong grid-based composition that makes the menu feel organized
- + Includes a logo and diverse vibrant accents like the yellow footer
- − The text is mostly gibberish with inconsistent fonts for section headers
- − Food variety is lacking as most photos are just small variations of the same pizza
HiDream I1 Full
- + Successfully included specific headers for 'Appetizers / Pizza' and 'Mains'
- + Higher legibility of central text elements
- + Better color saturation in the food photos
- − Very basic layout that lacks the 'professional' feel of a modern menu
- − Alignment issues with the text overlapping images slightly at the bottom
- − Repetitive food photos despite having different category labels
Verdict: FLUX.1 Kontext [max] creates a much more convincing professional layout that looks like a real graphic design piece, though the text is garbled. HiDream I1 Full adheres better to the specific text headers requested but fails to deliver a 'modern minimalist' aesthetic, resulting in a cluttered and amateurish appearance. FLUX is preferred for its superior composition and professional vibe.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Perfect text accuracy for all requested menu items and prices.
- + Highly realistic chalk texture with smudge marks and natural handwriting styles.
- + Accurately interpreted the cutoff prompt to complete 'Brown Butter Chocolate Chip Cookies'.
- − The title uses a print-style block letter rather than the requested elegant cursive.
- − The background cafe details are slightly blurry and less defined than the board itself.
HiDream I1 Full
- + Elegant cursive title at the top matches the stylistic request better than Model A.
- + High visual clarity and clean composition with nice chalk illustrations.
- − Complete failure to follow the specific menu item text requested, resulting in gibberish.
- − The text appears too much like a digital font rather than natural handwriting.
- − Failed to include the specific pricing from the prompt.
Verdict: FLUX.1 Kontext [max] provides a highly functional and accurate result, correctly rendering all specific text and prices with a very realistic chalk-on-blackboard texture. While HiDream I1 Full captured the 'elegant cursive' request for the title better, it failed entirely on the specific content of the menu, producing nonsensical text instead of the requested items.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent anatomical detail and texture on the horse.
- + Consistent lighting between the astronaut and the environment.
- + Clean, high-resolution rendering without significant artifacts.
- − Failed the negative constraint; the astronaut is riding the horse instead of vice versa.
- − The composition is a bit centered and standard for this prompt.
HiDream I1 Full
- + Dynamic composition with the background planet and clouds.
- + Good use of color contrast between the white horse and the dark space.
- − Failed the negative constraint regarding the position of the character and animal.
- − The back legs of the horse are anatomically distorted.
- − Low-quality texture on the planet surface and clouds.
Verdict: Both FLUX.1 Kontext [max] and HiDream I1 Full failed the core logic test of the prompt, which specifically requested the horse on top of the astronaut. However, FLUX.1 Kontext [max] is the superior image due to its significantly higher technical quality, realistic horse anatomy, and more coherent lighting, whereas HiDream I1 Full exhibits major anatomical flaws in the horse's legs.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent photorealistic fur texture and lighting
- + Cinematic composition with realistic background bokeh
- + Accurate depiction of the passenger as requested
- − The capybara only has one visible paw on the wheel instead of both
- − Composition feels slightly cramped with the side of the car blocking the view
HiDream I1 Full
- + Successfully placed both paws on the steering wheel as requested
- + Clearer view of the entire scene including the passenger and phone
- + Well-rendered 'TAXI' text on the cap
- − The passenger's proportions feel slightly off compared to the seat
- − Lighting is a bit more artificial and 'HDR' style compared to image A
- − Anatomical weirdness where the capybara's right arm seems to merge with its body
Verdict: Both models followed the prompt's unique requirements very well. FLUX.1 Kontext [max] produced a more professional, photorealistic image with superior textures, while HiDream I1 Full adhered more closely to the specific instruction regarding the paws on the steering wheel despite having slightly less realistic lighting.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent text rendering with no spelling errors.
- + Perfect adherence to all specific event details included in the prompt.
- + Atmospheric and polished cinematic lighting that fits the 'gothic' theme.
- − The 'Arches, NYC' location is repeated twice at the bottom.
- − The background is very dark, making the twisted trees slightly difficult to see.
HiDream I1 Full
- + Strong composition with a clear parchment paper effect in the center.
- + Good use of the spiderweb and thorn border elements described.
- − Failed significantly on text accuracy, producing gibberish like 'Yootmber' and '30 cm'.
- − Missing the word 'Invitation' from the main title.
- − The overall style is more 'cartoony' than the requested 'cinematic' and 'vintage gothic' look.
Verdict: FLUX.1 Kontext [max] is the clear winner as it accurately rendered almost every piece of specific text requested, including the date, time, and location. While HiDream I1 Full captured the parchment aesthetic well, its failure to generate legible or accurate event details makes it unusable as a functional invitation.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Successfully added a full head of hair that looks natural and thick.
- + Excellent preservation of the original facial identity and expression.
- + Maintains the lighting and background of the source image perfectly.
HiDream I1 Full
- + Successfully changed the hairstyle to a mullet-style length at the back.
- − Completely changed the background and environment.
- − Introduced severe artifacts like a bizarre black object protruding from the beard.
- − Modified the facial features, making the person unrecognizable from the original.
Verdict: FLUX.1 Kontext [max] performed a near-perfect edit, seamlessly adding realistic hair while keeping every other aspect of the image identical to the source. HiDream I1 Full failed significantly, generating a completely new background, altering the subject's face, and adding nonsensical artifacts over the beard.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent 3D soft textured aesthetic that perfectly matches the 'cartoon scene' prompt.
- + Accurate text rendering with a stylized font that fits the theme.
- + High-quality material rendering on the sushi and diorama base.
- − Missing the requested flag icon.
- − The 'diorama base' is a bit plain compared to Model B.
HiDream I1 Full
- + Stronger isometric diorama composition with more variety in sushi pieces.
- + Very clean, bold white typography.
- + Good adherence to the 45-degree top-down perspective.
- − Missing the requested flag icon.
- − The text 'SUSHI' is relatively small and less impactful than Model A.
- − Lighting feels slightly more flat and less 'refined' than Model A.
Verdict: Both models failed to include the flag icon, but FLUX.1 Kontext [max] captured the requested 'soft refined textures' and 3D cartoon aesthetic much better than HiDream I1 Full. While HiDream I1 Full offered a more complex sushi arrangement, FLUX.1 Kontext [max] produced a more visually cohesive and professionally rendered image that felt more like a miniature 3D scene.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent adherence to all prompts, including a clear news desk, a microphone, a dog, and a hockey stick.
- + Captures the 'caricature' art style perfectly with exaggerated features and vibrant colors.
- + Preserves the subject's outfit (denim shirt and black top) accurately from the source.
- − The hockey stick is positioned awkwardly in the top right corner without much context.
HiDream I1 Full
- + Successfully creates a stylized caricature version of the person in the source image.
- + Includes a dog and preserves the facial structure of the original subject well.
- − Fails to include any clear elements of a TV show anchor (no microphone, desk, or studio).
- − Completely misses the 'hockey' requirement of the prompt.
- − The 'hockey puck' on the dog's mouth looks more like a weird mouthpiece or gag.
Verdict: FLUX.1 Kontext [max] is the clear winner as it successfully incorporated every element of the prompt (anchor job, dog, and hockey) into a cohesive and humorous caricature. HiDream I1 Full failed to include the news anchor theme or hockey elements, resulting in a generic portrait of a woman with a dog.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Successfully included all four requested animals (dog, cat, rabbit, fox).
- + Excellent rendering of soft, backlit fur and natural meadows textures.
- + Consistent lighting and coherent 'god rays' following the sun's position.
- − The rabbit's hands/paws are slightly anatomically awkward.
HiDream I1 Full
- + Vibrant colors and a very cute, symmetrical composition.
- + High contrast and sharp focus on the central animals.
- − Missed the rabbit entirely, instead providing two kittens.
- − Animals look more like digital illustrations than the requested 'hyper-photorealistic' style.
- − The sun/glare is overly blown out and lacks detail.
Verdict: FLUX.1 Kontext [max] followed the prompt more accurately by including all four distinct animal types, whereas HiDream I1 Full omitted the rabbit and duplicated the kitten. Moreover, FLUX.1 Kontext [max] achieved a much more realistic, photographic aesthetic compared to the somewhat airbrushed, illustrative look of HiDream I1 Full.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent preservation of the original subjects' poses, clothing patterns, and composition.
- + Perfectly captures the Studio Ghibli aesthetic with soft watercolor textures and gentle line work.
- + Maintains the specific facial expressions of the 'distracted boyfriend' meme characters while stylizing them.
- − The background is slightly more washed out than typically seen in lush Ghibli films.
HiDream I1 Full
- + Successfully adopts a cute anime art style with vibrant colors.
- + Adds creative details like floral patterns and a hat that fit a summer anime theme.
- − Fails to preserve the core composition and distinct character dynamics of the source image.
- − The man's beard and the woman's hair style are significant departures from the original photographic subjects.
- − Loss of the iconic 'shocked' expression on the girlfriend character.
Verdict: FLUX.1 Kontext [max] is the clear winner as it successfully transforms the specific image provided into the requested style while keeping the meme's identity intact. HiDream I1 Full produces a generic anime-inspired scene that loses the character likenesses and the specific narrative moment of the distracted boyfriend meme.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent preservation of the source subject and background.
- + Accurately adds motion to hair while maintaining the woman's identity.
- + Includes the dog and leash from the original image.
- − The added leaves are small and relatively sparse.
- − The left arm positioning is slightly altered, becoming stiff.
HiDream I1 Full
- + Great sense of dynamic motion through hair and clothing pose.
- + Highly energetic and lively interpretation of the prompt.
- − Fails to preserve the source image, completely removing the dog.
- − Changes the woman's facial features and identity significantly.
- − Background details like the flowers and bridge have been lost or changed.
Verdict: FLUX.1 Kontext [max] successfully performs the edit by adding wind-blown hair and leaves while keeping the original woman and her dog intact. HiDream I1 Full fails as an image editor because it completely regenerates the scene, resulting in the loss of the dog and a change in the woman's appearance, despite capturing the requested 'energy' well.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent typography including the circumflex accent on 'Caffè'
- + Superior texture that gives a genuine vintage paper feel
- + Complete adherence to all prompt elements including the specific name.
- − The steam is a bit small relative to the cloche size.
HiDream I1 Full
- + Clean vector-style execution
- + Good use of the 'Est. 1720' banner
- − Completely failed to include the brand name 'Caffè Florian'
- − The steam appearing through the solid cloche creates a confusing visual logic
- − Lacks the requested 'subtle texture' compared to Model A.
Verdict: FLUX.1 Kontext [max] perfectly followed the prompt, including the brand name and the specific vintage texture requested. In contrast, HiDream I1 Full completely omitted the main text 'Caffè Florian' and produced a more generic, flat vector that lacked the authentic retro charm of the first image.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
FLUX.1 Kontext [max]
- + Excellent typography rendering for 'APOLLO 11' and the astronaut names.
- + Matches the requested navy, white, and muted red NASA palette perfectly.
- + Includes clever vector silhouettes of the three astronauts at the bottom.
- − Failed to provide all 6 specific steps, missing 'Launch' and 'Descent'.
- − The logical flow of the infographic is confusing, with arrows pointing in multiple directions.
- − The rocket icon looks like a generic toy rather than a Saturn V.
HiDream I1 Full
- + Successfully captured more of the requested infographic steps with clear labels.
- + The Saturn V rocket icon is much more accurate to the actual mission hardware.
- + Strong use of high-contrast flat vector style that feels like professional graphic design.
- − Contains significant text errors like '+ UNG' and '+ ICON' which appear to be hallucinations of the prompt text.
- − Composition is a bit cluttered with overlapping elements in the center.
- − Failed to include the lunar module landing and descent phases.
Verdict: Both models failed to include all six specific steps requested in the prompt. FLUX.1 Kontext produced a much cleaner, more aesthetically pleasing poster with accurate text for the astronauts, but the 'Translunar' and 'Lunar' graphics are illogical. HiDream I1 Full followed the 'Saturn V' instruction better and attempted more steps, but was marred by nonsensical garbled text within the infographic.
Explore each model
HiDream AI's 17B parameter text-to-image model using sparse diffusion transformer with mixture of experts, achieving state-of-the-art image generation quality with strong prompt following