OpenAI's previous image generation model that accepts both text and image inputs and produces image outputs
Settled by community votes across 20 shared challenges, with an AI judge weighing in on each.
GPT Image 1
#32 of 62 in Text-to-Image
Seedream 4.5
#9 of 62 in Text-to-Image
Where the votes landed
GPT Image 1
0%
win rate
Ties
0%
Seedream 4.5
0%
win rate
Challenge by challenge
The strongest take from each model on every shared challenge, with the AI judge's read.
Geometric Composition
Text-to-Image“A glass cube on a wooden table. Inside the cube is a small blue sphere. On top of the cube sits a red book. A green plant is behind the cube, partially visible through the glass. Soft window light from the left.”
AI Judge Analysis
GPT Image 1
- + Excellent photographic quality and realistic material textures.
- + Balanced composition with high clarity.
- + Accurate interpretation of 'small blue sphere' and red book.
- − The sphere appears to be floating slightly rather than resting on the base.
- − The glass cube has extremely thick, almost frame-like edges.
Seedream 4.5
- + Perfect adherence to the lighting instruction with realistic shadows and caustics.
- + Highly realistic glass properties including refraction and reflections.
- + Accurate placement of the sphere resting on the bottom surface.
- − The plant in the background is quite blurry compared to Image A.
- − The blue sphere is closer to a marble, which might be seen as 'tiny' rather than 'small'.
Verdict: Both models followed all prompt instructions perfectly. GPT Image 1 has a very clean, studio-like aesthetic with bold colors, while Seedream 4.5 exhibits superior physics, specifically regarding how light interacts with the glass cube and the sphere. Seedream 4.5 is the winner for its more natural lighting and more realistic glass rendering.
Man and Car in California
Editing“Make a photo of the man driving the car down the California coastline”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the car's body, lighting, and specific model details.
- + High-quality, realistic integration of the person into the driver's seat.
- + Dynamic motion blur on the wheels and road suggests actual movement.
- − The person's facial features and distinctive hair style are slightly modified from the source image.
Seedream 4.5
- + Strong preservation of the person's identity, including his specific plaid coat, scarf, and hairstyle.
- + Captures the full outfit of the person including his green cargo pants and boots.
- − The car door logic is broken, appearing to open outwards despite the car being in motion.
- − Composition is awkward with the car door cutting off the front of the vehicle.
Verdict: GPT Image 1 produces a more cohesive and professional-looking final image with realistic motion blur and lighting, successfully placing the car in a California coastline setting while keeping the car's model consistent. Seedream 4.5 does a better job of preserving the subject's exact clothing and identity, but fails significantly on the structural logic of the car itself, resulting in a nonsensical open door while driving.
Candid Street Photography
Text-to-Image“A candid street photo of an elderly Japanese man repairing a red bicycle in light rain, reflections on wet pavement, shallow depth of field, 50mm lens, natural skin texture, imperfect framing, motion blur from passing cars, cinematic but realistic, no stylization.”
AI Judge Analysis
GPT Image 1
- + Excellent natural skin texture and realistic wetness on clothing.
- + Captures a melancholy, cinematic mood with authentic lighting.
- + Strong adherence to the shallow depth of field and 50mm lens look.
- − Lacks the requested motion blur from passing cars; cars in background are static.
- − The bicycle mechanics are slightly physically nonsensical around the hub.
Seedream 4.5
- + Perfectly captures the requested motion blur of a passing car.
- + Dynamic composition with vibrant neon reflections on wet pavement.
- + Includes visible raindrops on the man's jacket, enhancing the 'light rain' setting.
- − The man's hands have structural errors and anatomical distortions.
- − Slightly less 'candid' and more 'stylized' compared to the prompt's request for no stylization.
Verdict: GPT Image 1 excels in skin texture and the realistic rendering of a person in rain, but completely misses the requested 'motion blur from passing cars.' Seedream 4.5 captures the motion blur and the urban atmosphere much better, though it suffers from significant anatomical issues with the hands. Seedream 4.5 is the winner for following all technical aspects of the prompt, including the motion blur and environmental reflections.
Fantasy Warrior
Text-to-Image“Close portrait of a battle-worn paladin in ornate engraved plate armor, hair braided with small beads, faint scars and dirt on the skin, warm torchlight reflecting off metal, shallow depth of field, bokeh sparks, lifelike eyes, highly detailed texture on leather straps and cloth underlayer.”
AI Judge Analysis
GPT Image 1
- + Excellent depiction of weathered, battle-worn plate armor with realistic patina.
- + Dramatic, moody lighting that emphasizes the character's grit and expression.
- + High-quality skin texture including subtle scars and dirt smears.
- − The beads in the hair are very subtle and blend in with the braid texture.
- − The overall image is slightly darker, losing some of the specified cloth and leather details in the shadows.
Seedream 4.5
- + Outstanding adherence to specific details like the beads in the hair and stitched leather straps.
- + Bright, warm torchlight effects with clear bokeh and fire elements.
- + High-fidelity rendering of skin pores, facial hair, and lifelike eyes.
- − The armor looks slightly more 'golden' and pristine than typically expected for 'battle-worn' plate.
- − The focus is extremely shallow, making the hair beads on the right side of the head look a bit detached.
Verdict: Both models followed the prompt exceptionally well, but Seedream 4.5 captures the technical details of the leather straps and the hair beads more distinctly than GPT Image 1. While GPT Image 1 excels at the 'battle-worn' atmosphere with grittier armor, Seedream 4.5 provides a sharper, more vibrant image that highlights all the requested textures and lighting elements.
Modern Clean Menu
Text-to-Image“Modern minimalist restaurant menu design, white background with colorful food photos in grid, sections for appetizers/pizza/mains, bold sans-serif fonts, vibrant accents, clean professional layout for casual dining.”
AI Judge Analysis
GPT Image 1
- + Excellent typography rendering with almost perfect spelling for headers.
- + High-quality, vibrant food photography that follows a consistent top-down style.
- + Balanced grid layout that feels like a real professional menu.
- − Placeholder description text is repetitive and nonsensical.
- − The layout is cropped at the bottom, cutting off the lower food images.
Seedream 4.5
- + Clean minimalist separation of sections using color-coded borders.
- + Follows the request for 'vibrant accents' through the frame colors.
- + Coherent food photography that matches the specified categories.
- − Text rendering is poor with many typos in the item lists.
- − Layout is less efficient with significant whitespace on the left and right.
- − The prices are unrealistic for a casual dining menu ($79 - $99).
Verdict: GPT Image 1 is the clear winner as it produces a much more professional and realistic menu layout with superior text legibility and high-quality food photography. While Seedream 4.5 interprets the 'vibrant accents' well through colored borders, its struggles with basic spelling and overpriced menu items make it less usable as a design template.
Magic Burger Explosion: Fiery Photorealism Challenge
Text-to-Image“Ad for 'Magic Burger'. Dynamic, exploded burger with all components (bun, patty, cheese, lettuce, tomato, sauce) suspended in mid-air. Emphasize photorealistic detail and a sense of motion. Dark, fiery background with glowing embers. Integrate text: 'MAGIC BURGER' as a prominent title, 'LIMITED TIME ONLY' as a secondary message, and '€6.99' in a starburst, all rendered with a fiery, glowing effect.”
AI Judge Analysis
GPT Image 1
- + Excellent layout with a clearly exploded view showing each distinct layer
- + Clean and legible fiery text effects
- + Great texture on the meat patty and fresh vegetables
- − The price tag text is incorrect, displaying '.99' instead of '€6.99'
- − The composition is a bit static compared to the requested motion
Seedream 4.5
- + Highly dynamic composition with motion-blurred ingredients and a tilted angle
- + Accurate text rendering including the full price and currency
- + Impressive fire effects within the lettering and background
- − Less of an 'exploded' view as the main stack remains largely intact
- − The lettuce and bottom bun are slightly blurry compared to the top halves
Verdict: GPT Image 1 succeeds in providing a clear exploded view and very high detailed textures, but fails significantly on the specific pricing text. Seedream 4.5 captures the 'sense of motion' much better through dynamic angles and motion blur while also correctly rendering all required text elements. Seedream 4.5 is the preferred choice for a commercial ad due to its energy and accuracy.
Chalkboard Menu
Text-to-Image“Handwritten-style chalkboard menu in a cozy café, all text rendered in the exact same realistic chalk handwriting style with natural variations in letter size, slight slant, and chalk texture — no printed or digital fonts anywhere on the board. Title at the top in elegant cursive chalk handwriting: ‘TODAY’S SPECIALS – APRIL 30, 2026’. Below it, three menu items also in the same handwritten chalk style: ‘Truffle Mushroom Risotto – $24’, ‘Grilled Octopus with Lemon & Herbs – $28’, ‘Brown Butter Chocolate Chip Cookies – $9’. At the very bottom, smaller text in the identical handwritten chalk style (slightly smaller but still clearly legible with the same handwriting characteristics): ‘All items made fresh daily • Ask about our gluten-free options’. Warm ambient café lighting, visible chalk dust and smudges, realistic handwriting imperfections, no clean printed text anywhere.”
AI Judge Analysis
GPT Image 1
- + Excellent text legibility and accuracy
- + Perfectly followed the specific menu item list including the cut-off 'Brown But' as a prompt instruction
- + Convincing chalk texture on the individual letters
- − The handwriting style is a bit too uniform, bordering on looking like a digital font
- − The cursive requested for the title is not truly present, opting for a clean print style instead
Seedream 4.5
- + Features a more authentic, varied chalk handwriting style with a nice slant
- + Includes a better 'cursive' interpretation for the top heading as requested
- + More creative composition showing the café environment around the board
- − Repeats the price '$24' unnecessarily on the first item
- − Failed to properly handle the cut-off prompt, completing the word 'Cookies' instead of stopping where the prompt ended
- − Slightly less crisp text rendering compared to Model A
Verdict: GPT Image 1 followed the technical constraints of the prompt more accurately, specifically adhering to the cut-off text 'Brown But' and accurately rendering all listed prices without repetition. Seedream 4.5 captured the 'handwritten' aesthetic much better with more stylistic flair and a more realistic board texture, though it struggled with logic by repeating text segments and ignoring the prompt's structural ending.
Pose & Character Mashup
Editing“Use Image 1 as the exact pose reference and Image 2 as the character reference. Recreate the person/character from Image 2 in the exact dynamic pose and body position from Image 1. Keep the exact face, hair, clothing style/details, and expression from Image 2. Match the lighting and environment of Image 1. The final image must show the character from Image 2 performing the precise action/pose from Image 1 with perfect anatomy and natural integration.”
AI Judge Analysis
GPT Image 1
- + Matches the specific 'black sweatshirt and checkered scarf' attire from Image 2
- + Adheres to the yellow background and red stand environment
- − The facial likeness is poor and distorted
- − Anatomy is broken with a severed hand and missing lower leg segment
- − The checkered scarf lacks the intricate circular patterns seen in Image 2
Seedream 4.5
- + Excellent facial likeness and hair recreation from Image 2
- + Accurately recreates the circular patterns on the scarf and the text on the sweatshirt
- + Anatomy is much more coherent and natural despite the complex pose
- − The arm position is slightly less dynamic than the source pose
- − One foot has six toes
Verdict: Seedream 4.5 is the clear winner as it successfully captures the specific character details from Image 2, including facial features, clothing text, and scarf patterns, while maintaining a high-quality image. GPT Image 1 fails significantly on anatomy, resulting in severing limbs and providing a poor likeness of the subject.
The Reversed Rodeo
Text-to-Image“Horse riding astronaut in space — horse on top, not vice versa. Surreal, highly detailed, cinematic.”
AI Judge Analysis
GPT Image 1
- + Excellent anatomical details on the horse's musculature and textures.
- + Cinematic lighting with a gritty, realistic aesthetic.
- + Well-integrated composition that feels like a real film still.
- − Failed the negative constraint; the astronaut is riding the horse, not the requested 'horse on top' position.
Seedream 4.5
- + Vibrant color palette with beautiful nebulae and lighting effects.
- + Dynamic pose with 'star dust' trailing from the horse's hooves.
- + Good attention to the golden visor reflection.
- − Failed the negative constraint; the horse is being ridden by the astronaut.
- − Minor anatomical issues where the astronaut's leg meets the horse.
Verdict: Both models failed the specific spatial instruction for the horse to be 'on top' of the astronaut, instead defaulting to the common trope of an astronaut riding a horse. GPT Image 1 provides a more detailed and realistically textured composition, while Seedream 4.5 offers a more fantasy-like, colorful aesthetic with higher visual contrast.
Outfit Transfer Challenge
Editing“Use Image 1 as the base person. Dress them in the exact elaborate outfit from Image 2 (including all layers, accessories, jewelry, and shoes). Carefully adapt the clothing to the body shape and pose in Image 1 while maintaining realistic fabric behavior, correct proportions, and perfect lighting/shadow matching. Keep the person’s exact face, hair, and background completely unchanged.”
AI Judge Analysis
GPT Image 1
- + Excellent fabric textures and realistic lighting integration
- + High identity preservation of the face
- + Captures the intricate pattern of the scarf accurately
- − Significantly alters the vitiligo pattern on the forehead
- − Crop is much tighter than the original source image
- − Background is slightly simplified compared to the source
Seedream 4.5
- + Perfectly preserves the person's face, hair, and vitiligo patterns
- + Maintains the original image composition and wide background
- + Captures all accessories including the gold watch and ring from Image 2
- − Logic error in clothing layering where the skin is visible behind the scarf despite wearing a black shirt
- − Lower image resolution/clarity compared to Model A
- − The lighting on the coat feels slightly flat compared to the scene
Verdict: Seedream 4.5 is the winner for its superior ability to maintain the source person's unique features and the original image's composition, although it makes a mistake by showing skin through the shirt. GPT Image 1 produces a more polished and higher-quality render, but it fails the 'source preservation' requirement by changing the subject's vitiligo pattern and cropping the image.
The Capybara Taxi Driver
Text-to-Image“Photorealistic scene inside a yellow New York taxi at night. A capybara is driving, wearing a yellow taxi driver cap and a dark jacket. It has a calm, professional expression and both front paws on the steering wheel. In the back seat sits a human businesswoman in a coat, looking at her phone with a completely normal, bored expression (as if this is just another normal ride). Through the windows you can see the streets of Manhattan at night with blurred lights. Realistic taxi interior, photorealistic, detailed fur and fabric, 35mm lens, night lighting with reflections, shallow depth of field.”
AI Judge Analysis
GPT Image 1
- + Excellent close-up detail on the capybara's fur and the leather texture of the hat
- + Cinematic lighting that accurately reflects a nighttime city environment
- + Strong adherence to the 'calm, professional expression' request
- − The passenger is slightly more out of focus compared to the other model
Seedream 4.5
- + Perfect adherence to all prompt elements including both paws on the wheel and the seatbelt
- + Clearer visibility of the Manhattan background through the front windshield
- + The passenger's 'bored' expression is very well executed
- − The capybara's fur looks slightly more matted or less realistic than Model A
- − The cap appears a bit more like a basic baseball cap than a traditional driver's cap
Verdict: Both models performed exceptionally well on this complex prompt. GPT Image 1 offers a more cinematic and photorealistic texture for the capybara itself, while Seedream 4.5 provides a better overall composition that captures more of the car interior and background detail with perfect logic for the passenger's expression. Seedream 4.5 is the likely winner for better framing and clarity of the human passenger.
The Halloween Invitation
Text-to-Image“Vintage gothic Halloween party invitation. Dark parchment poster, spooky border with webs and thorns, central glowing jack-o-lantern, bats, twisted trees, moody night sky. Add elegant gothic title text saying "Halloween Party Invitation", a small scroll banner saying "You are invited to a night of frights", and event details at the bottom: Date: 30.10.2026 Time: 7pm Location: The Arches, NYC Spooky but polished, cinematic lighting, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent typography style that perfectly fits the vintage gothic theme
- + Atmospheric and moody lighting that feels cohesive with the parchment texture
- + High level of detail in the border illustrations including webs and thorns
- − Missed the 'Time' detail by merging it with the location line
- − Formatting of the bottom text is slightly cramped
Seedream 4.5
- + Accurately followed all text instructions including the specific time
- + High visual clarity and sharp, cinematic rendering of the jack-o-lantern
- + Composition feels balanced with the thorn and web motifs clearly placed
- − The font choice for the bottom text is a bit modern and generic for a 'vintage gothic' prompt
- − The parchment texture is less apparent compared to its competitor
Verdict: Both models performed exceptionally well, but Seedream 4.5 is the overall winner for its perfect adherence to the text prompts, including the 'Time' field which GPT Image 1 stumbled on. While GPT Image 1 captured a more authentic vintage parchment aesthetic, Seedream 4.5 delivered a more polished and usable graphic design with better clarity.
Bald man challenge
Image Editing“Give the person a full, thick head of natural hair with realistic texture, density, and a natural hairline. Preserve facial features and lighting.”
AI Judge Analysis
GPT Image 1
- + Adds a very thick, dense head of hair as requested.
- + Preserves the background and clothing perfectly.
- − The hairline is unnaturally low, cutting deep into the forehead.
- − The lighting on the new hair doesn't match the directional sun in the rest of the scene.
- − The hair texture looks coarse and slightly artificial.
Seedream 4.5
- + Features a very realistic, natural hairline that respects the person's anatomy.
- + The lighting on the hair perfectly matches the soft, directional sunlight of the original photo.
- + Preserves facial features and the rest of the image with high fidelity.
- − The hair could be considered slightly less 'full' than 'thick' in terms of volume compared to the other model.
Verdict: GPT Image 1 adds a significant amount of hair but fails on the integration, creating an anatomically incorrect low hairline and mismatched lighting. Seedream 4.5 provides a much more professional edit, with a realistic hairline and lighting that makes the hair appear as if it were part of the original photograph.
Isometric Miniature Diorama Scenes
Text-to-Image“Create a clear, 45° top-down isometric miniature 3D cartoon scene of Japan's signature dish: sushi, with soft refined textures, realistic PBR materials, gentle lighting, on a small raised diorama base with minimal garnish and plate. Solid light blue background. At top-center: 'JAPAN' in large bold text, 'SUSHI' below it, small flag icon. Perfectly centered, ultra-clean, high-clarity, square format.”
AI Judge Analysis
GPT Image 1
- + Excellent text rendering with clean, centered alignment.
- + Soft, pleasing 3D cartoon textures that perfectly match the 'refined' prompt.
- + High clarity and clean composition on the diorama base.
- − The light blue background has a slight gradient instead of being a perfectly flat 'solid' color.
Seedream 4.5
- + Successfully creates the isometric miniature style on a diorama base.
- + Good material contrast between the rice and the fish.
- + Adheres to the color requirements and layout prompts.
- − Text is poorly aligned with 'SUSHI' being off-center relative to 'JAPAN'.
- − The diorama base texture is a bit noisy and looks like rough concrete rather than 'soft refined textures'.
- − The flag icon is placed awkwardly to the side of the text instead of below it as requested.
Verdict: GPT Image 1 is the clear winner as it perfectly executes the graphic design elements, particularly the centered typography and the clean cartoon aesthetic. While Seedream 4.5 captures the 3D miniature feel, its text rendering and layout are unbalanced and the textures feel a bit too gritty for a 'soft refined' scene.
Over-the-top cartoon caricature
Editing“Create a caricature of me and my job. Make it exaggerated and humorous, incorporating my profession as a tv show anchor and my love for dogs and hockey.”
AI Judge Analysis
GPT Image 1
- + Successfully captures a classic hand-drawn watercolor caricature style
- + Includes multiple clever story-telling elements like the hockey-playing dog on the monitor
- + Preserves the subject's outfit (denim shirt and black top) accurately
- − The facial features are extremely distorted in a way that loses the subject's likeness
- − The 'The News' text on the desk is slightly uneven
Seedream 4.5
- + Excellent preservation of the subject's facial likeness despite the caricature proportions
- + Maintains the blurred background from the source image to ground the edit
- + High resolution with very clean rendering of hockey equipment and the microphone
- − The caricature style is very subtle (mostly just a large head) compared to the more expressive art style of Model A
- − The subject's hands are quite small and slightly mutated
Verdict: GPT Image 1 (Model A) provides a much more traditional and humorous caricature art style with creative integration of the requested elements, though it sacrifices facial likeness. Seedream 4.5 (Model B) creates a high-quality 3D-style edit that perfectly preserves the subject's face and background, but feels more like a 'big head' filter than a full caricature. Model B is likely preferred for users wanting to still look like themselves, while Model A is better for those wanting a humorous, stylized piece of art.
Adorable Baby Animals in Sunny Meadow
Text-to-Image“Hyper-photorealistic scene of fluffy baby animals—a golden retriever puppy, tabby kitten, baby bunny, and red fox kit—with big expressive eyes and ultra-detailed soft fur, playfully chasing butterflies and tumbling together in a lush wildflower meadow, warm golden sunrise light with god rays and dew sparkles, joyful wholesome vibe, 8K masterpiece.”
AI Judge Analysis
GPT Image 1
- + Excellent anatomical consistency in the animals' paws and features
- + Superior lighting and god rays that feel integrated into the atmosphere
- + Natural-looking fur textures and soft color palette
- − The bunny is slightly smaller and less central to the action than others
Seedream 4.5
- + Vibrant colors with high contrast that pop
- + Includes both the requested dew sparkles and additional fantasy-like droplets
- + Strong character expressions
- − The fox's front right leg is missing a paw/joint, appearing like a stump
- − The 'dew/sparkles' look like floating bubbles or lens artifacts rather than moisture on plants
- − The bunny's anatomy is a bit stiff and upright compared to the motion of others
Verdict: GPT Image 1 is the superior image because it maintains high anatomical accuracy across all four animals while perfectly capturing the requested soft lighting and god rays. Seedream 4.5 has a noticeable anatomical error with the fox's leg and the 'sparkles' feel like artificial overlays rather than environmental dew.
Studio Ghibli Anime Style
Editing“Transform this photo into a Studio Ghibli–inspired illustration. Use soft pastel colors, hand-painted textures, gentle lighting, dreamy backgrounds, and a warm, nostalgic mood”
AI Judge Analysis
GPT Image 1
- + Excellent hand-painted texture that feels like a physical medium.
- + Captures a warm, nostalgic, and dreamy mood perfectly.
- + Highly creative interpretation of the scene in a soft-focus style.
- − The faces lose most of the specific expressions from the source image.
- − The plaid pattern on the shirt is significantly simplified/lost.
Seedream 4.5
- + Exceptional resemblance to specific Studio Ghibli character designs.
- + Perfectly preserves the composition and poses of the source image.
- + Accurately replicates the cell-shaded line art and watercolor backgrounds of the studio.
- − The transition from the foreground character to the background is a bit sharp compared to the source's bokeh.
Verdict: Both models did an excellent job, but Seedream 4.5 is the clear winner for its incredible mimicry of the actual Ghibli art style and character designs while perfectly preserving the 'Distracted Boyfriend' meme composition. GPT Image 1 captures the color and texture mood well but loses too much detail in the faces and clothing patterns.
Golden Hour Stroll
Image Editing“Add dynamic motion to this photo: make hair blow in the wind, add leaves flying, energetic and lively feel.”
AI Judge Analysis
GPT Image 1
- + Excellent preservation of the original subjects' faces and clothing
- + Subtle but effective hair motion that looks natural
- + Maintains the bridge and background detail perfectly
- − The flying leaves are very small and somewhat blurry
- − The leash has a strange loop artifact near the hand
Seedream 4.5
- + Stronger sense of motion with larger, more dynamic flying leaves
- + Highly energetic hair flow that clearly meets the prompt
- + Good composition with leaves in the foreground for depth
- − Significant changes to the woman's face and original features
- − Changed the background bridge and path layout substantially
- − The dog's proportions and tail changed noticeably from the source
Verdict: GPT Image 1 is much better at being a photo editor, as it preserves nearly all the original image's details while adding the requested effects. Seedream 4.5 creates a more dynamic and 'energetic' aesthetic but fails as an editor by essentially regenerating the entire scene, losing the likeness of the woman and the specific background details.
Vintage Cafe Logo
Text-to-Image“Vintage minimalist restaurant logo for "Caffè Florian", retro cloche dome with steam and "Est. 1720" banner, classic typography, warm brown and cream tones, subtle texture on light background, vector emblem style.”
AI Judge Analysis
GPT Image 1
- + Perfect text rendering of the name and establishment date
- + Clean minimalist vector style suitable for a modern brand
- − Failed to provide a light background as requested
- − Texture is very grain-heavy rather than subtle
Seedream 4.5
- + Excellent adherence to the color palette and background request
- + Better vintage aesthetic with classic engraving-style details
- + Dynamic banner design that balances the composition well
- − Minimal kerning issue on the main text
- − The cream tones are slightly inconsistent across the banner versus the cloche highlight
Verdict: Seedream 4.5 followed the prompt much more accurately by providing the light background and warm brown/cream tones requested. While GPT Image 1 has excellent typography, its failure to use a light background makes it less aligned with the desired vintage restaurant aesthetic than Seedream 4.5's classic logo design.
Apollo 11: Journey to Tranquility
Text-to-Image“Create a clean, modern vector infographic poster about the Apollo 11 mission. NASA-inspired palette (navy, white, muted red, light gray). Flat-vector style, crisp lines, consistent iconography, subtle gradients only. Steps (stop at landing): 1. Launch (Saturn Vicon) 2. Earth Orbit (Earth + orbit ring icon) 3. Translunar (trajectory arc icon) 4. Lunar Orbit (Moon + orbit ring icon) 5. Descent (lunar module descending icon) 6. Landing (lunar module on the surface icon) Small supporting elements (minimal text): • Crew strip: three silhouette icons with only last names: Armstrong, Aldrin, Collins. • Landing site marker: Moon pin labeled "Tranquility" only. Layout constraints: generous margins, large readable labels, clean background with subtle stars. Vector-only, print-poster look, high resolution.”
AI Judge Analysis
GPT Image 1
- + Excellent adherence to the NASA color palette
- + Clean flat-vector style across all icons
- + Accurate name labeling for the crew members
- − Poor layout where labels do not align with their respective icons
- − Significant spelling errors such as 'EARLLUNAR'
- − Distorted iconography for the Lunar Modules
Seedream 4.5
- + Perfectly organized timeline layout that clearly follows the requested steps
- + High-quality vector icons that match the technical subject matter
- + Flawless text rendering and alignment
- − Step 5 (Descent) features a generic satellite icon instead of a lunar module descending
Verdict: Seedream 4.5 is the clear winner as it successfully follows the complex infographic structure requested, providing a logical timeline with zero spelling errors. While GPT Image 1 uses the requested color palette well, its layout is chaotic and the text includes multiple typos and misalignments.
Explore each model
ByteDance's latest image generation model unifying text-to-image and image editing in a single architecture, with improved text rendering and 30-40% faster generation than v4.0