Seedream 4.0
NewReleased Sep 9, 2025
₹2.88
per image, from
Billed at provider rates converted at ₹96 per US dollar, with 0% markup.
Price per image
Output options
- Resolutions
- 1024x1024, 1280x720, 720x1280
- Editing
- Yes
- Reference images
- up to 3
About Seedream 4.0
ByteDance's Seedream 4.0, a strong general-purpose image model with particular strength in photographic realism and Chinese-language text rendering.
Seedream 4.0 is the release where ByteDance's Seed team folded its separate generation and editing models into one. Seedream 3.0 made images and SeedEdit 3.0 changed them; 4.0 does both from a single architecture, and the team says training the two jobs together produced better results than training either alone on instruction following, image quality and aesthetic appeal. That is the structural fact everything else in the post follows from: the model takes text, an image, or any mixture, and the same weights handle text-to-image, image-to-image, single-image editing, multi-image editing and composition.
Text rendering is where ByteDance claims the clearest break from earlier generations. The team says 4.0 renders text correctly and legibly while laying out genuinely difficult content - formulas, tables, chemical structures and statistical charts - which is what makes it usable for teaching material, academic figures and infographics rather than only for pictures. It also supports editing the text afterwards and swapping fonts, so a rendered layout is a starting point rather than a finished frame.
Two capabilities are unusual enough to be worth naming. Visual-signal control is native: edge, depth and mask conditioning work inside the model, without the external control networks that pipelines normally bolt on, and a rough sketch or a set of guide lines is enough to steer a composition, which matters for pose control, architectural work and interface prototypes. And the model reasons before it draws - ByteDance shows it solving puzzles, completing crosswords and continuing comic strips, tasks that need it to hold a physical or temporal constraint rather than only to match a description.
On its own evaluations, ByteDance reports 4.0 leading on every dimension of its MagicBench human-assessment benchmark for both text-to-image generation and image editing, and taking the highest Elo rating for single-image editing on MagicArena. The team is specific about where the edge sits against other models: image texture, lighting and colour. Elsewhere in the post it notes that the generation resolution ceiling rose from 2K to 4K in this generation and that the underlying transformer runs more than ten times faster than Seedream 3.0's, through distillation, quantisation and cache reuse rather than a smaller model.
ByteDance is measured about what it built. It calls 4.0 an early-stage prototype of a general-purpose multimodal creative engine rather than a finished one, and the comparative claims in the post rest on the team's own human-scored benchmarks. Treat the reasoning demonstrations as an indication of a new capability rather than a guarantee it holds on your prompt.
What ByteDance announced at launch
- Generation and editing, one model
- Seedream 3.0's text-to-image and SeedEdit's editing were merged into a single architecture, and ByteDance says joint training beat training either task alone on instruction following and image quality.
- Formulas, tables and charts
- The team says 4.0 renders difficult content legibly - equations, tables, chemical structures, statistical charts - and supports editing that text and replacing fonts afterwards.
- Control without extra models
- Edge, depth and mask conditioning are native to the model rather than bolted on with an external control network, and a rough sketch or guide lines can steer the composition.
- Reasoning before drawing
- ByteDance shows the model solving puzzles, completing crosswords and continuing comic strips, tasks that need it to hold a physical or temporal constraint rather than match a description.
- Composition from many references
- ByteDance says the model takes up to a dozen reference images at a time, pulling character features, scene styles and object structures from them while keeping scale and physical structure coherent.
- What it is not for
- ByteDance calls this an early-stage prototype of a general multimodal creative engine, and its comparative claims come from the team's own human-scored benchmarks rather than an outside harness.
Frequently Asked Questions
Frequently asked questions about Seedream 4.0.
When was Seedream 4.0 released?
ByteDance released Seedream 4.0 on Sep 9, 2025.
Who built Seedream 4.0?
Seedream 4.0 is developed by ByteDance. 99Models AI connects directly to it at the provider's published rate.
How much does an image from Seedream 4.0 cost?
Images start from ₹2.88 each, billed at the provider rate with 0% markup. You pay only for the images you generate.
What sizes and aspect ratios does Seedream 4.0 support?
Available options are 1024x1024, 1280x720, 720x1280, which you can select before generating an image.
Can Seedream 4.0 edit an existing image?
Yes. This specifies whether the model supports modifying an uploaded image in addition to creating images from text prompts.
More models from ByteDance
Catalog updated Sep 9, 2026