99MODELS

Seedream 4.0

New

Released Sep 9, 2025

2.88

per image, from

Billed at provider rates converted at ₹96 per US dollar, with 0% markup.

Price per image

Price per image
OptionINRUSD
1024x10242.88$0.03
1280x7202.88$0.03
720x12802.88$0.03

Output options

Resolutions
1024x1024, 1280x720, 720x1280
Editing
Yes
Reference images
up to 3

About Seedream 4.0

ByteDance's Seedream 4.0, a strong general-purpose image model with particular strength in photographic realism and Chinese-language text rendering.

Seedream 4.0 is the release where ByteDance's Seed team folded its separate generation and editing models into one. Seedream 3.0 made images and SeedEdit 3.0 changed them; 4.0 does both from a single architecture, and the team says training the two jobs together produced better results than training either alone on instruction following, image quality and aesthetic appeal. That is the structural fact everything else in the post follows from: the model takes text, an image, or any mixture, and the same weights handle text-to-image, image-to-image, single-image editing, multi-image editing and composition.

Text rendering is where ByteDance claims the clearest break from earlier generations. The team says 4.0 renders text correctly and legibly while laying out genuinely difficult content - formulas, tables, chemical structures and statistical charts - which is what makes it usable for teaching material, academic figures and infographics rather than only for pictures. It also supports editing the text afterwards and swapping fonts, so a rendered layout is a starting point rather than a finished frame.

Two capabilities are unusual enough to be worth naming. Visual-signal control is native: edge, depth and mask conditioning work inside the model, without the external control networks that pipelines normally bolt on, and a rough sketch or a set of guide lines is enough to steer a composition, which matters for pose control, architectural work and interface prototypes. And the model reasons before it draws - ByteDance shows it solving puzzles, completing crosswords and continuing comic strips, tasks that need it to hold a physical or temporal constraint rather than only to match a description.

On its own evaluations, ByteDance reports 4.0 leading on every dimension of its MagicBench human-assessment benchmark for both text-to-image generation and image editing, and taking the highest Elo rating for single-image editing on MagicArena. The team is specific about where the edge sits against other models: image texture, lighting and colour. Elsewhere in the post it notes that the generation resolution ceiling rose from 2K to 4K in this generation and that the underlying transformer runs more than ten times faster than Seedream 3.0's, through distillation, quantisation and cache reuse rather than a smaller model.

ByteDance is measured about what it built. It calls 4.0 an early-stage prototype of a general-purpose multimodal creative engine rather than a finished one, and the comparative claims in the post rest on the team's own human-scored benchmarks. Treat the reasoning demonstrations as an indication of a new capability rather than a guarantee it holds on your prompt.

What ByteDance announced at launch

Generation and editing, one model
Seedream 3.0's text-to-image and SeedEdit's editing were merged into a single architecture, and ByteDance says joint training beat training either task alone on instruction following and image quality.
Formulas, tables and charts
The team says 4.0 renders difficult content legibly - equations, tables, chemical structures, statistical charts - and supports editing that text and replacing fonts afterwards.
Control without extra models
Edge, depth and mask conditioning are native to the model rather than bolted on with an external control network, and a rough sketch or guide lines can steer the composition.
Reasoning before drawing
ByteDance shows the model solving puzzles, completing crosswords and continuing comic strips, tasks that need it to hold a physical or temporal constraint rather than match a description.
Composition from many references
ByteDance says the model takes up to a dozen reference images at a time, pulling character features, scene styles and object structures from them while keeping scale and physical structure coherent.
What it is not for
ByteDance calls this an early-stage prototype of a general multimodal creative engine, and its comparative claims come from the team's own human-scored benchmarks rather than an outside harness.

Frequently Asked Questions

Frequently asked questions about Seedream 4.0.

When was Seedream 4.0 released?

ByteDance released Seedream 4.0 on Sep 9, 2025.

Who built Seedream 4.0?

Seedream 4.0 is developed by ByteDance. 99Models AI connects directly to it at the provider's published rate.

How much does an image from Seedream 4.0 cost?

Images start from ₹2.88 each, billed at the provider rate with 0% markup. You pay only for the images you generate.

What sizes and aspect ratios does Seedream 4.0 support?

Available options are 1024x1024, 1280x720, 720x1280, which you can select before generating an image.

Can Seedream 4.0 edit an existing image?

Yes. This specifies whether the model supports modifying an uploaded image in addition to creating images from text prompts.

More models from ByteDance

Catalog updated Sep 9, 2026