DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

ByteDance logo

ByteDance/

Seedream-5.0-Pro

Partner

$0.0495 - $0.099 / image

*

ByteDance's flagship image model, with precise interactive editing, stronger multi-reference fusion and notably accurate text rendering at up to 2K.

Public
ByteDance/Seedream-5.0-Pro cover image
api

Input

Prompt

Text prompt for image generation. Describe the aspect ratio or purpose in the prompt and the model picks the final dimensions within the `size` tier.

Please upload an image file

You need to log in to use this model

Log In

Settings

Size

Resolution tier ('1K', '1.5K' or '2K'), or explicit 'widthxheight' pixels. '1.5K' costs the same as '1K' and generates better images.. (Default: 2K)

Please upload an image file

Please upload an image file

Please upload an image file

Output Format

File format of the generated image

Background

'transparent' generates an image with an alpha channel. Only supported for image-to-image with a single input image that itself has an alpha channel, and not with 'jpeg' output

Optimize Prompt Mode

Prompt optimization mode. 'standard' gives higher quality, 'fast' lower latency

Output

generated image #0
Model Information

Seedream 5.0 Pro is ByteDance's flagship image generation model and the successor to Seedream 4.5. It advances instruction following, multi-reference composition and text rendering, and adds interactive editing that can be steered by coordinates and sketch markers placed on the input image.

Accurate text rendering

Renders headlines and body copy cleanly and across languages, which makes it practical for posters, packaging and other production-ready visual assets rather than illustration alone.

Multi-reference fusion

Combines subjects, styles and backgrounds drawn from several reference images while preserving each subject's identity and detail.

Precise editing

Edits the region you name instead of re-rolling the whole frame, holding lighting, colour tone and fine detail steady everywhere else.

Resolution

Generates at the 1K, 1.5K and 2K tiers, or at explicit pixel dimensions. 1.5K is priced the same as 1K and produces noticeably better images, so it is usually the better default of the two.