doubao-seedream-5-0-pro-260628.
Key capabilities
- Text to image — Generate one high-precision image from a text prompt
- Image editing and fusion — Edit one reference or combine content and style from up to 10 references
- Interactive editing — Point to regions with natural language, markers, points, bounding boxes, arrows, or sketches
- Transparent background — Preserve an alpha channel in a single-image editing workflow
- Layer decomposition — Split one image into a background plus up to 16 editable transparent PNG layers
- Multilingual text — In addition to Chinese and English, supports Russian, Arabic, Filipino, Thai, Turkish, Korean, Malay, Spanish, Portuguese, Indonesian, French, German, Vietnamese, and Japanese prompts
- Text to image
- Image editing
- Transparent background
- Layer decomposition
Generate from text
size accepts 1K, 1.5K, 2K, or a valid explicit widthxheight value. optimize_prompt_options.mode accepts standard or fast; both modes have been verified through AnyFast.Output controls
Size and input limits
For normal generation, explicit dimensions must contain between 921,600 and 4,624,220 total pixels and use an aspect ratio in[1/16, 16]. Representative sizes include 1024x1024 for 1K, 1536x1536 for 1.5K, and 2048x2048 for 2K.
Normal reference images support jpeg, png, webp, bmp, tiff, gif, heic, and heif. Each image must be smaller than 30 MB, have an aspect ratio in [1/16, 16], have width and height greater than 14 px, and contain from 196 through 36,000,000 pixels.
Layer decomposition accepts one PNG or JPEG from 512x512 through 6000x6000 total-pixel bounds, with the same aspect-ratio and 30 MB limits.
The examples and response fields above are aligned with the current Volcengine Image Generation API and successful AnyFast tests for text generation, multi-image fusion, PNG/Base64 output,
standard and fast prompt optimization, transparent-background output, and layer decomposition.API Reference
View all Seedream 5.0 Pro workflows.