Create Video
Submit a Seedance 2.0 multimodal video generation task using `POST /v1/video/generations`.
https://zx1.deepwl.net/v1/video/generations{
"model": "doubao-seedance-2-0-260128",
"content": [
{
"type": "text",
"text": "First-person view fruit tea ad: 0-2s picking apples by hand; 2-4s cut to pouring into a shaker cup and shaking; 4-6s close-up of pouring into a clear cup; 6-8s raise the cup toward the camera"
},
{
"type": "image_url",
"image_url": { "url": "https://example.com/apple.jpg" },
"role": "reference_image"
},
{
"type": "image_url",
"image_url": { "url": "https://example.com/cup.jpg" },
"role": "reference_image"
},
{
"type": "video_url",
"video_url": { "url": "https://example.com/pov_reference.mp4" },
"role": "reference_video"
},
{
"type": "audio_url",
"audio_url": { "url": "https://example.com/bgm.mp3" },
"role": "reference_audio"
}
],
"metadata": {
"duration": 8,
"resolution": "720p",
"ratio": "16:9",
"generate_audio": true,
"watermark": false
}
}{
"task_id": "cgt-20260412163502-x8k2m"
}Submit a Seedance 2.0 video generation task. It supports text-to-video, first-frame/first-and-last-frame, reference image/video/audio, video continuation, video editing, and multimodal composition modes. For assets in the media library, it is recommended to reference them in content using asset://{assetId} (see Upload Assets).
Request Headers
Authorizationstring必填Authentication header. Uses a Bearer token, e.g. Bearer YOUR_API_KEY.
Content-Typestring必填Request content type, must be application/json.
Request Body
modelstring必填Model name:
doubao-seedance-2-0-260128: Standard version, optimized for the best visual quality and complex shot planningdoubao-seedance-2-0-fast-260128: Fast version, optimized for low latency and cost-sensitive scenarios
contentarray<object>必填Multimodal input array; the order affects role assignment.
typestring必填Content type: text, image_url, video_url, audio_url, draft_task.
textstringRequired when type=text; prompt text.
image_urlobjectUsed when type=image_url; must include url.
urlstring必填Public image URL or asset reference asset://{assetId}.
video_urlobjectUsed when type=video_url; must include url.
urlstring必填Public video URL or asset://{assetId}.
audio_urlobjectUsed when type=audio_url; must include url.
urlstring必填Public audio URL or asset://{assetId}.
draft_taskobjectUsed when type=draft_task; must include id, and it must be the only element in content.
idstring必填Draft task ID, used to continue generation from a draft.
rolestringMedia role:
first_frame: first frame (image)last_frame: last frame (image)reference_image: reference imagereference_video: reference/source video (continuation, editing)reference_audio: reference audio (requiresmetadata.generate_audio=true)
metadataobjectVideo generation parameters; all are optional.
durationintegerVideo duration in seconds. Valid range [4, 15] or -1 (automatically determined by the model), default 5.
resolutionstringResolution: 480p, 720p, 1080p, 4k, default 720p.
ratiostringAspect ratio: 16:9, 9:16, 1:1, 4:3, adaptive, default 16:9.
framesintegerTotal number of video frames. Mutually exclusive with duration; if frames is provided, it takes precedence over duration.
seedintegerRandom seed. The same seed plus the same input can produce similar results.
camera_fixedbooleanWhether to keep the camera fixed (suppress camera movement), default false.
watermarkbooleanWhether to add a watermark in the bottom-right corner of the video, default true.
generate_audiobooleanWhether to generate or synthesize audio. Must be true when using reference_audio, default false.
return_last_framebooleanWhether to return the final frame image URL for subsequent continuation, default false.
draftbooleanDraft mode: faster generation with slightly lower quality, suitable for previews, default false.
service_tierstringService tier, default default.
execution_expires_afterintegerMaximum task execution time in seconds, range [3600, 259200] (1 hour to 3 days), default 172800.
callback_urlstringCallback URL when the task is completed.
Request Examples
Response Examples
After submission, use Query Task to poll task_id.
Response Fields
task_idstringTask ID, used for Query Task.
content Mixing Rules
Violating the following rules may return 400:
reference_imagecannot appear together withfirst_frame/last_frameaudio_urlcannot be the only input incontent; it must be paired with at least an image or videodraft_taskmust be the only element in thecontentarray
Generation Mode Comparison
| Mode | Request Example Label | content Key Points |
|---|---|---|
| Text to Video | Text to Video | text + optional reference image/video/audio |
| First-Frame Image to Video | First-Frame Image to Video | text + first_frame |
| First-and-Last-Frame Image to Video | First-and-Last-Frame Image to Video | text + first_frame + last_frame |
| Reference Image to Video | Reference Image to Video | text + reference_image |
| Video Continuation | Video Continuation | text + reference_video |
| Video Editing | Video Editing | text + reference_video + reference_image |
| Multimodal Composition | Multimodal Composition | text + multiple types of references |
| Reference Assets | Reference Assets | Each URL uses asset://{assetId} |