Skip to main content
POST
cURL

Authorizations

Authorization
string
header
required

Use your APIPod API key as a Bearer token in the Authorization header.

Body

application/json
model
string
required

Public APIPod model ID.

Allowed value: "seedance-2.5"
prompt
string
required

Generation or editing instructions, up to 30000 characters. It is recommended to keep the prompt to no more than 500 Chinese characters or 1,000 English words. Lengthy text will lead to scattered information, and the model may ignore details and only focus on key points, resulting in missing elements in the generated video.

Maximum string length: 30000
image_urls
string[]

Image URLs. Scenario inference applies: when images are the only media input and 1-2 URLs are provided they are treated as first/last frames (image-to-video); with other reference media present or more than 2 URLs, all images are treated as reference images. Pair with the role field to pin the semantics explicitly. Up to 30 images are supported; each image must be smaller than 30 MB.

Maximum array length: 30
role
enum<string>

Explicit interpretation of image_urls. reference_images: all images are treated as reference images regardless of count. first_last_frame: images are treated as the first frame plus an optional last frame (max 2; cannot be combined with reference videos, audios, or reference_images). When omitted, the scenario is inferred from the materials.

Available options:
reference_images,
first_last_frame
reference_images
string[]

Explicit reference image URLs. Always treated as reference images (omni reference generation) regardless of count; use this field when you need strict reference semantics with few images. Up to 30 images are supported; each image must be smaller than 30 MB.

Maximum array length: 30
video_urls
string[]

Reference video URLs. Up to 10 videos are supported; each video lasts 2-30 seconds and the total duration of all videos must not exceed 30 seconds.

Maximum array length: 10
audio_urls
string[]

Reference audio URLs. Up to 10 audios are supported; each audio lasts 2-30 seconds and the total duration of all audios must not exceed 30 seconds. Audio-only input is allowed.

Maximum array length: 10
duration
enum<integer>
default:4

Video duration in seconds, from 4 to 30.

Available options:
4,
5,
6,
7,
8,
9,
10,
11,
12,
13,
14,
15,
16,
17,
18,
19,
20,
21,
22,
23,
24,
25,
26,
27,
28,
29,
30
Required range: 4 <= x <= 30
resolution
enum<string>
default:720p

Generate video resolution.

Available options:
480p,
720p,
1080p
aspect_ratio
enum<string>
default:adaptive

Aspect ratio of the generated video. When the ratio is configured as adaptive, the model will automatically adapt the aspect ratio according to the generation scenario.

Available options:
adaptive,
16:9,
4:3,
1:1,
3:4,
9:16,
21:9
generate_audio
boolean
default:true

Controls whether the generated video contains audio synchronized with the visuals.

return_last_frame
boolean
default:false

Whether to return the last frame image of the generated video. Using this parameter enables the generation of multiple consecutive videos: taking the last frame of the previous generated video as the first frame of the next video task.

real_person
boolean
default:true

Whether the reference materials contain real people (human faces). Defaults to true: materials are submitted to the asset review service before task creation. Set to false only when the materials are confirmed face-free to skip review and pass material URLs through as-is.

Response

200 - application/json

Task accepted

Standard response for task submission

code
integer
required
data
Task Submit Data · object
required

Data content for task submission

message
string
required