Wan 2.6 Text to Video
/v1/runAlibaba Wan 2.6 text-to-video model with cinematic visuals, configurable duration (5 or 10 seconds) and resolution.
Request body
Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.
Model identifier. Set to alibaba/wan/2.6/text-to-video.
Default: alibaba/wan/2.6/text-to-video
The text prompt to generate a video from.
Default: Rim light, low contrast, medium close-up, daylight, left-weighted composition, clean single-person shot, warm tones, soft light, sunny day, side light, daytime. A young girl sits in a field of tall grass, with two fluffy little donkeys standing behind her. The girl is about eleven or twelve, wearing a simple floral dress, hair in two braids, with an innocent smile. She sits cross-legged, gently playing with wildflowers beside her. The donkeys are sturdy with perked ears, curiously looking toward the camera. Sunlight bathes the field, creating a warm and natural atmosphere.
The duration of the generated video in seconds. Valid values are 5 and 10.
Allowed values: 5, 10
Default: 5
The URL of an audio file to guide video generation.
The aspect ratio of the generated video.
Allowed values: 1:1, 4:3, 3:4, 16:9, 9:16
Default: 16:9
The resolution of the video to generate.
Allowed values: 720p, 1080p
Default: 720p
Response Schema
The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.
Error message if the task failed. Empty on success.
Unique identifier for the generation task.
Model ID used for the prediction.
Array of generated content. Empty when status is not completed.
Status of the task: pending, running, completed, failed, or timeout.
Allowed values: pending, running, completed, failed, timeout
Model capabilities
Capabilities declared by the model registry.
Default: text-to-video
Execution mode declared by the model registry.
Default: async

