Skip to content

Vace

POST/v1/run

Wan Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

Request body

Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.

stringmodelrequired

Model identifier. Set to alibaba/wan-vace.

Default: alibaba/wan-vace

stringpromptrequired

The text prompt to guide video generation.

Optional<string>aspect_ratio

The aspect ratio of the generated image.

Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16

Optional<string>mask

URL to the guiding mask file. If provided, the model will use this mask as a reference to create masked video. If provided mask video url will be ignored.

Optional<string>video

URL to the source video file. If provided, the model will use this video as a reference.

Optional<string>resolution

Resolution of the generated video (480p,580p, or 720p).

Allowed values: 480p, 580p, 720p

Default: 720p

Optional<integer>num_inference_steps

Number of inference steps for sampling. Higher values give better quality but take longer.

Range: 2 to 40

Default: 30

Optional<integer>seed

Random seed for reproducibility. If None, a random seed is chosen.

Optional<boolean>preprocess

Whether to preprocess the input video.

Default: false

Optional<number>shift

Shift parameter for video generation.

Range: 1 to 10

Default: 5

Optional<string>mask_video_url

URL to the source mask file. If provided, the model will use this mask as a reference.

Optional<integer>num_frames

Number of frames to generate. Must be between 81 to 100 (inclusive). Works only with only reference images as input if source video or mask video is provided output len would be same as source up to 241 frames

Range: 81 to 240

Default: 81

Optional<string>task

Task type for the model.

Allowed values: depth, inpainting

Default: depth

Optional<array<string>>ref_image_urls

Urls to source reference image. If provided, the model will use this image as reference.

Optional<integer>frames_per_second

Frames per second of the generated video. Must be between 5 to 24.

Range: 5 to 24

Default: 16

Response Schema

The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.

Optional<string>error

Error message if the task failed. Empty on success.

stringidrequired

Unique identifier for the generation task.

Optional<string>model

Model ID used for the prediction.

Optional<array>outputs

Array of generated content. Empty when status is not completed.

stringstatusrequired

Status of the task: pending, running, completed, failed, or timeout.

Allowed values: pending, running, completed, failed, timeout

Model capabilities

array<string>capability_tagsrequired

Capabilities declared by the model registry.

Default: video-to-video

stringexecution_moderequired

Execution mode declared by the model registry.

Default: async