PixVerse V6 Text to Video
/v1/runV6 by PixVerse - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
Request body
Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.
Model identifier. Set to pixverse/v6/text-to-video.
Default: pixverse/v6/text-to-video
Text prompt that describes the image or asset to generate.
The aspect ratio of the generated image.
Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16
The resolution of the generated video
Allowed values: 360p, 540p, 720p, 1080p
Default: 720p
The duration of the generated video in seconds. v6 supports values from 1 to 15 seconds
Range: 1 to 15
Default: 5
The same seed and the same prompt given to the same version of the model will output the same video every time.
Enable audio generation (BGM, SFX, dialogue)
Default: false
The style of the generated video
Allowed values: anime, 3d_animation, clay, comic, cyberpunk
Response Schema
The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.
Error message if the task failed. Empty on success.
Unique identifier for the generation task.
Model ID used for the prediction.
Array of generated content. Empty when status is not completed.
Status of the task: pending, running, completed, failed, or timeout.
Allowed values: pending, running, completed, failed, timeout
Model capabilities
Capabilities declared by the model registry.
Default: text-to-video
Execution mode declared by the model registry.
Default: async

