Vidu Q2 Text to Video
/v1/runVidu Q2 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
Request body
Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.
Model identifier. Set to vidu/q2/text-to-video.
Default: vidu/q2/text-to-video
Text prompt for video generation, max 3000 characters
The aspect ratio of the generated image.
Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16
Output video resolution
Allowed values: 360p, 520p, 720p, 1080p
Default: 720p
Duration of the video in seconds
Allowed values: 2, 3, 4, 5, 6, 7, 8
Default: 4
Random seed for reproducibility. If None, a random seed is chosen.
The movement amplitude of objects in the frame
Allowed values: auto, small, medium, large
Default: auto
Whether to add background music to the video (only for 4-second videos)
Default: false
Response Schema
The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.
Error message if the task failed. Empty on success.
Unique identifier for the generation task.
Model ID used for the prediction.
Array of generated content. Empty when status is not completed.
Status of the task: pending, running, completed, failed, or timeout.
Allowed values: pending, running, completed, failed, timeout
Model capabilities
Capabilities declared by the model registry.
Default: text-to-video
Execution mode declared by the model registry.
Default: async

