Kling Video V3 Standard

kwaivgi/kling-video/v3/standard/text-to-video

Kling Video V3 Standard is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Base price
$0.63USD / run
Execution
async
Model type
video
Input fields
6

Try the model

Playground

Open playground
Input
Text prompt for video generation. Either prompt or multi_prompt must be provided, but not both.
The aspect ratio of the generated image. Allowed values: 16:9, 9:16, 1:1.
The duration of the generated video in seconds Allowed values: 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15.
Whether to generate native audio for the video. Supports Chinese and English voice output. Other languages are automatically translated to English. For English speech, use lowercase letters; for acronyms or proper nouns, use uppercase.
The type of multi-shot video generation. 'intelligent' lets the model automatically determine shot structure. Allowed values: customize, intelligent.
List of prompts for multi-shot video generation. If provided, overrides the single prompt and divides the video into multiple shots with specified prompts and durations.
OutputReady

Your output will appear here

Complete the inputs, then click Run.

Specifications

Pricing

Base price
$0.63 / run
Billing formula
params.duration * 0.126

Context & modalities

Input
Schema-defined
Output
video

Capabilities

Chat
Not supported
Vision
Not supported
Reasoning
Not supported
Structured output
Not supported
Function calling
Not supported
Audio input
Not supported

Access

Provider
KwaiVGI
Model ID
kwaivgi/kling-video/v3/standard/text-to-video
Execution
async
API
Unified Run API
Endpoint
/v1/run

API README

Kling Video V3 Standard Text-to-Video

Kling Video V3 Standard generates video directly from a written scene. The prompt describes the subject, action, camera, light, atmosphere, and sound without requiring an opening image. Landscape, portrait, and square frame shapes are available.

The route produces clips from 3 to 15 seconds and returns a video URL. Native audio is enabled by default, with documented language handling in the local contract. Custom and intelligent shot structure are exposed as request settings.

Highlights

  • Cinematic visual quality. The exact Standard text-to-video route is documented for cinematic visuals. This is the route's stated visual positioning.
  • Fluid motion. Fluid motion is explicitly identified for Standard text-to-video. Movement is generated from the written scene direction.
  • Improved visual quality. V3 Standard text-to-video documentation identifies improved visual quality. The statement is scoped to this exact version, tier, and route.
  • Improved motion consistency. Motion consistency is also identified as improved for V3 Standard text-to-video. It describes the generated motion rather than a schema control.

Pricing

DurationPrice
3 seconds$0.378
4 seconds$0.504
5 seconds$0.630
6 seconds$0.756
7 seconds$0.882
8 seconds$1.008
9 seconds$1.134
10 seconds$1.260
11 seconds$1.386
12 seconds$1.512
13 seconds$1.638
14 seconds$1.764
15 seconds$1.890

When to Use

✅ Good fit❌ Consider alternatives
Turn a written scene into a short video without an image input.Preserve a specific opening composition; use image-to-video.
Describe action and camera movement in one continuous prompt.Require a clip longer than 15 seconds in one request.
Deliver landscape, vertical, or square video.Require an aspect ratio outside the documented set.
Include generated ambience, effects, or speech.Require voice behavior outside the documented language handling.
Use the Standard tier for this text-led workflow.Choose Pro when its exact tier capabilities are required.

Prompt Guide

Write one continuous scene in temporal order. Start with the required prompt; add duration, aspect ratio, audio, or shot settings only when the request needs them.

Scene: [setting, time, atmosphere]
Subject: [appearance and action]
Camera: [framing and movement]
Lighting: [direction and change]
Motion: [subject and environment]
Audio: [ambience, effects, or speech when enabled]
{
  "prompt": "Cinematic drone shot through ancient stone ruins at golden hour. The camera rises through archways and reveals a misty valley.",
  "duration": 5,
  "aspect_ratio": "16:9",
  "generate_audio": true
}

Technical Specs

SpecValue
Required inputprompt
Prompt lengthUp to 2,500 characters
Duration3–15 seconds; default 5
Aspect ratios16:9, 9:16, 1:1
Shot typescustomize, intelligent
Multi-prompt fieldExposed locally, but prompt remains required by the request contract
Native audioOptional; default enabled
OutputVideo URL with optional file metadata

Related

Start building

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

Production API
Unified Run API endpoint
https://api.sandbase.ai/v1/run
Model ID
kwaivgi/kling-video/v3/standard/text-to-video
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
  -X POST "https://api.sandbase.ai/v1/run" \
  -H "Authorization: Bearer $SANDBASE_API_KEY" \
  -H "Content-Type: application/json" \
  --data-binary @- <<'SANDBASE_JSON'
{
  "model": "kwaivgi/kling-video/v3/standard/text-to-video",
  "prompt": "Cinematic drone shot flying through ancient stone ruins covered in moss and vines at golden hour. Camera starts low, rises through crumbling archways, revealing a vast misty valley beyond. Volumetric light rays pierce through gaps in the stone. Epic scale, photorealistic, 8K quality.",
  "duration": 5,
  "shot_type": "customize",
  "generate_audio": true
}
SANDBASE_JSON
)
run_id=$(printf '%s' "$result" | jq -r .id)
for attempt in $(seq 1 120); do
  status=$(printf '%s' "$result" | jq -r .status)
  case "$status" in completed|failed|timeout) break ;; esac
  sleep 2
  result=$(curl --fail-with-body --silent \
    -H "Authorization: Bearer $SANDBASE_API_KEY" \
    "https://api.sandbase.ai/v1/run/$run_id")
done
status=$(printf '%s' "$result" | jq -r .status)
[ "$status" = completed ] || { echo "Generation ended: $status" >&2; exit 1; }
printf '%s\n' "$result"

Choose your model

Compare models

All language models
ModelContextInput / 1MOutput / 1MReleased
KwaiVGI———Feb 4, 2026
KwaiVGI———Jul 6, 2026
KwaiVGI———Jul 6, 2026
KwaiVGI———Jul 3, 2026
KwaiVGI———Jul 3, 2026
KwaiVGI———Jul 3, 2026