Kling Video O3 4k

kwaivgi/kling-video/o3/4k/text-to-video

Kling Video O3 4k is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Base price
$2.10USD / run
Execution
async
Model type
video
Input fields
6

Try the model

Playground

Open playground
Input
Text prompt for video generation. Required unless multi_prompt is provided.
The aspect ratio of the generated image. Allowed values: 16:9, 9:16, 1:1.
Video duration in seconds (3-15s). Allowed values: 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15.
Whether to generate native audio for the video.
List of prompts for multi-shot video generation.
The type of multi-shot video generation. 'intelligent' lets the model automatically determine shot structure. Allowed values: customize, intelligent.
OutputReady

Your output will appear here

Complete the inputs, then click Run.

Specifications

Pricing

Base price
$2.10 / run
Billing formula
params.duration * 0.42

Context & modalities

Input
Schema-defined
Output
video

Capabilities

Chat
Not supported
Vision
Not supported
Reasoning
Not supported
Structured output
Not supported
Function calling
Not supported
Audio input
Not supported

Access

Provider
KwaiVGI
Model ID
kwaivgi/kling-video/o3/4k/text-to-video
Execution
async
API
Unified Run API
Endpoint
/v1/run

API README

Kling Video O3 4k

Kling Video O3 4k is a text-led video-generation route in the Kling Video O3 family. It turns written scene direction into a multi-second sequence with controlled subject action, camera movement, shot structure, and optional synchronized sound.

The 4K route is intended for creators who need dependable temporal continuity across a directed clip. Prompts can describe staging, motion, lens behavior, atmosphere, and audio cues; multi-prompt controls can divide the sequence into several shots while reference inputs keep important visual elements anchored.

Highlights

  • Prompt-led scene creation. Builds the scene, subjects, action, and camera language from a written production brief.
  • Multi-shot direction. Supports structured prompt segments for sequences that need more than one planned shot.
  • Native audio option. Can generate sound together with the visible action when the route exposes audio generation.
  • High-resolution delivery. Targets a 4K delivery route for work that needs additional spatial detail.

Pricing

ConfigurationBilling unitPrice
3 seconds4K route$1.260
4 seconds4K route$1.680
5 seconds4K route$2.100
6 seconds4K route$2.520
7 seconds4K route$2.940
8 seconds4K route$3.360
9 seconds4K route$3.780
10 seconds4K route$4.200
11 seconds4K route$4.620
12 seconds4K route$5.040
13 seconds4K route$5.460
14 seconds4K route$5.880
15 seconds4K route$6.300

When to Use

✅ Good fit❌ Consider alternatives
The project needs creating a directed video sequenceThe goal is a different media task or endpoint
The available inputs match the required local schemaRequired source media or permissions are unavailable
The brief can specify subject, composition, style, and deliveryThe result must be deterministic at pixel or sample level
The supported formats and controls match final placementDelivery requires unsupported dimensions, codecs, or duration
An asynchronous generated result fits the workflowA live, frame-synchronous, or real-time response is mandatory

Prompt Guide

Write the request as a production brief: identify the main subject or source material, state the intended transformation, describe composition or timing, and finish with style, atmosphere, and delivery constraints. Keep preservation requirements separate from requested changes, and use only fields exposed by this route.

{
  "prompt": "A cinematic, precisely directed scene with clear subject action, camera movement, lighting, and atmosphere",
  "duration": 3,
  "shot_type": "customize",
  "aspect_ratio": "16:9"
}

Technical Specs

SpecValue
Model IDkwaivgi/kling-video/o3/4k/text-to-video
Input fieldsprompt (string)<br>duration (integer; 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15)<br>shot_type (string; customize, intelligent)<br>aspect_ratio (string; 16:9, 9:16, 1:1)<br>multi_prompt (array)<br>generate_audio (boolean)
Required inputprompt
Output fieldsurl, content_type
ExecutionAsynchronous job

Related Models

  • kwaivgi/kling-video/o3/pro/image-to-video
  • kwaivgi/kling-video/o3/standard/reference-to-video
  • kwaivgi/kling-video/o3/standard/text-to-video

Start building

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

Production API
Unified Run API endpoint
https://api.sandbase.ai/v1/run
Model ID
kwaivgi/kling-video/o3/4k/text-to-video
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
  -X POST "https://api.sandbase.ai/v1/run" \
  -H "Authorization: Bearer $SANDBASE_API_KEY" \
  -H "Content-Type: application/json" \
  --data-binary @- <<'SANDBASE_JSON'
{
  "model": "kwaivgi/kling-video/o3/4k/text-to-video",
  "prompt": "A mecha lands on the ground to save the city, and says \"I'm here\", in anime style",
  "duration": 5,
  "shot_type": "customize",
  "generate_audio": false
}
SANDBASE_JSON
)
run_id=$(printf '%s' "$result" | jq -r .id)
for attempt in $(seq 1 120); do
  status=$(printf '%s' "$result" | jq -r .status)
  case "$status" in completed|failed|timeout) break ;; esac
  sleep 2
  result=$(curl --fail-with-body --silent \
    -H "Authorization: Bearer $SANDBASE_API_KEY" \
    "https://api.sandbase.ai/v1/run/$run_id")
done
status=$(printf '%s' "$result" | jq -r .status)
[ "$status" = completed ] || { echo "Generation ended: $status" >&2; exit 1; }
printf '%s\n' "$result"

Choose your model

Compare models

All language models
ModelContextInput / 1MOutput / 1MReleased
KwaiVGI———Apr 23, 2026
KwaiVGI———Jul 6, 2026
KwaiVGI———Jul 6, 2026
KwaiVGI———Jul 3, 2026
KwaiVGI———Jul 3, 2026
KwaiVGI———Jul 3, 2026