Veo 3.1 Fast
google/veo3.1/fast/reference-to-videoVeo3.1 Fast by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
- Base price
- $1.20USD / run
- Execution
- async
- Model type
- video
- Input fields
- 6
Try the model
Playground
PNG, JPEG, WebP, or GIF · 20 MiB maximum each
Your output will appear here
Complete the inputs, then click Run.
Specifications
Pricing
- Base price
- $1.20 / run
- Billing formula
- params.resolution == "4k" ? (params.generate_audio ? params.duration * 0.35 : params.duration * 0.30) : (params.generate_audio ? params.duration * 0.15 : params.duration * 0.10)
Context & modalities
- Input
- Schema-defined
- Output
- video
Capabilities
- Chat
- Not supported
- Vision
- Not supported
- Reasoning
- Not supported
- Structured output
- Not supported
- Function calling
- Not supported
- Audio input
- Not supported
Access
- Provider
- Model ID
- google/veo3.1/fast/reference-to-video
- Execution
- async
- API
- Unified Run API
- Endpoint
- /v1/run
API README
Veo 3.1 Fast
google/veo3.1/fast/reference-to-video uses multiple visual references to establish characters, objects, environments, or style in a newly generated video. Veo 3.1 improves instruction following, physical consistency, character continuity, and audiovisual quality while supporting more ways to anchor or continue a shot. This combination makes the model a practical choice when the creative outcome depends on those qualities rather than on a generic media conversion.
For production work with google/veo3.1/fast/reference-to-video, this Fast reference-guided generation workflow prioritizes turnaround for concept comparison, previsualization, social creative, and high-volume variation before a final direction is selected. The result is most reliable when the source material and creative brief clearly describe the intended subject, progression, visual or sonic character, and the qualities that must remain unchanged.
Highlights
Multi-reference understanding carries supplied identities, objects, and style into a coherent moving scene.
Improved physical coherence strengthens object interactions, character movement, and cause-and-effect across shots.
Fast creative iteration produces cinematic alternatives more quickly for review and selection.
Native audio generation coordinates dialogue, ambience, and sound effects with the visual action.
Pricing
| Resolution | Audio off, per second | Audio on, per second |
|---|---|---|
| 720p or 1080p | $0.100 | $0.150 |
| 4K | $0.300 | $0.350 |
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| The model's named workflow matches the source material and intended output | A different input modality or model route is required |
| A managed asynchronous result is suitable for the production pipeline | A synchronous, interactive editor is essential |
| The documented controls cover the required duration, framing, or format | The project needs controls outside this endpoint's schema |
| Creative iteration benefits from a repeatable request structure | Exact deterministic pixels, frames, geometry, or samples are mandatory |
| A finished downloadable media asset is the desired deliverable | Editable source layers or a native project file are required |
Prompt Guide
For generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.
{
"aspect_ratio": "21:9",
"duration": 8,
"images": [
"https://static.sandbase.ai/examples/google/veo3.1/fast/reference-to-video/input_images_0.png"
],
"prompt": "A chimpanzee wearing overalls frolics in the grassy field, gently playing with the butterflies. In the background, a circus tent and carousel beckon.",
"resolution": "720p"
}
Technical Specs
| Spec | Value |
|---|---|
| Model ID | google/veo3.1/fast/reference-to-video |
| Inputs | aspect_ratio, duration, generate_audio, images, prompt, resolution |
| Required inputs | prompt |
| Output fields | content_type, url |
| Execution | Async (submit, then poll for result) |
| Resolution | 720p / 1080p / 4k |
| Aspect Ratio | 21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16 |
Related Models
Start building
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
-X POST "https://api.sandbase.ai/v1/run" \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
-H "Content-Type: application/json" \
--data-binary @- <<'SANDBASE_JSON'
{
"model": "google/veo3.1/fast/reference-to-video",
"images": [
"https://static.sandbase.ai/examples/google/veo3.1/fast/reference-to-video/input_images_0.png",
"https://static.sandbase.ai/examples/google/veo3.1/fast/reference-to-video/input_images_1.png",
"https://static.sandbase.ai/examples/google/veo3.1/fast/reference-to-video/input_images_2.png"
],
"prompt": "A chimpanzee wearing overalls frolics in the grassy field, gently playing with the butterflies. In the background, a circus tent and carousel beckon.",
"duration": 8,
"resolution": "720p",
"generate_audio": true
}
SANDBASE_JSON
)
run_id=$(printf '%s' "$result" | jq -r .id)
for attempt in $(seq 1 120); do
status=$(printf '%s' "$result" | jq -r .status)
case "$status" in completed|failed|timeout) break ;; esac
sleep 2
result=$(curl --fail-with-body --silent \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
"https://api.sandbase.ai/v1/run/$run_id")
done
status=$(printf '%s' "$result" | jq -r .status)
[ "$status" = completed ] || { echo "Generation ended: $status" >&2; exit 1; }
printf '%s\n' "$result"Choose your model
Compare models
| Model | Context | Input / 1M | Output / 1M | Released |
|---|---|---|---|---|
Veo 3.1 FastThis model Google | — | — | — | May 11, 2026 |
Google | — | — | — | Jun 30, 2026 |
Google | — | — | — | Jun 30, 2026 |
Google | — | — | — | Jun 30, 2026 |
Google | — | — | — | Jun 30, 2026 |
Google | — | — | — | Mar 31, 2026 |
