SandBase is live — $1 in free credits on signupStart free ›

KwaiVGI modelsvideo generation api

kwaivgi/kling-video/v1.6/standard/text-to-video

Kling Video V1.6 Standard is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Input
The aspect ratio of the generated image. Allowed values: 16:9, 9:16, 1:1.
The duration of the generated video in seconds Allowed values: 5, 10.
Idle

Example output — click Run to generate your own

API README

Kling Video V1.6 Standard

text to video in the Kling Video lineup is built to turn scene direction into a moving shot with deliberate temporal progression. The kling video v1.6 standard text to video configuration combines that transformation with the quality and motion profile represented by this exact model version, so teams can choose it deliberately among adjacent family variants. It is useful when the input already establishes part of the creative intent and the model must supply a polished result rather than a generic media conversion, with subject identity, scene logic, visual hierarchy, and the delivery goal kept explicit.

For production work with kwaivgi/kling-video/v1.6/standard/text-to-video, begin with the non-negotiable content, then describe the desired change, framing, action, atmosphere, and finishing cues in that order. Separate what must remain recognizable from what may vary, and prefer concrete nouns and observable actions over abstract praise. That structure makes outputs easier to compare across storyboard passes, campaign variants, catalog assets, and other repeatable creative pipelines.

Highlights

Prompt-to-motion synthesis. Builds the opening composition and temporal development from scene direction.

Coordinated subject action. Connects movement, environment response, and camera behavior across the shot.

Cinematic shot language. Understands framing, lens feel, camera travel, pacing, and atmosphere direction.

kling video v1.6 standard text to video temporal visual continuity. Reduces distracting changes in identity, geometry, texture, and lighting. This is the defining creative strength of the kling video v1.6 standard text to video configuration.

Pricing

DurationPrice per generated clip
5 seconds$0.1400
10 seconds$0.2800

Billing follows params.duration * 0.028: multiply generated seconds by the $0.0280 per-second rate.

When to Use

✅ Good fit❌ Consider alternatives
The scene should be created entirely from written directionA source image must anchor the opening frame
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "16:9",
  "duration": 5,
  "prompt": "A stylish woman walks down a Tokyo street filled with warm glowing neon and animated city signage. She wears a black leather jacket, a long red dress, and black boots, and carries a black purse."
}

Technical Specs

SpecValue
Model IDkwaivgi/kling-video/v1.6/standard/text-to-video
Inputsaspect_ratio, duration, prompt
Required inputsprompt
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Duration5 / 10
Aspect Ratio16:9 / 9:16 / 1:1

Related Models

Related Models

kwaivgi/kling-video/v1.6/standard/image-to-videoKling Video V1.6 Standard by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/v1.6/pro/text-to-videoKling Video V1.6 Pro by KwaiVGI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.kwaivgi/kling-video/v1.6/pro/image-to-videoKling Video V1.6 Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.kwaivgi/kling-video/1.6/pro/elementsKling Video 1.6 Pro by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/3.0/omni/pro/image-to-videoKling Video Omni Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.kwaivgi/kling-video/3.0/omni/pro/reference-to-videoGenerate a video from multiple reference images and text guidance with Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/editModify an existing video with natural-language instructions using Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/referenceUse a source video plus image references to generate a guided video variation with Kling 3.0 Omni Pro.