SandBase is live — $1 in free credits on signupStart free ›

KwaiVGI modelsvideo generation api

kwaivgi/kling-video/v2/master/text-to-video

Kling Video V2 Master is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Input
The aspect ratio of the generated image. Allowed values: 16:9, 9:16, 1:1.
The duration of the generated video in seconds Allowed values: 5, 10.
Idle

Example output — click Run to generate your own

API README

Kling Video V2 Master

Kling Video V2 Master is a text-to-video creation endpoint in the Kling family. It is built for creators who need to turn a concrete creative brief into a controlled visual sequence: the request establishes the source material, the intended subject behavior, the camera language, and the atmosphere of the finished shot. The standard route keeps that workflow explicit instead of hiding its input assumptions behind a generic video-generation label.

In practice, this route accepts a written prompt as creative context and exposes duration, aspect ratio for delivery planning. That makes it suitable for shot-based pipelines where teams must preserve a source, direct a transformation, or control the final format without losing sight of the model's central task. Write prompts as a compact shot plan—subject, action, setting, camera, light, and timing—then use the structured fields for constraints that should remain deterministic across iterations.

Highlights

  • Prompt-to-scene generation. Builds the shot from written direction, including subject action, environment, lighting, lens language, and pacing.
  • Controllable shot length. Duration or frame-count controls make timing an explicit part of the generation brief.
  • Format-aware framing. Selectable canvases cover cinematic, landscape, square, and portrait delivery formats.
  • Temporal coherence. The Kling generation path is designed around continuous motion across a shot, not a collection of unrelated frames.

Pricing

The request price is calculated with params.duration * 0.28.

ConfigurationPrice
duration=5$1.400000
duration=10$2.800000

The model card records a base price of $1.400000; the formula above determines usage-priced requests.

When to Use

ScenarioRecommendation
Choose this routeUse it when the deliverable specifically calls for text-to-video creation, rather than a neighboring generation mode.
Prepare the sourceProvide a precise written brief in the format described by the request schema.
Direct the shotDescribe the subject, action, environment, camera movement, lighting, and temporal progression in that order.
Control continuityUse endpoint frames, reference media, strength, or audio controls when those fields are available instead of burying hard constraints in prose.
Plan deliverySet duration, frame count, resolution, and aspect ratio explicitly when the schema exposes them, then compare iterations with a stable seed where supported.

Prompt Guide

For text-to-video creation, describe one coherent shot rather than a list of visual keywords. Put the main subject and action first, follow with location and staging, then add camera movement, lens or framing, lighting, mood, and any timed change. Keep URLs and hard delivery choices in their dedicated fields.

{
  "prompt": "A slow-motion drone shot descending from above a maze of neon-lit Tokyo alleyways at night during heavy rainfall. The camera gradually focuses on a lone figure in a luminescent white raincoat standing perfectly still amid the bustling crowd, all carrying black umbrellas. As the camera continues its downward journey, we see the raindrops creating rippling patterns on puddles that reflect the kaleidoscope of colors from the surrounding signs, creating a mirror world beneath the city.",
  "duration": 5,
  "aspect_ratio": "16:9"
}

Technical Specs

SpecificationValue
Model IDkwaivgi/kling-video/v2/master/text-to-video
Required inputsprompt
ExecutionAsynchronous generation job
Request controls3 documented fields
Outputurl, content_type

Request fields

FieldType and constraints
promptstring; Required
durationinteger; Optional; Options: 5, 10; Default: 5
aspect_ratiostring; Optional; Options: 16:9, 9:16, 1:1

Related Models

Related Models

kwaivgi/kling-video/v2/master/image-to-videoKling Video V2 Master by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/1.6/elementsKling Video 1.6 Elements by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/1.6/pro/effectsKling Video 1.6 Pro is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.kwaivgi/kling-video/1.6/pro/elementsKling Video 1.6 Pro by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/3.0/omni/pro/image-to-videoKling Video Omni Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.kwaivgi/kling-video/3.0/omni/pro/reference-to-videoGenerate a video from multiple reference images and text guidance with Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/editModify an existing video with natural-language instructions using Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/referenceUse a source video plus image references to generate a guided video variation with Kling 3.0 Omni Pro.