SandBase is live — $1 in free credits on signupStart free ›

KwaiVGI modelsvideo generation api

kwaivgi/kling-video/v3/4k/text-to-video

Kling Video V3 4k is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Input
Text prompt for video generation. Either prompt or multi_prompt must be provided, but not both.
The aspect ratio of the generated image. Allowed values: 16:9, 9:16, 1:1.
The duration of the generated video in seconds Allowed values: 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15.
Whether to generate native audio for the video. Supports Chinese and English voice output. Other languages are automatically translated to English. For English speech, use lowercase letters; for acronyms or proper nouns, use uppercase.
List of prompts for multi-shot video generation. If provided, overrides the single prompt and divides the video into multiple shots with specified prompts and durations.
The type of multi-shot video generation. 'intelligent' lets the model automatically determine shot structure. Allowed values: customize, intelligent.
Idle

Example output — click Run to generate your own

API README

Kling Video V3 4K Text-to-Video

Kling Video V3 4K Text-to-Video converts a written creative brief into a rendered video clip. The endpoint is the text-first option in the Kling Video V3 4K model family, so no image upload is needed to begin generation.

The workflow keeps narrative direction and video generation in one request. Teams can move from a script-like description to a downloadable result while choosing the framing and run time appropriate to the intended sequence.

Highlights

Native professional-grade 4K. Produces professional-quality 4K video directly, preserving the resolution expected for premium campaign, cinematic, and large-screen delivery.

One-step 4K generation. Creates the high-resolution result in a single generation step, so the workflow does not depend on a separate upscale pass.

Cinema-grade clarity. Maintains cinematic visual clarity in the generated 4K frame, helping fine scene detail remain legible at final delivery resolution.

Multi-shot composition. Builds a sequenced clip from distinct per-shot prompts and durations; creators can define each cut with customize or let intelligent plan the shot structure.

Pricing

DurationPrice
3 seconds$1.260
4 seconds$1.680
5 seconds$2.100
6 seconds$2.520
7 seconds$2.940
8 seconds$3.360
9 seconds$3.780
10 seconds$4.200
11 seconds$4.620
12 seconds$5.040
13 seconds$5.460
14 seconds$5.880
15 seconds$6.300

When to Use

✅ Good fit❌ Consider alternatives
The scene should be created entirely from written directionA source image must anchor the opening frame
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "16:9",
  "duration": 5,
  "prompt": "Close-up of glowing fireflies dancing in a dark forest at twilight. Soft bioluminescent particles float through the air. Shallow depth of field, bokeh lights in background. Magical atmosphere, gentle movement."
}

Technical Specs

SpecValue
Model IDkwaivgi/kling-video/v3/4k/text-to-video
Inputsaspect_ratio, duration, generate_audio, multi_prompt, prompt, shot_type
Required inputsprompt
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Duration3 / 4 / 5 / 6 / 7 / 8 / 9 / 10 / 11 / 12 / 13 / 14 / 15
Aspect Ratio16:9 / 9:16 / 1:1

Related Models

Related Models

kwaivgi/kling-video/v3/4k/image-to-videoKling Video V3 4k by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/v3Kling Video V3 is KwaiVGI's unified video generation model. Generate or transform videos from prompts, images, reference videos, and multi-shot text while choosing standard or pro mode per request.kwaivgi/kling-video/v3/pro/image-to-videoKling Video V3 Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.kwaivgi/kling-video/v3/pro/text-to-videoKling Video V3 Pro by KwaiVGI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.kwaivgi/kling-video/v3/standard/image-to-videoKling Video V3 Standard by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/v3/standard/text-to-videoKling Video V3 Standard is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.kwaivgi/kling-video/3.0/omni/pro/video-to-video/editModify an existing video with natural-language instructions using Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/referenceUse a source video plus image references to generate a guided video variation with Kling 3.0 Omni Pro.