SandBase is live — $1 in free credits on signupStart free ›

KwaiVGI modelsvideo generation api

kwaivgi/kling-video/v2.6/pro/image-to-video

Kling Video V2.6 Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Input

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the image to be used for the video

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the image to be used for the end of the video
The duration of the generated video in seconds Allowed values: 5, 10.
Whether to generate native audio for the video. Supports Chinese and English voice output. Other languages are automatically translated to English. For English speech, use lowercase letters; for acronyms or proper nouns, use uppercase.
Optional Voice IDs for video generation. Reference voices in your prompt with <<<voice_1>>> and <<<voice_2>>> (maximum 2 voices per task). Get voice IDs from the kling video create-voice endpoint: https://fal.ai/models/fal-ai/kling-video/create-voice
Idle

Example output — click Run to generate your own

API README

Kling Video V2.6 Pro Image-to-Video

Kling Video V2.6 Pro animates a supplied start image from a written direction. The image establishes the opening composition while the prompt describes motion, camera, and sound. An optional end image can guide the last frame.

The route produces either a 5- or 10-second video and returns a video URL. Native audio is enabled by default, and optional voice IDs can be referenced in the prompt. The request remains centered on the required image and prompt.

Highlights

  • Cinematic image animation. The exact V2.6 Pro route is documented for cinematic visuals. It applies that positioning to a supplied opening image.
  • Fluid motion. Fluid motion is explicitly named for this version, tier, and route. Motion is generated from the image and written action.
  • Improved visual quality. The exact API documentation identifies improved visual quality for V2.6 Pro image-to-video. This is a model result rather than a schema setting.
  • Improved motion consistency. The same exact documentation identifies improved motion consistency. It complements the route's fluid-motion positioning.

Pricing

DurationPrice
5 seconds$0.70
10 seconds$1.40

When to Use

✅ Good fit❌ Consider alternatives
Animate a supplied still into a short action.Generate without a starting image; use text-to-video.
Direct movement and camera behavior from the opening frame.Require a duration other than 5 or 10 seconds.
Guide the final state with an optional end image.Require a selectable aspect ratio independent of the image.
Generate speech and scene sound together with the motion.Require more than two optional voice bindings.
Use this exact V2.6 Pro workflow.Require controls that are not present in the local contract.

Prompt Guide

Describe motion relative to the supplied opening image. State the subject action, camera path, environmental movement, dialogue or sound, and any final-frame transition.

Subject action: [movement over time]
Camera: [framing and movement]
Environment: [light and background motion]
Timing: [pace and progression]
Dialogue and sound: [spoken lines, ambience, effects]
Final state: [arrival at the optional end image]
{
  "image": "<start-image-url>",
  "prompt": "A king walks slowly and says \"My people, here I am! I am here to save you all.\"",
  "duration": 5,
  "generate_audio": true
}

Technical Specs

SpecValue
Required inputsimage, prompt
Prompt lengthUp to 2,500 characters
Duration5 or 10 seconds; default 5
Frame guidanceRequired image; optional end_image
Native audioOptional; default enabled
Voice IDsOptional array; maximum two documented locally
OutputVideo URL with optional file metadata

Related

Related Models

kwaivgi/kling-video/v2.6/pro/text-to-videoKling Video V2.6 Pro by KwaiVGI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.kwaivgi/kling-video/1.6/elementsKling Video 1.6 Elements by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/1.6/pro/effectsKling Video 1.6 Pro is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.kwaivgi/kling-video/1.6/pro/elementsKling Video 1.6 Pro by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.kwaivgi/kling-video/3.0/omni/pro/image-to-videoKling Video Omni Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.kwaivgi/kling-video/3.0/omni/pro/reference-to-videoGenerate a video from multiple reference images and text guidance with Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/editModify an existing video with natural-language instructions using Kling 3.0 Omni Pro.kwaivgi/kling-video/3.0/omni/pro/video-to-video/referenceUse a source video plus image references to generate a guided video variation with Kling 3.0 Omni Pro.