kwaivgi/kling-video/ai-avatar/2.0/standard
Kling Video Ai Avatar 2.0 by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
PNG, JPEG, WebP, or GIF · 20 MiB maximum
Example output — click Run to generate your own
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
https://api.sandbase.ai/v1/runkwaivgi/kling-video/ai-avatar/2.0/standardInput Schema
3 parameters · 2 required · 1 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
image | string | Required | The URL of the image to use as your avatar |
prompt | string | Required | The prompt to use for the video generation. · Default: "." |
audio | string | Optional | The URL of the audio file. |
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the generation task |
status | string | Task status: pending, running, completed, failed, timeout |
model | string | Model used for the generation |
outputs | array | Array of output items |
outputs[].url | string | URL of the generated artifact |
outputs[].content_type | string | MIME type (e.g. image/png, video/mp4) |
error | object | null | Error details if failed, null on success |
error.type | string | Machine-readable error type code |
error.message | string | Human-readable error description |
Async Workflow
This model uses asynchronous execution. Submit a request and poll for the result.
- Submit — POST to /v1/run, receive an
id - Poll — GET /v1/run/{id} until status is
completed,failed, ortimeout - Retrieve — Read
outputsfrom the completed response
Code Examples
Ready-to-run snippets
# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "kwaivgi/kling-video/ai-avatar/2.0/standard",
"audio": "https://static.sandbase.ai/examples/kwaivgi/kling-video/ai-avatar/2.0/standard/output_url_1.mp3",
"image": "https://static.sandbase.ai/examples/kwaivgi/kling-video/ai-avatar/2.0/standard/input_image_0.jpg",
"prompt": "."
}'
# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
-H "Authorization: Bearer YOUR_API_KEY"API README
Kling AI Avatar 2.0 Standard
standard in the Kling Video lineup is built to turn a portrait and speech into a coordinated presenter performance. The kling video ai avatar 2.0 standard configuration combines that transformation with the quality and motion profile represented by this exact model version, so teams can choose it deliberately among adjacent family variants. It is useful when the input already establishes part of the creative intent and the model must supply a polished result rather than a generic media conversion, with subject identity, scene logic, visual hierarchy, and the delivery goal kept explicit.
For production work with kwaivgi/kling-video/ai-avatar/2.0/standard, begin with the non-negotiable content, then describe the desired change, framing, action, atmosphere, and finishing cues in that order. Separate what must remain recognizable from what may vary, and prefer concrete nouns and observable actions over abstract praise. That structure makes outputs easier to compare across storyboard passes, campaign variants, catalog assets, and other repeatable creative pipelines.
Highlights
Audio-driven performance. Turns speech into coordinated mouth shapes, timing, expression, and presenter motion.
Portrait identity retention. Keeps the speaker recognizable while animating a still source.
Expressive facial behavior. Adds blinks, gaze changes, and natural micro-movements beyond mouth replacement.
kling video ai avatar 2.0 standard presenter-ready motion. Creates composed delivery for explainers, localized messages, and digital hosts. This is the defining creative strength of the kling video ai avatar 2.0 standard configuration.
Pricing
| Duration | Price per generated clip |
|---|---|
| Per generated second | $0.0562 |
Billing follows params.duration * 0.0562: multiply generated seconds by the $0.0562 per-second rate.
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| The model's named workflow matches the source material and intended output | A different input modality or model route is required |
| A managed asynchronous result is suitable for the production pipeline | A synchronous, interactive editor is essential |
| The documented controls cover the required duration, framing, or format | The project needs controls outside this endpoint's schema |
| Creative iteration benefits from a repeatable request structure | Exact deterministic pixels, frames, geometry, or samples are mandatory |
| A finished downloadable media asset is the desired deliverable | Editable source layers or a native project file are required |
Prompt Guide
For audio transformation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.
{
"audio": "https://static.sandbase.ai/examples/kwaivgi/kling-video/ai-avatar/2.0/standard/output_url_1.mp3",
"image": "https://static.sandbase.ai/examples/kwaivgi/kling-video/ai-avatar/2.0/standard/input_image_0.jpg",
"prompt": "."
}
Technical Specs
| Spec | Value |
|---|---|
| Model ID | kwaivgi/kling-video/ai-avatar/2.0/standard |
| Inputs | audio, image, prompt |
| Required inputs | prompt, image |
| Output fields | content_type, url |
| Execution | Async (submit, then poll for result) |

