alibaba/wan/2.2/image-to-video
Wan 2.2 by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
PNG, JPEG, WebP, or GIF · 20 MiB maximum
PNG, JPEG, WebP, or GIF · 20 MiB maximum
Example output — click Run to generate your own
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
https://api.sandbase.ai/v1/runalibaba/wan/2.2/image-to-videoInput Schema
6 parameters · 2 required · 4 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
image | string | Required | URL of the input image. If the input image does not match the chosen aspect ratio, it is resized and center cropped. |
prompt | string | Required | The text prompt to guide video generation. |
seed | integer | Optional | Random seed for reproducibility. If None, a random seed is chosen. |
end_image | string | Optional | URL of the end image. |
resolution | string | Optional | Resolution of the generated video (480p, 580p, or 720p). · Options: 480p, 720p · Default: "720p" 480p720p |
aspect_ratio | string | Optional | The aspect ratio of the generated image. · Options: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16 21:916:93:24:35:41:14:53:42:39:16 |
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the generation task |
status | string | Task status: pending, running, completed, failed, timeout |
model | string | Model used for the generation |
outputs | array | Array of output items |
outputs[].url | string | URL of the generated artifact |
outputs[].content_type | string | MIME type (e.g. image/png, video/mp4) |
error | object | null | Error details if failed, null on success |
error.type | string | Machine-readable error type code |
error.message | string | Human-readable error description |
Async Workflow
This model uses asynchronous execution. Submit a request and poll for the result.
- Submit — POST to /v1/run, receive an
id - Poll — GET /v1/run/{id} until status is
completed,failed, ortimeout - Retrieve — Read
outputsfrom the completed response
Code Examples
Ready-to-run snippets
# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "alibaba/wan/2.2/image-to-video",
"image": "https://static.sandbase.ai/examples/alibaba/wan/2.2/image-to-video/input_image_0.jpg",
"prompt": "The white dragon warrior stands still, eyes full of determination and strength. The camera slowly moves closer or circles around the warrior, highlighting the powerful presence and heroic spirit of the character.",
"resolution": "720p"
}'
# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
-H "Authorization: Bearer YOUR_API_KEY"API README
Wan 2.2 Image to Video
Wan 2.2 Image-to-Video animates a supplied keyframe while using the prompt to direct subject movement, camera behavior, atmosphere, and the scene's progression. The dedicated I2V-A14B route is built for workflows in which an existing visual must remain the foundation rather than being reinvented from text.
Wan 2.2 uses a mixture-of-experts denoising design that assigns high-noise and low-noise stages to different experts, separating broad composition from later detail refinement. The endpoint supports 480p and 720p delivery, optional end-frame guidance, seed control, and a broad set of aspect ratios for adapting one source image to different placements.
Highlights
Dedicated image-to-video model. I2V-A14B is a separate checkpoint for animating source imagery at both 480p and 720p.
Two-expert denoising. Different experts handle early structural planning and later detail refinement while activating about 14B parameters per step.
Cinematic aesthetic control. Training annotations cover lighting, composition, contrast, color tone, and other visual qualities used in film-oriented prompting.
Start-and-end guidance. An optional ending image can define where the motion should arrive while the first image anchors its opening.
Pricing
| Resolution | Price |
|---|---|
| 480p | $0.040 per second |
| 720p | $0.080 per second |
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| A still image should anchor the video's subject and composition | The whole scene should be invented from text alone |
| A managed asynchronous result is suitable for the production pipeline | A synchronous, interactive editor is essential |
| The documented controls cover the required duration, framing, or format | The project needs controls outside this endpoint's schema |
| Creative iteration benefits from a repeatable request structure | Exact deterministic pixels, frames, geometry, or samples are mandatory |
| A finished downloadable media asset is the desired deliverable | Editable source layers or a native project file are required |
Prompt Guide
For image-conditioned generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.
{
"aspect_ratio": "21:9",
"image": "https://static.sandbase.ai/examples/alibaba/wan/2.2/image-to-video/input_image_0.jpg",
"prompt": "The white dragon warrior stands still, eyes full of determination and strength. The camera slowly moves closer or circles around the warrior, highlighting the powerful presence and heroic spirit of the character.",
"resolution": "720p"
}
Technical Specs
| Spec | Value |
|---|---|
| Model ID | alibaba/wan/2.2/image-to-video |
| Inputs | aspect_ratio, end_image, image, prompt, resolution, seed |
| Required inputs | prompt, image |
| Output fields | content_type, url |
| Execution | Async (submit, then poll for result) |
| Resolution | 480p / 720p |
| Aspect Ratio | 21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16 |
Related Models
alibaba/wan/2.2/image-to-video/lora— Compare a nearby route in the same local model family.alibaba/wan/2.2/image-to-video/turbo— Compare a nearby route in the same local model family.alibaba/wan/2.2/5b/fast/text-to-video— Compare a nearby route in the same local model family.

