SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsvideo generation api

alibaba/wan/2.2/animate/replace

Wan 2.2 Animate is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

Input

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the input image. If the input image does not match the chosen aspect ratio, it is resized and center cropped.
URL of the input video.
Resolution of the generated video (480p, 580p, or 720p). Allowed values: 480p, 720p.
If true, applies quality enhancement for faster generation with improved quality. When enabled, parameters are automatically optimized for best results.
Random seed for reproducibility. If None, a random seed is chosen.
Idle

Example output — click Run to generate your own

API README

Wan 2.2 Animate Replace

Wan 2.2 Animate Replace substitutes a reference character into an existing driving video while retaining the source scene, staging, and surrounding interactions. It is intended for character replacement rather than free-form generation: the video supplies performance and environment, and the image supplies the new performer's visual identity.

The Wan-Animate pipeline jointly follows holistic body motion and facial expression, using prepared pose, face, background, and mask signals to keep the replacement connected to the shot. Because large body-proportion changes can disrupt contact with props or other actors, replacement favors spatial continuity over aggressive pose retargeting.

Highlights

Scene-aware character replacement. Preserves the driving video's environment and staging while substituting the reference character.

Body and face performance transfer. Replicates holistic movement together with facial expression instead of treating them as unrelated passes.

Foreground-aware synthesis. Uses pose, face, background, and mask information to integrate the replacement with the original scene.

Interaction-conscious continuity. Prioritizes spatial relationships with props, surfaces, and nearby subjects during the substitution.

Pricing

ResolutionPrice
480p$0.040 per second
720p$0.080 per second

When to Use

✅ Good fit❌ Consider alternatives
The model's named workflow matches the source material and intended outputA different input modality or model route is required
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For image-conditioned generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "image": "https://static.sandbase.ai/examples/alibaba/wan/2.2/animate/replace/input_image_0.png",
  "resolution": "480p",
  "video": "https://static.sandbase.ai/examples/alibaba/wan/2.2/animate/replace/input_video_1.mp4"
}

Technical Specs

SpecValue
Model IDalibaba/wan/2.2/animate/replace
Inputsimage, resolution, seed, use_turbo, video
Required inputsimage
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Resolution480p / 720p

Related Models

Related Models