SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsvideo generation api

alibaba/wan/move

Wan Move by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

Input
Text prompt to guide the video generation.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the input image. If the input image does not match the chosen aspect ratio, it is resized and center cropped.
A list of trajectories. Each trajectory list means the movement of one object.
Random seed for reproducibility. If None, a random seed is chosen.
Idle

Example output — click Run to generate your own

API README

Wan Move

Wan Move brings a supplied photograph or illustration into motion, generating cinematic subject action, camera travel, and environmental response around the established frame. It is designed for image-led animation where the source should remain visually recognizable as the clip develops.

Describe the movement that begins after the source frame: subject gesture, secondary motion such as hair or fabric, camera path, depth changes, ambience, and any sound intention. It fits character moments, product showcases, landscape animation, and social clips built from finished still artwork.

Highlights

  • Cinematic photo animation. Develops an approved still into a composed moving shot rather than a simple slideshow effect.
  • Natural secondary motion. Animates hair, fabric, particles, reflections, and environment around the primary action.
  • Source identity preservation. Keeps subjects and art direction recognizable throughout the generated movement.
  • Camera-and-audio staging. Can combine directed camera behavior with optional sound for a more complete scene.

Pricing

Billing unitPrice
Per request$0.2

When to Use

✅ Good fit❌ Consider alternatives
The model's named workflow matches the source material and intended outputA different input modality or model route is required
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For image-conditioned generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "image": "https://example.com/start-frame.png",
  "prompt": "In a boxing gym, the camera captures a half-body view of a single man practicing shadow boxing inside the ring. The camera sways slightly, emphasizing movement and intensity. He wears black boxing gloves, black shorts, and boxing shoes. With no opponent present, he throws a controlled straight punch into the air, his lead arm fully extended while the other hand stays tight near his face in a defensive guard. His body rotates subtly through the hips and shoulders, and his feet shift lightly across the mat to maintain balance and rhythm. Sweat highlights his muscular form as he focuses forward, fully immersed in his training. The background features blurred gym equipment and hanging heavy bags, reinforcing the sense of motion and concentration."
}

Technical Specs

SpecValue
Model IDalibaba/wan/move
Inputsimage, prompt, seed, trajectories
Required inputsprompt, image
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)

Related Models

Related Models

alibaba/wan/2.1/flf-to-videoWan 2.1 Flf To Video by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.alibaba/wan/2.1/image-to-videoWan 2.1 by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.alibaba/wan/2.1/image-to-video/loraWan 2.1 Lora is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.alibaba/wan/2.1/text-to-videoWan 2.1 is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.alibaba/wan/2.1/text-to-video/loraWan 2.1 Lora by Alibaba - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.alibaba/wan/2.1/vaceWan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.alibaba/wan/2.1/vace/depthWan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.alibaba/wan/2.1/vace/inpaintingWan 2.1 Vace is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.