SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsvideo generation api

alibaba/wan/ati

Wan Ati is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Input
The text prompt to guide video generation.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the input image.
Resolution of the generated video (480p, 580p, 720p). Allowed values: 480p, 720p.
Motion tracks to guide video generation. Each track is a sequence of points defining a motion trajectory. Multiple tracks can control different elements or objects in the video. Expected format: array of tracks, where each track is an array of points with 'x' and 'y' coordinates (up to 121 points per track). Points will be automatically padded to 121 if fewer are provided. Coordinates should be within the image dimensions.
Random seed for reproducibility. If None, a random seed is chosen.
Idle

Example output — click Run to generate your own

API README

Wan Ati

Wan Ati animates a still image with Wan’s image-to-video modeling, turning the source composition into a dynamic clip while maintaining recognizable characters and scene elements. The model adds fluid action and camera movement rather than merely applying a pan or zoom to the original frame.

Choose a source with a clear subject and describe how pose, expression, environment, and camera should evolve. It is useful for portrait animation, product reveals, illustration motion, and social creative where the first frame already carries approved art direction and the task is to introduce convincing movement.

Highlights

  • Fluid still-image animation. Transforms a static composition into evolving subject and environmental motion.
  • Character appearance continuity. Keeps recognizable visual traits tied to the supplied image during animation.
  • Camera-movement synthesis. Introduces directed reveals, pushes, pans, or orbiting viewpoints as part of the generated shot.
  • Scene-aware dynamics. Adds movement that responds to the visible setting instead of treating foreground and background independently.

Pricing

ConfigurationBilling unitPrice
Base generationPer request$0.15

When to Use

✅ Good fit❌ Consider alternatives
The project needs this exact audio-to-image workflowThe intended task belongs to a different media route
Available source media matches every required fieldRequired assets or usage rights are unavailable
The brief can define subject, action, camera, and styleOutput must be deterministic at frame or pixel level
Supported duration, resolution, and framing fit deliveryFinal placement requires unsupported specifications
An asynchronous generation job fits productionA live or frame-synchronous response is mandatory

Prompt Guide

Describe the result as a shot or design brief: identify subjects and references, state the action or transformation, specify environment and composition, then add camera behavior, lighting, pacing, style, sound, and preservation constraints where relevant. Use exact reference identifiers exposed by the local schema.

{
  "image": "https://example.com/start-frame.png",
  "prompt": "A cinematic scene with clearly directed subject action, camera movement, lighting, pacing, and atmosphere",
  "resolution": "480p",
  "seed": 1,
  "track": [
    "Example"
  ]
}

Technical Specs

SpecValue
Model IDalibaba/wan/ati
Input fieldsseed (integer)<br>image (string)<br>track (array)<br>prompt (string)<br>resolution (string; 480p, 720p)
Required inputprompt, image
Output fieldsurl, content_type
ExecutionAsynchronous job

Related Models

Related Models

alibaba/wan/2.1/flf-to-videoWan 2.1 Flf To Video by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.alibaba/wan/2.1/image-to-videoWan 2.1 by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.alibaba/wan/2.1/image-to-video/loraWan 2.1 Lora is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.alibaba/wan/2.1/text-to-videoWan 2.1 is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.alibaba/wan/2.1/text-to-video/loraWan 2.1 Lora by Alibaba - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.alibaba/wan/2.1/vaceWan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.alibaba/wan/2.1/vace/depthWan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.alibaba/wan/2.1/vace/inpaintingWan 2.1 Vace is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.