Wan 2.6 Flash Image to Video
alibaba/wan/2.6/flash/image-to-videoWan 2.6 Flash by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
- Base price
- $0.25USD / run
- Execution
- async
- Model type
- video
- Input fields
- 7
Try the model
Playground
PNG, JPEG, WebP, or GIF · 20 MiB maximum
Your output will appear here
Complete the inputs, then click Run.
Specifications
Pricing
- Base price
- $0.25 / run
- Billing formula
- (params.resolution ?? "720p") == "1080p" ? (params.duration ?? 5) * 0.075 : (params.duration ?? 5) * 0.05
Context & modalities
- Input
- Schema-defined
- Output
- video
Capabilities
- Chat
- Not supported
- Vision
- Not supported
- Reasoning
- Not supported
- Structured output
- Not supported
- Function calling
- Not supported
- Audio input
- Not supported
Access
- Provider
- Alibaba
- Model ID
- alibaba/wan/2.6/flash/image-to-video
- Execution
- async
- API
- Unified Run API
- Endpoint
- /v1/run
API README
Wan 2.6 Flash Image to Video
Wan 2.6 Flash Image to Video animates a supplied keyframe with the speed-oriented Wan 2.6 image-to-video model. The reference establishes character identity, framing, and scene design, while the model develops action, camera movement, environmental dynamics, and synchronized sound into a coherent short sequence.
Choose it for rapid storyboards, social variants, product motion, and character tests that must remain visibly connected to an approved still. Describe how the scene changes over time, including action beats, camera path, pacing, ambience, and sound, rather than repeating what is already visible in the source image.
Highlights
- Fast image-anchored animation. Builds fluid motion from a still reference through the lower-latency Flash generation path.
- Multi-shot scene development. Can turn one image-led idea into a sequence with planned shot changes and narrative progression.
- Synchronized audiovisual output. Coordinates generated sound with visible events when audio is part of the request.
- Reference identity continuity. Keeps the source character, product, or composition recognizable while action and viewpoint evolve.
Pricing
| Configuration | Billing unit | Price |
|---|---|---|
| 5 seconds | 720p | $0.250 |
| 5 seconds | 1080p | $0.375 |
| 10 seconds | 720p | $0.500 |
| 10 seconds | 1080p | $0.750 |
| 15 seconds | 720p | $0.750 |
| 15 seconds | 1080p | $1.125 |
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| The project needs this exact image-to-video workflow | The intended task belongs to a different media route |
| Available source media matches every required field | Required assets or usage rights are unavailable |
| The brief can define subject, action, camera, and style | Output must be deterministic at frame or pixel level |
| Supported duration, resolution, and framing fit delivery | Final placement requires unsupported specifications |
| An asynchronous generation job fits production | A live or frame-synchronous response is mandatory |
Prompt Guide
Describe the result as a shot or design brief: identify subjects and references, state the action or transformation, specify environment and composition, then add camera behavior, lighting, pacing, style, sound, and preservation constraints where relevant. Use exact reference identifiers exposed by the local schema.
{
"audio": "example",
"duration": 5,
"image": "https://example.com/start-frame.png",
"prompt": "A cinematic scene with clearly directed subject action, camera movement, lighting, pacing, and atmosphere",
"seed": 1
}
Technical Specs
| Spec | Value |
|---|---|
| Model ID | alibaba/wan/2.6/flash/image-to-video |
| Input fields | seed (integer)<br>audio (string)<br>image (string)<br>prompt (string)<br>duration (integer; 5, 10, 15)<br>resolution (string; 720p, 1080p)<br>multi_shots (boolean) |
| Required input | prompt, image |
| Output fields | url, content_type |
| Execution | Asynchronous job |
Related Models
alibaba/wan/2.6/image-to-video— Compare this concrete local family route.alibaba/wan/2.6/text-to-image— Compare this concrete local family route.alibaba/wan/2.6/image-to-image— Compare this concrete local family route.
Start building
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
-X POST "https://api.sandbase.ai/v1/run" \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
-H "Content-Type: application/json" \
--data-binary @- <<'SANDBASE_JSON'
{
"model": "alibaba/wan/2.6/flash/image-to-video",
"image": "https://static.sandbase.ai/examples/bfl/flux-2/dev/edit/input_images_0.png",
"prompt": "A comedic cinematic demo where typed prompts physically transform reality. Photoreal, strong match cuts, coherent main character, no subtitles.\n\nShot 1 [0-4s] Continue from first frame. The creator presses \"PRINT\". The machine clunks like a spaceship. Creator whispers: \"Okay… I'm pressing enter.\"\nShot 2 [4-8s] Smash cut: the printed paper flies into the air and unfolds into a full desert canyon scene around the desk, like reality is being unrolled. Creator says: \"Wait—my prompt has physics?\"\nShot 3 [8-12s] Hard cut: the paper tears and reveals a tropical jungle behind it, perfectly lit, cinematic sun. Creator laughs: \"This is exactly why we do AI.\"\nShot 4 [12-15s] Hard cut back to studio. The printer prints a final line (not shown clearly). Creator looks to camera: \"Multi-scene. Single prompt.\"",
"duration": 5,
"resolution": "720p",
"multi_shots": false
}
SANDBASE_JSON
)
run_id=$(printf '%s' "$result" | jq -r .id)
for attempt in $(seq 1 120); do
status=$(printf '%s' "$result" | jq -r .status)
case "$status" in completed|failed|timeout) break ;; esac
sleep 2
result=$(curl --fail-with-body --silent \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
"https://api.sandbase.ai/v1/run/$run_id")
done
status=$(printf '%s' "$result" | jq -r .status)
[ "$status" = completed ] || { echo "Generation ended: $status" >&2; exit 1; }
printf '%s\n' "$result"Choose your model
Compare models
| Model | Type | Default price | Released |
|---|---|---|---|
Wan 2.6 Flash Image to VideoThis model Alibaba | Video | $0.25 / run | Jan 18, 2026 |
Alibaba | Video | $0.50 / run | Oct 5, 2026 |
Alibaba | Video | $0.50 / run | Oct 5, 2026 |
Alibaba | Video | $1.40 / run | Oct 4, 2026 |
Alibaba | Video | $0.70 / run | Oct 4, 2026 |
Alibaba | Video | $1.00 / run | Oct 4, 2026 |
Questions
FAQ
How do I call Wan 2.6 Flash Image to Video through SandBase?
Create a SandBase API key, then send requests with the model ID "alibaba/wan/2.6/flash/image-to-video" to Unified Run API (/v1/run) at https://api.sandbase.ai. The request examples on this page show the exact payload.
Do I need a separate Alibaba account?
No. One SandBase API key and balance gives you access to Wan 2.6 Flash Image to Video and the other models in the catalog; you do not need to sign up with Alibaba separately.
