alibaba/wan/2.6/image-to-video/flash
Wan 2.6 Flash by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
PNG, JPEG, WebP, or GIF · 20 MiB maximum
Example output — click Run to generate your own
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
https://api.sandbase.ai/v1/runalibaba/wan/2.6/image-to-video/flashInput Schema
7 parameters · 2 required · 5 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
image | string | Required | URL of the image to use as the first frame. Must be publicly accessible or base64 data URI. Image dimensions must be between 240 and 7680. |
prompt | string | Required | The text prompt describing the desired video motion. Max 800 characters. · Min length: 1 |
seed | integer | Optional | Random seed for reproducibility. If None, a random seed is chosen. |
audio | string | Optional | URL of the audio to use as the background music. Must be publicly accessible. Limit handling: If the audio duration exceeds the duration value (5, 10, or 15 seconds), the audio is truncated to the first N seconds, and the rest is discarded. If the audio is shorter than the video, the remaining part of the video will be silent. For example, if the audio is 3 seconds long and the video duration is 5 seconds, the first 3 seconds of the output video will have sound, and the last 2 seconds will be silent. - Format: WAV, MP3. - Duration: 3 to 30 s. - File size: Up to 15 MB. |
duration | integer | Optional | Duration of the generated video in seconds. Choose between 5, 10 or 15 seconds. · Options: 5, 10, 15 · Default: 5 51015 |
resolution | string | Optional | Video resolution. Valid values: 720p, 1080p · Options: 720p, 1080p · Default: "1080p" 720p1080p |
multi_shots | boolean | Optional | When true, enables intelligent multi-shot segmentation. Only active when enable_prompt_expansion is True. Set to false for single-shot generation. · Default: false |
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the generation task |
status | string | Task status: pending, running, completed, failed, timeout |
model | string | Model used for the generation |
outputs | array | Array of output items |
outputs[].url | string | URL of the generated artifact |
outputs[].content_type | string | MIME type (e.g. image/png, video/mp4) |
error | object | null | Error details if failed, null on success |
error.type | string | Machine-readable error type code |
error.message | string | Human-readable error description |
Async Workflow
This model uses asynchronous execution. Submit a request and poll for the result.
- Submit — POST to /v1/run, receive an
id - Poll — GET /v1/run/{id} until status is
completed,failed, ortimeout - Retrieve — Read
outputsfrom the completed response
Code Examples
Ready-to-run snippets
# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "alibaba/wan/2.6/image-to-video/flash",
"image": "https://static.sandbase.ai/examples/bfl/flux-2/dev/edit/input_images_0.png",
"prompt": "A comedic cinematic demo where typed prompts physically transform reality. Photoreal, strong match cuts, coherent main character, no subtitles.\n\nShot 1 [0-4s] Continue from first frame. The creator presses \"PRINT\". The machine clunks like a spaceship. Creator whispers: \"Okay… I'm pressing enter.\"\nShot 2 [4-8s] Smash cut: the printed paper flies into the air and unfolds into a full desert canyon scene around the desk, like reality is being unrolled. Creator says: \"Wait—my prompt has physics?\"\nShot 3 [8-12s] Hard cut: the paper tears and reveals a tropical jungle behind it, perfectly lit, cinematic sun. Creator laughs: \"This is exactly why we do AI.\"\nShot 4 [12-15s] Hard cut back to studio. The printer prints a final line (not shown clearly). Creator looks to camera: \"Multi-scene. Single prompt.\"",
"duration": 5,
"resolution": "1080p",
"multi_shots": false
}'
# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
-H "Authorization: Bearer YOUR_API_KEY"API README
Wan 2.6 Image to Video Flash
Wan 2.6 Image to Video Flash animates a supplied keyframe with the speed-oriented Wan 2.6 image-to-video model. The reference establishes character identity, framing, and scene design, while the model develops action, camera movement, environmental dynamics, and synchronized sound into a coherent short sequence.
Choose it for rapid storyboards, social variants, product motion, and character tests that must remain visibly connected to an approved still. Describe how the scene changes over time, including action beats, camera path, pacing, ambience, and sound, rather than repeating what is already visible in the source image.
Highlights
- Fast image-anchored animation. Builds fluid motion from a still reference through the lower-latency Flash generation path.
- Multi-shot scene development. Can turn one image-led idea into a sequence with planned shot changes and narrative progression.
- Synchronized audiovisual output. Coordinates generated sound with visible events when audio is part of the request.
- Reference identity continuity. Keeps the source character, product, or composition recognizable while action and viewpoint evolve.
Pricing
| Configuration | Billing unit | Price |
|---|---|---|
| 5 seconds | 720p | $0.250 |
| 5 seconds | 1080p | $0.375 |
| 10 seconds | 720p | $0.500 |
| 10 seconds | 1080p | $0.750 |
| 15 seconds | 720p | $0.750 |
| 15 seconds | 1080p | $1.125 |
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| The project needs this exact image-to-video workflow | The intended task belongs to a different media route |
| Available source media matches every required field | Required assets or usage rights are unavailable |
| The brief can define subject, action, camera, and style | Output must be deterministic at frame or pixel level |
| Supported duration, resolution, and framing fit delivery | Final placement requires unsupported specifications |
| An asynchronous generation job fits production | A live or frame-synchronous response is mandatory |
Prompt Guide
Describe the result as a shot or design brief: identify subjects and references, state the action or transformation, specify environment and composition, then add camera behavior, lighting, pacing, style, sound, and preservation constraints where relevant. Use exact reference identifiers exposed by the local schema.
{
"audio": "example",
"duration": 5,
"image": "https://example.com/start-frame.png",
"prompt": "A cinematic scene with clearly directed subject action, camera movement, lighting, pacing, and atmosphere",
"seed": 1
}
Technical Specs
| Spec | Value |
|---|---|
| Model ID | alibaba/wan/2.6/image-to-video/flash |
| Input fields | seed (integer)<br>audio (string)<br>image (string)<br>prompt (string)<br>duration (integer; 5, 10, 15)<br>resolution (string; 720p, 1080p)<br>multi_shots (boolean) |
| Required input | prompt, image |
| Output fields | url, content_type |
| Execution | Asynchronous job |
Related Models
alibaba/wan/2.6/image-to-video— Compare this concrete local family route.alibaba/wan/2.6— Compare this concrete local family route.alibaba/wan/2.6/edit— Compare this concrete local family route.

