SandBase is live — $1 in free credits on signupStart free ›
Use in agentVidu models

Vidu modelsvideo generation api

vidu/q3/image-to-video/turbo

Vidu Q3 Turbo by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

Input
Text prompt for video generation, max 2000 characters

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL or base64 image to use as the starting frame

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the image to use as the ending frame. When provided, generates a transition video between start and end frames.
Whether to use direct audio-video generation. When true, outputs video with sound (including dialogue and sound effects).
Output video resolution. Note: 360p is not available when end_image_url is provided. Allowed values: 360p, 540p, 720p, 1080p.
116
Duration of the video in seconds (1-16 for Q3 models) Range: 1 to 16.
Random seed for reproducibility. If None, a random seed is chosen.
Idle

Example output — click Run to generate your own

API README

Vidu Q3 Image to Video Turbo

Vidu Q3 Image to Video Turbo is a specialized SandBase endpoint built to animate a supplied still with a faster, lower-cost Q3 workflow designed for iteration. It is most useful when teams need rapid motion exploration, optional endpoint framing, native audio, and the same broad duration envelope, with the route’s inputs defining a repeatable production contract instead of leaving critical delivery choices to an ad-hoc manual workflow.

Use this exact route when its input mode matches the creative asset already in hand: describe the visual or spoken result clearly, set only the controls that support the intended delivery, and keep the subject, action, environment, camera or performance direction internally consistent. The result is returned asynchronously as a downloadable media URL suitable for review, automation, or downstream finishing.

Highlights

  • Fast still animation. Turbo is suited to trying several motion directions from one approved keyframe without rebuilding the visual premise.
  • Endpoint-aware transitions. An optional end image lets the sequence converge on a chosen composition, useful for reveals and matched cuts.
  • Audio generated with motion. Enable native sound when dialogue, environmental texture, or effects should follow the action in the clip.
  • Scalable delivery settings. Choose 1–16 seconds across 360p, 540p, 720p, or 1080p to balance preview speed and delivery detail.

Pricing

Billing follows the selected duration and resolution. The standard rate is $0.0350 per second at 360p/540p and $0.0770 per second at 720p/1080p.

DurationResolutionPrice
1 sec360p$0.0350
1 sec540p$0.0350
1 sec720p$0.0770
1 sec1080p$0.0770
2 sec360p$0.0700
2 sec540p$0.0700
2 sec720p$0.1540
2 sec1080p$0.1540
3 sec360p$0.1050
3 sec540p$0.1050
3 sec720p$0.2310
3 sec1080p$0.2310
4 sec360p$0.1400
4 sec540p$0.1400
4 sec720p$0.3080
4 sec1080p$0.3080
5 sec360p$0.1750
5 sec540p$0.1750
5 sec720p$0.3850
5 sec1080p$0.3850
6 sec360p$0.2100
6 sec540p$0.2100
6 sec720p$0.4620
6 sec1080p$0.4620
7 sec360p$0.2450
7 sec540p$0.2450
7 sec720p$0.5390
7 sec1080p$0.5390
8 sec360p$0.2800
8 sec540p$0.2800
8 sec720p$0.6160
8 sec1080p$0.6160
9 sec360p$0.3150
9 sec540p$0.3150
9 sec720p$0.6930
9 sec1080p$0.6930
10 sec360p$0.3500
10 sec540p$0.3500
10 sec720p$0.7700
10 sec1080p$0.7700
11 sec360p$0.3850
11 sec540p$0.3850
11 sec720p$0.8470
11 sec1080p$0.8470
12 sec360p$0.4200
12 sec540p$0.4200
12 sec720p$0.9240
12 sec1080p$0.9240
13 sec360p$0.4550
13 sec540p$0.4550
13 sec720p$1.0010
13 sec1080p$1.0010
14 sec360p$0.4900
14 sec540p$0.4900
14 sec720p$1.0780
14 sec1080p$1.0780
15 sec360p$0.5250
15 sec540p$0.5250
15 sec720p$1.1550
15 sec1080p$1.1550
16 sec360p$0.5600
16 sec540p$0.5600
16 sec720p$1.2320
16 sec1080p$1.2320

When to Use

ScenarioWhy this route fits
Concept developmentChoose Vidu Q3 Image to Video Turbo when its turbo workflow matches the starting material and you need several clearly directed variations.
Production iterationUse explicit duration, resolution, ratio, seed, or quality controls to compare versions without changing the core creative brief.
Channel adaptationGenerate directly in the landscape, square, portrait, or vertical format required by the destination whenever that control is available.
Automated pipelinesIntegrate the asynchronous media URL into review queues, asset libraries, publishing tools, or a later finishing stage.
Alternative routePick a related text-, image-, reference-, edit-, or turbo route when the available source media or required degree of control is different.

Prompt Guide

Lead with the main subject or source asset, then describe the intended action or transformation, environment, composition, camera or vocal delivery, lighting and mood. Keep instructions concrete and compatible; use the route’s explicit fields for duration, resolution, ratio, quality, voice, or reproducibility instead of burying those settings in prose.

{
  "prompt": "A cinematic product reveal with deliberate subject motion, coherent lighting, and a slow camera push.",
  "image": "https://example.com/input.jpg",
  "duration": 5,
  "resolution": "720p",
  "audio": true
}

Technical Specs

PropertyDetails
Model IDvidu/q3/image-to-video/turbo
Required inputsprompt, image
ExecutionAsynchronous; poll the returned generation ID
OutputDownloadable media URL
seedinteger; optional
audioboolean; optional; default: true
imagestring; required
promptstring; required; default:
durationinteger; optional; range: 1–16; default: 5
end_imagestring; optional
resolutionstring; optional; choices: 360p, 540p, 720p, 1080p; default: 720p

Related Models

ModelBest for
vidu/q3/text-to-videoAlternative text to video workflow
vidu/q3/image-to-videoAlternative image to video workflow
vidu/q3/reference-to-video/mixAlternative mix workflow
vidu/q3/text-to-video/turboAlternative turbo workflow

Related Models

vidu/q3/image-to-videoVidu Q3 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q3/reference-to-video/mixVidu Q3 Mix by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q3/text-to-videoVidu Q3 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.vidu/q3/text-to-video/turboVidu Q3 Turbo is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.vidu/q1/image-to-videoVidu Q1 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/reference-to-videoVidu Q1 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/start-end-to-videoVidu Q1 Start End To Video by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/text-to-videoVidu Q1 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.