SandBase is live — $1 in free credits on signupStart free ›
Use in agentVidu models

Vidu modelsvideo generation api

vidu/q3/text-to-video/turbo

Vidu Q3 Turbo is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Input
Text prompt for video generation, max 2000 characters
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
Whether to use direct audio-video generation. When true, outputs video with sound.
Output video resolution Allowed values: 360p, 540p, 720p, 1080p.
116
Duration of the video in seconds Range: 1 to 16.
Random seed for reproducibility. If None, a random seed is chosen.
Idle

Example output — click Run to generate your own

API README

Vidu Q3 Text to Video Turbo

Vidu Q3 Text to Video Turbo is a specialized SandBase endpoint built to create prompt-led video drafts quickly and at a lower per-second rate than the standard Q3 route. It is most useful when teams need fast concept iteration, integrated sound, flexible framing, and long-enough timing for complete beats, with the route’s inputs defining a repeatable production contract instead of leaving critical delivery choices to an ad-hoc manual workflow.

Use this exact route when its input mode matches the creative asset already in hand: describe the visual or spoken result clearly, set only the controls that support the intended delivery, and keep the subject, action, environment, camera or performance direction internally consistent. The result is returned asynchronously as a downloadable media URL suitable for review, automation, or downstream finishing.

Highlights

  • Rapid concept exploration. Turbo makes it practical to compare alternate scenes, actions, and camera ideas before committing to a higher-cost render path.
  • Complete text-led generation. Build the subject, environment, lighting, and motion from a prompt with no image preparation step.
  • Sound in the first draft. Native audio can establish dialogue, ambience, and effects early enough to judge the whole creative direction.
  • Delivery-shaped outputs. Set 1–16 seconds, ten aspect ratios, and four resolution tiers to test the intended placement rather than a generic canvas.

Pricing

Billing follows the selected duration and resolution. The standard rate is $0.0350 per second at 360p/540p and $0.0770 per second at 720p/1080p.

DurationResolutionPrice
1 sec360p$0.0350
1 sec540p$0.0350
1 sec720p$0.0770
1 sec1080p$0.0770
2 sec360p$0.0700
2 sec540p$0.0700
2 sec720p$0.1540
2 sec1080p$0.1540
3 sec360p$0.1050
3 sec540p$0.1050
3 sec720p$0.2310
3 sec1080p$0.2310
4 sec360p$0.1400
4 sec540p$0.1400
4 sec720p$0.3080
4 sec1080p$0.3080
5 sec360p$0.1750
5 sec540p$0.1750
5 sec720p$0.3850
5 sec1080p$0.3850
6 sec360p$0.2100
6 sec540p$0.2100
6 sec720p$0.4620
6 sec1080p$0.4620
7 sec360p$0.2450
7 sec540p$0.2450
7 sec720p$0.5390
7 sec1080p$0.5390
8 sec360p$0.2800
8 sec540p$0.2800
8 sec720p$0.6160
8 sec1080p$0.6160
9 sec360p$0.3150
9 sec540p$0.3150
9 sec720p$0.6930
9 sec1080p$0.6930
10 sec360p$0.3500
10 sec540p$0.3500
10 sec720p$0.7700
10 sec1080p$0.7700
11 sec360p$0.3850
11 sec540p$0.3850
11 sec720p$0.8470
11 sec1080p$0.8470
12 sec360p$0.4200
12 sec540p$0.4200
12 sec720p$0.9240
12 sec1080p$0.9240
13 sec360p$0.4550
13 sec540p$0.4550
13 sec720p$1.0010
13 sec1080p$1.0010
14 sec360p$0.4900
14 sec540p$0.4900
14 sec720p$1.0780
14 sec1080p$1.0780
15 sec360p$0.5250
15 sec540p$0.5250
15 sec720p$1.1550
15 sec1080p$1.1550
16 sec360p$0.5600
16 sec540p$0.5600
16 sec720p$1.2320
16 sec1080p$1.2320

When to Use

ScenarioWhy this route fits
Concept developmentChoose Vidu Q3 Text to Video Turbo when its turbo workflow matches the starting material and you need several clearly directed variations.
Production iterationUse explicit duration, resolution, ratio, seed, or quality controls to compare versions without changing the core creative brief.
Channel adaptationGenerate directly in the landscape, square, portrait, or vertical format required by the destination whenever that control is available.
Automated pipelinesIntegrate the asynchronous media URL into review queues, asset libraries, publishing tools, or a later finishing stage.
Alternative routePick a related text-, image-, reference-, edit-, or turbo route when the available source media or required degree of control is different.

Prompt Guide

Lead with the main subject or source asset, then describe the intended action or transformation, environment, composition, camera or vocal delivery, lighting and mood. Keep instructions concrete and compatible; use the route’s explicit fields for duration, resolution, ratio, quality, voice, or reproducibility instead of burying those settings in prose.

{
  "prompt": "A cinematic product reveal with deliberate subject motion, coherent lighting, and a slow camera push.",
  "duration": 5,
  "resolution": "720p",
  "aspect_ratio": "21:9",
  "audio": true
}

Technical Specs

PropertyDetails
Model IDvidu/q3/text-to-video/turbo
Required inputsprompt
ExecutionAsynchronous; poll the returned generation ID
OutputDownloadable media URL
seedinteger; optional
audioboolean; optional; default: true
promptstring; required
durationinteger; optional; range: 1–16; default: 5
resolutionstring; optional; choices: 360p, 540p, 720p, 1080p; default: 720p
aspect_ratiostring; optional; choices: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16

Related Models

ModelBest for
vidu/q3/text-to-videoAlternative text to video workflow
vidu/q3/image-to-videoAlternative image to video workflow
vidu/q3/reference-to-video/mixAlternative mix workflow
vidu/q3/image-to-video/turboAlternative turbo workflow

Related Models

vidu/q3/text-to-videoVidu Q3 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.vidu/q3/image-to-video/turboVidu Q3 Turbo by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q3/reference-to-video/mixVidu Q3 Mix by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q3/image-to-videoVidu Q3 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/image-to-videoVidu Q1 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/reference-to-videoVidu Q1 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/start-end-to-videoVidu Q1 Start End To Video by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.vidu/q1/text-to-videoVidu Q1 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.