Seedance 2.0 Fast Text to Video

Bytedancevideo

Published April 1, 2026

About

ByteDance's most advanced text-to-video model in its fast tier delivering lower latency and cost without compromising on cinematic output, native audio, multi-shot editing, and director-level camera control for professional-grade video creation.

Documentation

Seedance 2.0 Fast Text to Video

ByteDance's fast-tier text-to-video model — delivers lower latency and ~20% cost savings compared to the standard version, without compromising on cinematic output, native audio, and camera control.

Highlights

Lower latency — Optimized for speed, ideal for iterative workflows and rapid prototyping.

Native audio generation — Automatically creates synchronized sound effects, ambient audio, and lip-synced speech.

Flexible duration — Generate 4 to 15 second clips. Use -1 for model-intelligent duration selection.

Up to 720p — Supports 480p and 720p output (1080p not available in fast tier).

Pricing

ResolutionPrice per second
480p$0.08
720p$0.16

Default: 5 seconds at 720p = $0.80. Audio generation included at no extra cost.

Technical Specs

SpecValue
InputText prompt
OutputMP4 video with optional audio
Resolution480p / 720p
Aspect ratios16:9, 9:16, 1:1, 4:3, 3:4, 21:9
Duration4–15 seconds
ExecutionAsync (submit → poll for result)
Typical latency1–5 minutes

Related

Try Seedance 2.0 Fast Text to Video

Test this model in the Sandbase Playground with your own prompts.

Open in Playground

Related Models