SandBase is live — $1 in free credits on signupStart free ›

sandbase-ai modelsvideo generation api

sandbase-ai/ai-avatar-multi-text

Ai Avatar Multi Text by sandbase-ai - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

Input
The text prompt to guide video generation.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the input image. If the input image does not match the chosen aspect ratio, it is resized and center cropped.
Resolution of the video to generate. Must be either 480p or 720p. Allowed values: 480p, 720p.
The first person's voice to use for speech generation Allowed values: Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill.
The text input to guide video generation.
The text input to guide video generation.
The second person's voice to use for speech generation Allowed values: Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill.
04294967295
Random seed for reproducibility. If None, a random seed is chosen. Range: 0 to 4294967295.
Idle

Example output — click Run to generate your own

API README

AI Avatar Multi Text

AI Avatar Multi Text occupies the reference-led-video role in the model catalog for application builders. Its practical value is to separate this job from adjacent family routes: an upstream service can confirm that it has prompt, image, submit one well-scoped job, and receive a rendered video asset for downstream approval. This positioning explains when the exact sandbase-ai/ai-avatar-multi-text route belongs in an architecture without borrowing promises from another version, tier, or task variant.

A useful acceptance workflow begins by checking the request intent, asset eligibility, and downstream placement. The caller records required values separately from optional choices such as seed, voice1, voice2, resolution, first_text_input, second_text_input, runs a small representative sample, and stores the route name with the resulting artifact. Reviewers then compare that artifact with the agreed acceptance criteria, document any rejected setup, and promote only a successful request shape into larger batches. This workflow gives sandbase-ai/ai-avatar-multi-text measurable business context while leaving the Highlights section to explain model behavior.

Highlights

  • Two written dialogue channels. first_text_input and second_text_input provide independent lines for the pictured participants.
  • Per-speaker voice selection. voice1 and voice2 assign distinct synthesized voices to the two text tracks.
  • Multi-person still-image animation. The required image establishes the participants and scene before their written dialogue is rendered.
  • Resolution choice. The route exposes its declared resolution setting alongside dialogue and voice decisions.

Pricing

Billing optionPrice
Billing formulaparams.duration * 0.2
Each generated second$0.200000

When to Use

ScenarioWhy it fits
Choose AI Avatar Multi TextUse it when the job specifically calls for ai avatar multi text and the route’s documented input contract matches assets already available in your workflow.
Keep the request reproducibleUse the required image, prompt explicitly, then record optional choices alongside the prompt for repeatable reruns.
Build controlled creative variationsChange one declared control at a time when comparing framing, format, duration, resolution, style, or other route-supported output decisions.
Integrate asynchronous deliveryUse this endpoint when an application can submit work, track completion, and collect the returned media URL instead of requiring an immediate inline artifact.
Validate before a large batchRun representative prompts and source assets first, confirm the visual or audio behavior, then lock the successful request shape for scaled production.

Prompt Guide

Lead with the subject or transformation, then add the composition, environment, style, lighting, motion, voice, or fidelity details relevant to this route. Keep every parameter inside the documented schema; the example below uses only fields declared for sandbase-ai/ai-avatar-multi-text.

{
  "image": "https://static.sandbase.ai/examples/bytedance/omnihuman/1.5/input_image_1.png",
  "prompt": "Two kids talking on a lunch.",
  "voice1": "Sarah",
  "voice2": "Roger",
  "resolution": "480p"
}

Technical Specs

SpecificationDetails
Model IDsandbase-ai/ai-avatar-multi-text
Execution modeasync
Task typevideo
seedinteger · optional · min 0 · max 4294967295
imagestring · required
promptstring · required
voice1string · optional · choices: Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill
voice2string · optional · choices: Aria, Roger, Sarah, Laura, Charlie, George, Callum, River, Liam, Charlotte, Alice, Matilda, Will, Jessica, Eric, Chris, Brian, Daniel, Lily, Bill
resolutionstring · optional · choices: 480p, 720p
first_text_inputstring · optional
second_text_inputstring · optional

Related Models

More Models by sandbase-ai