SandBase is live — $1 in free credits on signupStart free ›
Use in agentVEED models

VEED modelsvideo generation api

veed/avatars/audio-to-video

Avatars Audio To Video by VEED - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

Input
Format: complete URL.
The avatar to use for the video Allowed values: emily_vertical_primary, emily_vertical_secondary, marcus_vertical_primary, marcus_vertical_secondary, mira_vertical_primary, mira_vertical_secondary, jasmine_vertical_primary, jasmine_vertical_secondary, jasmine_vertical_walking, aisha_vertical_walking, elena_vertical_primary, elena_vertical_secondary, any_male_vertical_primary, any_female_vertical_primary, any_male_vertical_secondary, any_female_vertical_secondary, any_female_vertical_walking, emily_primary, emily_side, marcus_primary, marcus_side, aisha_walking, elena_primary, elena_side, any_male_primary, any_female_primary, any_male_side, any_female_side.
Idle

Example output — click Run to generate your own

API README

VEED Avatars Audio to Video

VEED Avatars Audio to Video handles the veed/avatars/audio-to-video workflow for audio to video work in video production. Teams can place this named operation inside review, asset preparation, and publishing pipelines while keeping the original request attached to every result. Avatars Audio To Video by VEED - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows. Because veed/avatars/audio-to-video is a separate catalog entry, evaluations should use this route’s own inputs, output expectations, and billing behavior. The endpoint is especially useful when a project has a defined transformation goal and needs a repeatable request contract instead of an ad-hoc desktop step. The route should therefore be evaluated on its own documented behavior.

For example, the route documentation includes “https://static.sandbase.ai/examples/veed/avatars/audio-to-video/input_audio_0.mp3”. The mandatory portion is no mandatory fields declared by the schema. avatar id controls the avatar to use for the video; avatar id accepts emily_vertical_primary, emily_vertical_secondary, marcus_vertical_primary, marcus_vertical_secondary, mira_vertical_primary, mira_vertical_secondary, jasmine_vertical_primary, jasmine_vertical_secondary, jasmine_vertical_walking, aisha_vertical_walking, elena_vertical_primary, elena_vertical_secondary, any_male_vertical_primary, any_female_vertical_primary, any_male_vertical_secondary, any_female_vertical_secondary, any_female_vertical_walking, emily_primary, emily_side, marcus_primary, marcus_side, aisha_walking, elena_primary, elena_side, any_male_primary, any_female_primary, any_male_side, any_female_side. VEED Avatars Audio to Video exposes audio, avatar_id as its documented request surface. Applications can submit that exact shape asynchronously, retain the chosen arguments in job metadata, and collect the resulting media URL for review or publishing.

Highlights

  • Portrait animation. VEED Avatars Audio to Video turns a supplied person image or avatar asset into a moving on-camera performance.
  • Speech-led facial motion. Audio or written dialogue drives mouth movement and expressive timing for a more convincing presenter result.
  • Recognizable identity. The workflow uses the source portrait as its anchor so the generated performance retains the intended subject.
  • Presenter-content workflow. It fits explainers, announcements, localized clips, and character-led media that do not require filming each delivery.

Pricing

Billing optionPrice
Per request$0.000575

When to Use

ScenarioWhy it fits
Choose VEED Avatars Audio to VideoUse it when the job specifically calls for audio to video and the route’s documented input contract matches assets already available in your workflow.
Keep the request reproducibleUse the required inputs explicitly, then record optional choices alongside the prompt for repeatable reruns.
Build controlled creative variationsChange one declared control at a time when comparing framing, format, duration, resolution, style, or other route-supported output decisions.
Integrate asynchronous deliveryUse this endpoint when an application can submit work, track completion, and collect the returned media URL instead of requiring an immediate inline artifact.
Validate before a large batchRun representative prompts and source assets first, confirm the visual or audio behavior, then lock the successful request shape for scaled production.

Prompt Guide

Lead with the subject or transformation, then add the composition, environment, style, lighting, motion, voice, or fidelity details relevant to this route. Keep every parameter inside the documented schema; the example below uses only fields declared for veed/avatars/audio-to-video.

{
  "avatar_id": "emily_vertical_primary"
}

Technical Specs

SpecificationDetails
Model IDveed/avatars/audio-to-video
Execution modeasync
Task typevideo
audiostring · optional
avatar_idstring · optional · choices: emily_vertical_primary, emily_vertical_secondary, marcus_vertical_primary, marcus_vertical_secondary, mira_vertical_primary, mira_vertical_secondary, jasmine_vertical_primary, jasmine_vertical_secondary, jasmine_vertical_walking, aisha_vertical_walking, elena_vertical_primary, elena_vertical_secondary, any_male_vertical_primary, any_female_vertical_primary, any_male_vertical_secondary, any_female_vertical_secondary, any_female_vertical_walking, emily_primary, emily_side, marcus_primary, marcus_side, aisha_walking, elena_primary, elena_side, any_male_primary, any_female_primary, any_male_side, any_female_side

Related Models

Related Models

veed/avatars/text-to-videoAvatars is VEED's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.veed/fabric-1.0/fast/image-to-videoFabric 1.0 Fast by VEED - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.veed/fabric-1.0/image-to-videoFabric 1.0 is VEED's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.veed/fabric-1.0/text-to-videoFabric 1.0 by VEED - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.veed/lipsyncLipsync is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.veed/lipsync/2.0Lipsync 2.0 is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.veed/video-bg-removalVideo Bg Removal by VEED - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.veed/video-bg-removal/fastVideo Bg Removal Fast is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.