API reference · meituan

meituan/longcat-single-avatar/audio-to-video

Integrate this model through SandBase's unified API, with production-ready schemas and examples.

IMAGEAsyncOpen model
Production endpoint

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

POSThttps://api.sandbase.ai/v1/run
Model IDmeituan/longcat-single-avatar/audio-to-video
01

Input Schema

8 parameters · 1 required · 7 optional

ParameterTypeRequiredDescription
promptstringRequiredThe prompt to guide the video generation. · Default: "A person is talking naturally with natural expressions and movements."
seedintegerOptionalThe seed for the random number generator.
audiostringOptionalThe URL of the audio file to drive the avatar.
resolutionstringOptionalResolution of the generated video (480p or 720p). Billing is per video-second (16 frames): 480p is 1 unit per second and 720p is 4 units per second. · Options: 480p, 720p · Default: "480p"
480p720p
num_segmentsintegerOptionalNumber of video segments to generate. Each segment adds ~5 seconds of video. First segment is ~5.8s, additional segments are 5s each. · Min: 1 · Max: 10 · Default: 1
num_inference_stepsintegerOptionalThe number of inference steps to use. · Min: 10 · Max: 100 · Default: 50
text_guidance_scalenumberOptionalThe text guidance scale for classifier-free guidance. · Min: 1 · Max: 10 · Default: 4
audio_guidance_scalenumberOptionalThe audio guidance scale. Higher values may lead to exaggerated mouth movements. · Min: 1 · Max: 10 · Default: 4
02

Output Schema

FieldTypeDescription
idstringUnique identifier for the generation task
statusstringTask status: pending, running, completed, failed, timeout
modelstringModel used for the generation
outputsarrayArray of output items
outputs[].urlstringURL of the generated artifact
outputs[].content_typestringMIME type (e.g. image/png, video/mp4)
errorobject | nullError details if failed, null on success
error.typestringMachine-readable error type code
error.messagestringHuman-readable error description

Async Workflow

This model uses asynchronous execution. Submit a request and poll for the result.

  1. Submit — POST to /v1/run, receive an id
  2. Poll — GET /v1/run/{id} until status is completed, failed, or timeout
  3. Retrieve — Read outputs from the completed response
03

Code Examples

Ready-to-run snippets

# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
  "model": "meituan/longcat-single-avatar/audio-to-video",
  "audio": "https://static.sandbase.ai/examples/meituan/longcat-single-avatar/audio-to-video/input_audio_0.mp3",
  "prompt": "A person is talking naturally with natural expressions and movements.",
  "resolution": "480p",
  "num_segments": 1,
  "num_inference_steps": 50,
  "text_guidance_scale": 4,
  "audio_guidance_scale": 4
}'

# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
  -H "Authorization: Bearer YOUR_API_KEY"