API reference · meituan
meituan/longcat-single-avatar/audio-to-video
Integrate this model through SandBase's unified API, with production-ready schemas and examples.
Production endpoint
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
POST
https://api.sandbase.ai/v1/runModel ID
meituan/longcat-single-avatar/audio-to-video01
Input Schema
8 parameters · 1 required · 7 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
prompt | string | Required | The prompt to guide the video generation. · Default: "A person is talking naturally with natural expressions and movements." |
seed | integer | Optional | The seed for the random number generator. |
audio | string | Optional | The URL of the audio file to drive the avatar. |
resolution | string | Optional | Resolution of the generated video (480p or 720p). Billing is per video-second (16 frames): 480p is 1 unit per second and 720p is 4 units per second. · Options: 480p, 720p · Default: "480p" 480p720p |
num_segments | integer | Optional | Number of video segments to generate. Each segment adds ~5 seconds of video. First segment is ~5.8s, additional segments are 5s each. · Min: 1 · Max: 10 · Default: 1 |
num_inference_steps | integer | Optional | The number of inference steps to use. · Min: 10 · Max: 100 · Default: 50 |
text_guidance_scale | number | Optional | The text guidance scale for classifier-free guidance. · Min: 1 · Max: 10 · Default: 4 |
audio_guidance_scale | number | Optional | The audio guidance scale. Higher values may lead to exaggerated mouth movements. · Min: 1 · Max: 10 · Default: 4 |
02
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the generation task |
status | string | Task status: pending, running, completed, failed, timeout |
model | string | Model used for the generation |
outputs | array | Array of output items |
outputs[].url | string | URL of the generated artifact |
outputs[].content_type | string | MIME type (e.g. image/png, video/mp4) |
error | object | null | Error details if failed, null on success |
error.type | string | Machine-readable error type code |
error.message | string | Human-readable error description |
Async Workflow
This model uses asynchronous execution. Submit a request and poll for the result.
- Submit — POST to /v1/run, receive an
id - Poll — GET /v1/run/{id} until status is
completed,failed, ortimeout - Retrieve — Read
outputsfrom the completed response
03
Code Examples
Ready-to-run snippets
# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "meituan/longcat-single-avatar/audio-to-video",
"audio": "https://static.sandbase.ai/examples/meituan/longcat-single-avatar/audio-to-video/input_audio_0.mp3",
"prompt": "A person is talking naturally with natural expressions and movements.",
"resolution": "480p",
"num_segments": 1,
"num_inference_steps": 50,
"text_guidance_scale": 4,
"audio_guidance_scale": 4
}'
# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
-H "Authorization: Bearer YOUR_API_KEY"
