bytedance/omnihuman/1.0
Omnihuman 1.0 is Bytedance's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
PNG, JPEG, WebP, or GIF · 20 MiB maximum
Example output — click Run to generate your own
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
https://api.sandbase.ai/v1/runbytedance/omnihuman/1.0Input Schema
2 parameters · 1 required · 1 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
image | string | Required | The URL of the image used to generate the video |
audio | string | Optional | The URL of the audio file to generate the video. Audio must be under 30s long. |
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the generation task |
status | string | Task status: pending, running, completed, failed, timeout |
model | string | Model used for the generation |
outputs | array | Array of output items |
outputs[].url | string | URL of the generated artifact |
outputs[].content_type | string | MIME type (e.g. image/png, video/mp4) |
error | object | null | Error details if failed, null on success |
error.type | string | Machine-readable error type code |
error.message | string | Human-readable error description |
Async Workflow
This model uses asynchronous execution. Submit a request and poll for the result.
- Submit — POST to /v1/run, receive an
id - Poll — GET /v1/run/{id} until status is
completed,failed, ortimeout - Retrieve — Read
outputsfrom the completed response
Code Examples
Ready-to-run snippets
# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "bytedance/omnihuman/1.0",
"audio": "https://static.sandbase.ai/examples/bytedance/omnihuman/1.0/input_audio_1.mp3",
"image": "https://static.sandbase.ai/examples/bytedance/omnihuman/1.0/input_image_0.png",
"prompt": "a beautiful sunset over mountains"
}'
# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
-H "Authorization: Bearer YOUR_API_KEY"API README
Bytedance Omnihuman 1.0
Bytedance Omnihuman 1.0 is built for audio-driven human animation from a reference portrait or character image. It gives creative and production teams a focused way to move from an approved brief or source asset to a reviewable result without fragmenting the job across unrelated tools. The model is most valuable when visual intent, brand suitability, and downstream usability all matter, because its output can enter an editorial, campaign, product, or content pipeline as a purposeful asset rather than an isolated experiment.
In practice, teams can use Bytedance Omnihuman 1.0 during a structured cycle of briefing, generation, comparison, and refinement. Establish the subject, audience, visual objective, and acceptance criteria first; prepare any reference media at suitable quality; then evaluate alternatives for composition, continuity, realism, and communication value before delivery. This workflow keeps creative judgment central while making repeated production easier to review, reproduce, and scale for the specific bytedance/omnihuman/1.0 task.
Highlights
Audio-synchronized performance. Audio-synchronized performance gives bytedance › omnihuman › 1.0 a recognizable technical advantage: reviewers can assess this property directly in the generated asset instead of inferring it from request mechanics.
Identity-preserving animation. bytedance › omnihuman › 1.0 applies identity-preserving animation to the visual or temporal result itself, helping artists make a meaningful quality decision during selection and refinement.
Natural facial expression. For bytedance › omnihuman › 1.0, natural facial expression supports coherent assets across the intended creative workflow and distinguishes this capability from a simple format or delivery option.
Coordinated body motion. The practical value of coordinated body motion is visible in the finished media from bytedance › omnihuman › 1.0, where it supports repeatable art direction rather than merely exposing another request setting.
Pricing
| Billing basis | Rate |
|---|---|
| Audio-driven output runtime | $0.140000 per second |
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| Use it when audio-driven human animation from a reference portrait or character image is the central production goal | Choose a different model when the required media task is fundamentally different |
| The team needs several reviewable creative alternatives | Exact deterministic reproduction is mandatory |
| Visual quality and practical downstream use both matter | Editable source layers or native project files are required |
| A managed generation step fits the delivery workflow | A live frame-by-frame interactive editor is essential |
| The documented inputs cover the available source assets | Required source media or controls fall outside the documented fields |
Prompt Guide
State the intended result first, then describe the subject, environment, action, visual treatment, and delivery constraints. Keep instructions concrete, avoid conflicting directions, and change one creative variable at a time when comparing outputs. For source-driven work, describe what should remain recognizable as clearly as what should change.
{
"audio": "https://static.sandbase.ai/examples/bytedance/omnihuman/1.0/input_audio_1.mp3",
"image": "https://static.sandbase.ai/examples/bytedance/omnihuman/1.0/input_image_0.png"
}
Technical Specs
| Property | Details |
|---|---|
audio | Type / options: string<br>Required: No<br>Description: The URL of the audio file to generate the video. Audio must be under 30s long. |
image | Type / options: string<br>Required: Yes<br>Description: The URL of the image used to generate the video |
Related Models
bytedance/dreamactor/2.0bytedance/lynxbytedance/omnihuman/1.5

