API reference · NVIDIA

nvidia/cosmos-3-super/image-to-video

Integrate this model through SandBase's unified API, with production-ready schemas and examples.

VIDEOAsyncOpen model
Production endpoint

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

POSThttps://api.sandbase.ai/v1/run
Model IDnvidia/cosmos-3-super/image-to-video
01

Input Schema

12 parameters · 2 required · 10 optional

ParameterTypeRequiredDescription
imagestringRequiredURL of the conditioning first-frame image for the video.
promptstringRequiredText prompt describing the motion and scene of the video to generate. · Min length: 1 · Max length: 4096
seedintegerOptionalThe same seed and prompt given to the same model version will produce the same video every time.
num_framesintegerOptionalNumber of frames to generate. More frames yield a longer video. · Min: 5 · Max: 189 · Default: 189
aspect_ratiostringOptionalThe aspect ratio of the generated image. · Options: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16
21:916:93:24:35:41:14:53:42:39:16
guidance_scalenumberOptionalClassifier-free guidance scale. Higher values increase prompt adherence at the cost of diversity. · Min: 0 · Max: 20 · Default: 6
frames_per_secondintegerOptionalFrames per second of the output video. · Min: 4 · Max: 60 · Default: 24
agentic_early_stopbooleanOptionalStop the agentic loop early when the critic score clears the strict quality threshold. · Default: true
num_inference_stepsintegerOptionalNumber of denoising steps. More steps yield higher quality but take longer. · Min: 1 · Max: 50 · Default: 28
agentic_max_iterationsintegerOptionalMaximum agentic prompt stages when agentic generation is enabled. · Min: 1 · Max: 3 · Default: 2
enable_agentic_generationbooleanOptionalEnable the iterative Cosmos agentic loop: prompt upsampling, candidate video generation, VLM critique of sampled frames, and prompt rewrite. Each candidate is a full render, so this is substantially slower and costlier than a single generation. · Default: false
agentic_samples_per_iterationintegerOptionalCandidate videos to generate and judge per agentic iteration. The best candidate advances to the next rewrite stage. · Min: 1 · Max: 3 · Default: 2
02

Output Schema

FieldTypeDescription
idstringUnique identifier for the generation task
statusstringTask status: pending, running, completed, failed, timeout
modelstringModel used for the generation
outputsarrayArray of output items
outputs[].urlstringURL of the generated artifact
outputs[].content_typestringMIME type (e.g. image/png, video/mp4)
errorobject | nullError details if failed, null on success
error.typestringMachine-readable error type code
error.messagestringHuman-readable error description

Async Workflow

This model uses asynchronous execution. Submit a request and poll for the result.

  1. Submit — POST to /v1/run, receive an id
  2. Poll — GET /v1/run/{id} until status is completed, failed, or timeout
  3. Retrieve — Read outputs from the completed response
03

Code Examples

Ready-to-run snippets

# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
  "model": "nvidia/cosmos-3-super/image-to-video",
  "image": "https://storage.googleapis.com/falserverless/example_inputs/hunyuan_i2v.jpg",
  "prompt": "The camera slowly pushes in as the subject turns their head toward the light, hair drifting in a gentle breeze, dust motes floating through warm afternoon sun.",
  "num_frames": 189,
  "guidance_scale": 6,
  "frames_per_second": 24,
  "agentic_early_stop": true,
  "num_inference_steps": 28,
  "agentic_max_iterations": 2,
  "enable_agentic_generation": false,
  "agentic_samples_per_iteration": 2
}'

# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
  -H "Authorization: Bearer YOUR_API_KEY"