Nemotron 3 Nano Omni
/v1/runNemotron 3 Nano Omni Video by NVIDIA - advanced AI model for video-to-text. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
Request body
Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.
Model identifier. Set to nvidia/nemotron-3-nano-omni/video.
Default: nvidia/nemotron-3-nano-omni/video
Text prompt to send to the model. English only.
URL of the video to reason about. mp4, up to 1080p, max 2 minutes.
Maximum number of tokens to generate.
Range: 1 to 20000
Default: 1024
Nucleus sampling probability mass.
Range: 0 to 1
Default: 0.95
Sampling temperature. Lower is more deterministic.
Range: 0 to 2
Default: 0.7
Optional system prompt to steer the model. Reasoning behavior is controlled by the separate `reasoning_mode` field.
Whether the model should emit an explicit reasoning trace. `no_think` returns a direct answer; `think` returns chain-of-thought followed by the final answer.
Allowed values: think, no_think
Default: no_think
Response Schema
The submit endpoint returns a run response. If its status is pending or running, poll GET /v1/run/{id} with the returned opaque ID until it reaches a terminal state.
Opaque SandBase run identifier. Use it exactly as returned; no prefix is guaranteed.
Current public run status.
Allowed values: pending, running, completed, failed, timeout
Public SandBase model name used for this run.
Present only for completed runs. Each object is capability-specific; inspect the selected model schema for its fields.
Present only for failed or timeout runs. Contains a public error type and sanitized message.
Stable public error category.
Sanitized error message safe to show to clients.
Usage details when available.
Model capabilities
Capabilities declared by the model registry.
Default: video-to-text
Execution mode declared by the model registry.
Default: async