lightricks/ltx-2.3-pro/text-to-video
Ltx 2.3 Pro is Lightricks's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
Example output — click Run to generate your own
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
https://api.sandbase.ai/v1/runlightricks/ltx-2.3-pro/text-to-videoInput Schema
5 parameters · 1 required · 4 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
prompt | string | Required | The prompt to use for the generated video · Min length: 1 · Max length: 5000 |
duration | integer | Optional | The duration of the generated video in seconds · Options: 6, 8, 10 · Default: 6 6810 |
resolution | string | Optional | The resolution of the generated video · Options: 1080p, 1440p, 2160p · Default: "1080p" 1080p1440p2160p |
aspect_ratio | string | Optional | The aspect ratio of the generated image. · Options: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16 21:916:93:24:35:41:14:53:42:39:16 |
generate_audio | boolean | Optional | Whether to generate audio for the generated video · Default: true |
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the generation task |
status | string | Task status: pending, running, completed, failed, timeout |
model | string | Model used for the generation |
outputs | array | Array of output items |
outputs[].url | string | URL of the generated artifact |
outputs[].content_type | string | MIME type (e.g. image/png, video/mp4) |
error | object | null | Error details if failed, null on success |
error.type | string | Machine-readable error type code |
error.message | string | Human-readable error description |
Async Workflow
This model uses asynchronous execution. Submit a request and poll for the result.
- Submit — POST to /v1/run, receive an
id - Poll — GET /v1/run/{id} until status is
completed,failed, ortimeout - Retrieve — Read
outputsfrom the completed response
Code Examples
Ready-to-run snippets
# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "lightricks/ltx-2.3-pro/text-to-video",
"prompt": "Through-the-veil shot of a bride's face during an Indian wedding ceremony, camera positioned behind the sheer red dupatta fabric, the embroidered pattern creating a textured overlay on her face, her eyes lined with kohl looking down at henna-covered hands, marigold garlands in soft background bokeh, 85mm f/1.2 focused through the fabric layer, warm tungsten and candlelight, Mira Nair Monsoon Wedding intimacy",
"duration": 6,
"resolution": "1080p",
"generate_audio": true
}'
# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
-H "Authorization: Bearer YOUR_API_KEY"API README
LTX 2.3 Pro Text to Video
LTX 2.3 Pro Text to Video is a LTX route built for text-to-video generation. It turns a shot description with subject, action, camera, and atmosphere into a coherent video interpretation of the written direction, giving creators a focused endpoint instead of forcing one generic workflow across materially different production tasks. The route belongs to a family known for audio-visual timing, cinematic motion, and shot-level controllability, so it is best evaluated as a creative system for intentional shots and assets rather than as a one-click novelty generator.
Use this endpoint when the input contract and deliverable match that job exactly. Its documented controls include duration choices (6, 8, 10); resolution choices (1080p, 1440p, 2160p); aspect_ratio choices (21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16); generate_audio (Whether to generate audio for the generated video). Together, these controls help teams plan predictable iterations, compare outputs under stable settings, and connect generation to iterative film, advertising, and social-video workflows without hiding the operational choices that shape the result.
Highlights
- Purpose-built Text-to-video generation. The route accepts a shot description with subject, action, camera, and atmosphere and produces a coherent video interpretation of the written direction; its interface is scoped to that transformation, keeping source assets and creative intent explicit.
- Creative direction. Prompts can describe subject behavior, composition, camera intent, lighting, material, atmosphere, and temporal progression so the result is driven by a shot plan rather than isolated keywords.
- Route-specific control. The request exposes duration choices (6, 8, 10); resolution choices (1080p, 1440p, 2160p); aspect_ratio choices (21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16); generate_audio (Whether to generate audio for the generated video), allowing the same concept to be tested systematically while preserving a repeatable production setup.
- Pipeline-ready output. The generated media asset is returned through the documented asynchronous output contract, which suits review queues, batch iteration, and downstream automation. Editors can review pacing, continuity, lens language, choreography, transitions, temporal artifacts, soundtrack alignment, color response, delivery framing, and cut compatibility before approval.
Pricing
| Configuration | Price |
|---|---|
| Billing rule | params.resolution == "2160p" ? params.duration * 0.32 : params.resolution == "1440p" ? params.duration * 0.16 : params.duration * 0.08 |
| duration=6, resolution=1080p | $0.480000 |
| duration=6, resolution=1440p | $0.960000 |
| duration=6, resolution=2160p | $1.920000 |
| duration=8, resolution=1080p | $0.640000 |
| duration=8, resolution=1440p | $1.280000 |
| duration=8, resolution=2160p | $2.560000 |
| duration=10, resolution=1080p | $0.800000 |
| duration=10, resolution=1440p | $1.600000 |
| duration=10, resolution=2160p | $3.200000 |
When to Use
| Scenario | Why this model fits |
|---|---|
| Create the exact route output | Choose it when you need text-to-video generation and already have a shot description with subject, action, camera, and atmosphere. |
| Develop controlled variations | Keep the main brief fixed while changing one documented setting at a time to compare motion, framing, quality, or asset behavior. |
| Build repeatable batches | Use a consistent request shape for catalog, campaign, storyboard, game-asset, or social-content production. |
| Preserve source intent | Prefer this route when the supplied reference material must remain the foundation of a coherent video interpretation of the written direction. |
| Connect a media pipeline | Use asynchronous results in an automated review, approval, post-production, or asset-management workflow. |
Prompt Guide
Start with the desired result, then describe the source relationship, subject action, composition or camera behavior, lighting, style, and timing. For text-to-video generation, state what must remain stable as clearly as what should change. Use only fields exposed by the schema; the example below is structurally valid for this route.
{
"prompt": "Through-the-veil shot of a bride's face during an Indian wedding ceremony, camera positioned behind the sheer red dupatta fabric, the embroidered pattern creating a textured overlay on her face, her eyes lined with kohl looking down at henna-covered hands, marigold garlands in soft background bokeh, 85mm f/1.2 focused through the fabric layer, warm tungsten and candlelight, Mira Nair Monsoon Wedding intimacy",
"duration": 6,
"resolution": "1080p",
"aspect_ratio": "21:9",
"generate_audio": true
}
Technical Specs
| Specification | Value |
|---|---|
| Model ID | lightricks/ltx-2.3-pro/text-to-video |
| Workflow | Text-to-video generation |
| Required inputs | prompt |
prompt | string |
duration | integer; options: 6, 8, 10 |
resolution | string; options: 1080p, 1440p, 2160p |
aspect_ratio | string; options: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16 |
generate_audio | boolean |

