SandBase is live — $1 in free credits on signupStart free ›

Lightricks modelsvideo generation api

lightricks/ltx-2.3-fast/image-to-video

Ltx 2.3 Fast is Lightricks's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Input
The prompt to generate the video from

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the image to generate the video from. Must be publicly accessible or base64 data URI. Supports PNG, JPEG, WebP, AVIF, and HEIF formats.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

The URL of the end image to use for the generated video. When provided, generates a transition video between start and end frames.
The resolution of the generated video Allowed values: 1080p, 1440p, 2160p.
The duration of the generated video in seconds. The fast model supports 6-20 seconds. Note: Durations longer than 10 seconds (12, 14, 16, 18, 20) are only supported with 25 FPS and 1080p resolution. Allowed values: 6, 8, 10, 12, 14, 16, 18, 20.
Whether to generate audio for the generated video
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
Idle

Example output — click Run to generate your own

API README

LTX 2.3 Fast Image to Video

LTX 2.3 Fast Image to Video is a image-to-video animation endpoint in the LTX 2.3 family. It is built for creators who need to turn a concrete creative brief into a controlled visual sequence: the request establishes the source material, the intended subject behavior, the camera language, and the atmosphere of the finished shot. The fast route keeps that workflow explicit instead of hiding its input assumptions behind a generic video-generation label.

In practice, this route accepts image, end image as creative context and exposes duration, resolution, aspect ratio, generate audio for delivery planning. That makes it suitable for shot-based pipelines where teams must preserve a source, direct a transformation, or control the final format without losing sight of the model's central task. Write prompts as a compact shot plan—subject, action, setting, camera, light, and timing—then use the structured fields for constraints that should remain deterministic across iterations.

Highlights

  • First-frame animation. Turns a supplied still into moving footage while the prompt specifies subject action, camera behavior, and atmosphere.
  • Synchronized audio path. Audio conditioning or generation is available in the same request, helping sound and picture share one creative plan.
  • Endpoint composition. An optional end frame can anchor where the shot lands, giving transitions a deliberate visual destination.
  • Resolution-aware delivery. Explicit output resolution lets the request target preview or finishing needs before generation.

Pricing

The request price is calculated with params.resolution == "2160p" ? params.duration * 0.24 : params.resolution == "1440p" ? params.duration * 0.12 : params.duration * 0.06.

ConfigurationPrice
duration=6, resolution=1080p$0.360000
duration=6, resolution=1440p$0.720000
duration=6, resolution=2160p$1.440000
duration=8, resolution=1080p$0.480000
duration=8, resolution=1440p$0.960000
duration=8, resolution=2160p$1.920000
duration=10, resolution=1080p$0.600000
duration=10, resolution=1440p$1.200000
duration=10, resolution=2160p$2.400000
duration=12, resolution=1080p$0.720000
duration=14, resolution=1080p$0.840000
duration=16, resolution=1080p$0.960000
duration=18, resolution=1080p$1.080000
duration=20, resolution=1080p$1.200000

The model card records a base price of $0.060000; the formula above determines usage-priced requests.

When to Use

ScenarioRecommendation
Choose this routeUse it when the deliverable specifically calls for image-to-video animation, rather than a neighboring generation mode.
Prepare the sourceProvide image in the format described by the request schema.
Direct the shotDescribe the subject, action, environment, camera movement, lighting, and temporal progression in that order.
Control continuityUse endpoint frames, reference media, strength, or audio controls when those fields are available instead of burying hard constraints in prose.
Plan deliverySet duration, frame count, resolution, and aspect ratio explicitly when the schema exposes them, then compare iterations with a stable seed where supported.

Prompt Guide

For image-to-video animation, describe one coherent shot rather than a list of visual keywords. Put the main subject and action first, follow with location and staging, then add camera movement, lens or framing, lighting, mood, and any timed change. Keep URLs and hard delivery choices in their dedicated fields.

{
  "prompt": "Snorkel lens ground-scraping tracking shot following a barefoot Ethiopian long-distance runner training on a dirt road at dawn, camera inches from the ground racing alongside, her feet kicking up red dust in slow motion at 240fps, Rift Valley landscape blurred in the background, the texture of earth and callused skin in hyper-detail, 40mm snorkel lens at ground height, Nike Running documentary meets Lubezki natural light, the poetry of human endurance",
  "image": "https://static.sandbase.ai/examples/lightricks/ltx-2.3-fast/image-to-video/input_image_0.png",
  "duration": 6,
  "resolution": "1080p",
  "aspect_ratio": "21:9",
  "generate_audio": true
}

Technical Specs

SpecificationValue
Model IDlightricks/ltx-2.3-fast/image-to-video
Required inputsprompt, image
ExecutionAsynchronous generation job
Request controls7 documented fields
Outputurl, content_type

Request fields

FieldType and constraints
imagestring; Required
promptstring; Required
durationinteger; Optional; Options: 6, 8, 10, 12, 14, 16, 18, 20; Default: 6
end_imagestring; Optional
resolutionstring; Optional; Options: 1080p, 1440p, 2160p; Default: "1080p"
aspect_ratiostring; Optional; Options: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16
generate_audioboolean; Optional; Default: true

Related Models

Related Models

lightricks/ltx-2.3-fast/text-to-videoLtx 2.3 Fast by Lightricks - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.lightricks/ltx-0.9.8-13b-distilled/image-to-videoLtx 0.9.8 13b Distilled by Lightricks - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.lightricks/ltx-2-19b/audio-to-video/loraLtx 2 19b Audio To Video Lora by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.lightricks/ltx-2-19b/distilled/extend-videoLtx 2 19b Distilled Extend Video by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.lightricks/ltx-2-19b/distilled/extend-video/loraLtx 2 19b Distilled Extend Video is Lightricks's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.lightricks/ltx-2-19b/distilled/image-to-videoLtx 2 19b Distilled by Lightricks - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.lightricks/ltx-2-19b/distilled/image-to-video/loraLtx 2 19b Distilled Lora is Lightricks's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.lightricks/ltx-2-19b/distilled/text-to-videoLtx 2 19b Distilled is Lightricks's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.