Grok Imagine Video 1.5 Text to Video
xai/grok-imagine-video/1.5/text-to-videoGrok Imagine Video 1.5 by xAI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
- Base price
- $0.84USD / run
- Execution
- async
- Model type
- video
- Input fields
- 4
Try the model
Playground
Your output will appear here
Complete the inputs, then click Run.
Specifications
Pricing
- Base price
- $0.84 / run
- Billing formula
- (params.duration ?? 6) * ((params.resolution ?? "720p") == "1080p" ? 0.25 : (params.resolution ?? "720p") == "480p" ? 0.08 : 0.14)
Context & modalities
- Input
- Schema-defined
- Output
- video
Capabilities
- Chat
- Not supported
- Vision
- Not supported
- Reasoning
- Not supported
- Structured output
- Not supported
- Function calling
- Not supported
- Audio input
- Not supported
Access
- Provider
- xAI
- Model ID
- xai/grok-imagine-video/1.5/text-to-video
- Execution
- async
- API
- Unified Run API
- Endpoint
- /v1/run
API README
Grok Imagine Video 1.5 Text To Video
Grok Imagine Video 1.5 Text To Video is a specialized SandBase endpoint built to generate a moving scene directly from written direction. It is most useful when teams need prompt-led shot creation with controllable duration, framing, and resolution, with the route’s inputs defining a repeatable production contract instead of leaving critical delivery choices to an ad-hoc manual workflow.
Use this exact route when its input mode matches the creative asset already in hand: describe the visual or spoken result clearly, set only the controls that support the intended delivery, and keep the subject, action, environment, camera or performance direction internally consistent. The result is returned asynchronously as a downloadable media URL suitable for review, automation, or downstream finishing.
Highlights
- Prompt-built shots. Create subject, setting, action, lens behavior, lighting, and mood from text when a source frame would constrain exploration.
- Temporal direction. Duration control gives a described action or camera beat enough time to begin, develop, and resolve in one generated clip.
- Multiple delivery frames. Wide, standard, square, portrait, and vertical ratios let composition begin in the format where the video will be watched.
- Resolution-aware output. Select the available resolution tier deliberately for fast iteration, general delivery, or higher-detail finishing.
Pricing
| Duration | Resolution | Price |
|---|---|---|
| 1 sec | 480p | $0.08 |
| 1 sec | 720p | $0.14 |
| 1 sec | 1080p | $0.25 |
| 2 sec | 480p | $0.16 |
| 2 sec | 720p | $0.28 |
| 2 sec | 1080p | $0.50 |
| 3 sec | 480p | $0.24 |
| 3 sec | 720p | $0.42 |
| 3 sec | 1080p | $0.75 |
| 4 sec | 480p | $0.32 |
| 4 sec | 720p | $0.56 |
| 4 sec | 1080p | $1.00 |
| 5 sec | 480p | $0.40 |
| 5 sec | 720p | $0.70 |
| 5 sec | 1080p | $1.25 |
| 6 sec | 480p | $0.48 |
| 6 sec | 720p | $0.84 |
| 6 sec | 1080p | $1.50 |
| 7 sec | 480p | $0.56 |
| 7 sec | 720p | $0.98 |
| 7 sec | 1080p | $1.75 |
| 8 sec | 480p | $0.64 |
| 8 sec | 720p | $1.12 |
| 8 sec | 1080p | $2.00 |
| 9 sec | 480p | $0.72 |
| 9 sec | 720p | $1.26 |
| 9 sec | 1080p | $2.25 |
| 10 sec | 480p | $0.80 |
| 10 sec | 720p | $1.40 |
| 10 sec | 1080p | $2.50 |
| 11 sec | 480p | $0.88 |
| 11 sec | 720p | $1.54 |
| 11 sec | 1080p | $2.75 |
| 12 sec | 480p | $0.96 |
| 12 sec | 720p | $1.68 |
| 12 sec | 1080p | $3.00 |
| 13 sec | 480p | $1.04 |
| 13 sec | 720p | $1.82 |
| 13 sec | 1080p | $3.25 |
| 14 sec | 480p | $1.12 |
| 14 sec | 720p | $1.96 |
| 14 sec | 1080p | $3.50 |
| 15 sec | 480p | $1.20 |
| 15 sec | 720p | $2.10 |
| 15 sec | 1080p | $3.75 |
Billed per second of requested output duration at xAI list price.
When to Use
| Scenario | Why this route fits |
|---|---|
| Concept development | Choose Grok Imagine Video 1.5 Text To Video when its text to video workflow matches the starting material and you need several clearly directed variations. |
| Production iteration | Use explicit duration, resolution, ratio, seed, or quality controls to compare versions without changing the core creative brief. |
| Channel adaptation | Generate directly in the landscape, square, portrait, or vertical format required by the destination whenever that control is available. |
| Automated pipelines | Integrate the asynchronous media URL into review queues, asset libraries, publishing tools, or a later finishing stage. |
| Alternative route | Pick a related text-, image-, reference-, edit-, or turbo route when the available source media or required degree of control is different. |
Prompt Guide
Lead with the main subject or source asset, then describe the intended action or transformation, environment, composition, camera or vocal delivery, lighting and mood. Keep instructions concrete and compatible; use the route’s explicit fields for duration, resolution, ratio, quality, voice, or reproducibility instead of burying those settings in prose.
{
"prompt": "A cinematic product reveal with deliberate subject motion, coherent lighting, and a slow camera push.",
"duration": 6,
"resolution": "720p",
"aspect_ratio": "21:9"
}
Technical Specs
| Property | Details |
|---|---|
| Model ID | xai/grok-imagine-video/1.5/text-to-video |
| Required inputs | prompt |
| Execution | Asynchronous; poll the returned generation ID |
| Output | Downloadable media URL |
prompt | string; required |
aspect_ratio | string; optional; choices: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16 |
resolution | string; optional; choices: 480p, 720p, 1080p; default: 720p |
duration | integer; optional; range: 1–15; default: 6 |
Related Models
| Model | Best for |
|---|---|
xai/grok-tts | Alternative grok tts workflow |
xai/grok-imagine-video/text-to-video | Alternative text to video workflow |
xai/grok-imagine-video/extend | Alternative extend workflow |
xai/grok-imagine-video/image-to-video | Alternative image to video workflow |
Start building
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
-X POST "https://api.sandbase.ai/v1/run" \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
-H "Content-Type: application/json" \
--data-binary @- <<'SANDBASE_JSON'
{
"model": "xai/grok-imagine-video/1.5/text-to-video",
"prompt": "Anime schoolgirl bursting out of house door, cherry blossoms blowing, morning light, speed lines indicating rush, chibi-ready expressions, classic shojo aesthetic, vibrant colors",
"duration": 6,
"resolution": "720p"
}
SANDBASE_JSON
)
run_id=$(printf '%s' "$result" | jq -r .id)
for attempt in $(seq 1 120); do
status=$(printf '%s' "$result" | jq -r .status)
case "$status" in completed|failed|timeout) break ;; esac
sleep 2
result=$(curl --fail-with-body --silent \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
"https://api.sandbase.ai/v1/run/$run_id")
done
status=$(printf '%s' "$result" | jq -r .status)
[ "$status" = completed ] || { echo "Generation ended: $status" >&2; exit 1; }
printf '%s\n' "$result"Choose your model
Compare models
| Model | Context | Input / 1M | Output / 1M | Released |
|---|---|---|---|---|
Grok Imagine Video 1.5 Text to VideoThis model xAI | — | — | — | Aug 1, 2026 |
xAI | — | — | — | Aug 1, 2026 |
xAI | — | — | — | May 31, 2026 |
xAI | — | — | — | Mar 24, 2026 |
xAI | — | — | — | Mar 24, 2026 |
xAI | — | — | — | Jan 29, 2026 |
