SandBase is live — $1 in free credits on signupStart free ›

PixVerse modelsvideo generation api

pixverse/v5/image-to-video

V5 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Input

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the image to use as the first frame
The resolution of the generated video Allowed values: 360p, 540p, 720p, 1080p.
The duration of the generated video in seconds. 8s videos cost double. 1080p videos are limited to 5 seconds Allowed values: 5, 8.
The style of the generated video Allowed values: anime, 3d_animation, clay, comic, cyberpunk.
The same seed and the same prompt given to the same version of the model will output the same video every time.
Idle

Example output — click Run to generate your own

API README

PixVerse V5 Image To Video

PixVerse V5 Image To Video is the PixVerse V5 route for image-to-video generation, turning a supplied still image into a production-ready result while keeping the operation distinct from neighboring endpoints. V5 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output. The workflow is designed for creators who need the model’s specific transformation to remain visible in the request, so the source, intended change, and finished artifact can be reviewed as one coherent creative decision.

In practical use, this route exposes seed, image, style, prompt, duration, resolution to shape the exact deliverable. Those controls let a team preserve the important source constraints, state subject behavior or material treatment precisely, choose supported timing or output characteristics, and reproduce successful settings across alternate takes. The result fits an iterative pipeline: establish the core brief, compare controlled variations, then pass the selected asset into editorial, design, localization, visualization, or publishing work.

Highlights

  • Source-frame animation. Begin from the supplied image and direct how its subjects, camera, and environment evolve through the clip. This capability belongs to the exact pixverse/v5/image-to-video workflow.
  • Directed motion. Translate the request into subject movement and camera behavior while maintaining temporal continuity.
  • Format control. Choose the documented duration and resolution options for the target placement.
  • Repeatable variation. Use seed and route controls to explore alternates without changing the core shot brief.

Pricing

ConfigurationPrice
Per request$0.150000

When to Use

ScenarioWhy it fits
Exact workflow fitChoose this route when the required deliverable is image-to-video generation, rather than a related route with different source media.
Directed creative iterationUse it when subject, motion, material, speech, framing, or finish should be expressed explicitly and compared across controlled variants.
Existing-asset continuityUse it when supplied images, video, audio, references, or styles must remain the anchor for the generated result.
Repeatable productionUse it when successful inputs need to be saved and rerun across a campaign, asset set, localization pass, or batch.
Pipeline handoffUse it when the returned artifact will move into editorial, compositing, visualization, review, storage, or publishing.

Prompt Guide

Start with the source or subject, then describe the intended transformation, movement or behavior, camera and composition, and the desired finish. Keep media URLs reachable, use only fields documented for this exact route, and change one major control at a time when comparing results. For source-led tasks, describe what should change as well as what must remain recognizable.

{
  "prompt": "A woman warrior with her hammer walking with his glacier wolf.",
  "image": "https://static.sandbase.ai/examples/pixverse/v5/image-to-video/output_url_0.png",
  "duration": 5,
  "resolution": "720p"
}

Technical Specs

PropertyValue
Model IDpixverse/v5/image-to-video
Execution modeasync
Required inputsprompt, image
seedinteger
imagestring
stylestring; options: anime, 3d_animation, clay, comic, cyberpunk
promptstring
durationinteger; options: 5, 8; default: 5
resolutionstring; options: 360p, 540p, 720p, 1080p; default: 720p

Related Models

  • pixverse/v5/effects
  • pixverse/v5/text-to-video
  • pixverse/v5/transition
  • pixverse/c1/image-to-video

Related Models

pixverse/v5/effectsV5 Effects by PixVerse - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.pixverse/v5/text-to-videoV5 by PixVerse - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.pixverse/v5/transitionV5 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.pixverse/c1/transitionC1 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.pixverse/extendExtend by PixVerse - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.pixverse/extend/fastExtend Fast is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.pixverse/lipsyncLipsync is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.pixverse/sound-effectsSound Effects is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.