Skip to content

Video Generation ​

SandBase currently publishes API reference pages for 377 enabled video generation models across 22 providers. Choose a provider in the left navigation, then open a model page for its exact API identifier, supported capabilities, and a working request.

Video Generation models use the async SandBase generation protocol declared in each model registry file. Submit a request, receive a task id, then poll the result endpoint until the generation is completed, failed, or timed out.

Providers ​

OpenAI ​

  • Sora 2 Characters — Sora 2 Characters is OpenAI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Sora 2 Remix — Sora 2 Video To Video Remix is OpenAI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Sora 2 Image to Video Pro — Sora 2 Pro is OpenAI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Sora 2 Text to Video Pro — Sora 2 Pro by OpenAI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Sora 2 — Sora 2 by OpenAI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Sora 2 — Sora 2 is OpenAI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Bytedance ​

  • Seedance 2.5 Image to Video — ByteDance's next-generation image-to-video model, animating a single still into a native clip up to 30 seconds at 720p with continuous, coherent motion, native audio, and director-level camera control.
  • Seedance 2.5 Reference to Video — ByteDance's next-generation reference-to-video model, generating video from multimodal references (images, videos, audio) and locking a character, set, and palette across a full take up to 30 seconds for production-grade consistency.
  • Seedance 2.5 Text to Video — ByteDance's next-generation text-to-video model, generating native single-shot clips up to 30 seconds at 720p with coherent motion, native audio, and director-level camera control for professional-grade video creation.
  • Seedance 2.0 Text to Video — ByteDance's most advanced text-to-video model delivering cinematic output, native audio, multi-shot editing, and director-level camera control for professional-grade video creation.
  • Seedance 2.0 Reference to Video — ByteDance's most advanced reference-to-video model generating cinematic video guided by reference content, with native audio, multi-shot editing, and director-level camera control for professional-grade video creation.
  • Seedance 2.0 Fast Text to Video — ByteDance's most advanced text-to-video model in its fast tier delivering lower latency and cost without compromising on cinematic output, native audio, multi-shot editing, and director-level camera control for professional-grade video creation.
  • Seedance 2.0 Image to Video — ByteDance's most advanced image-to-video model transforming still images into cinematic video with native audio, multi-shot editing, and director-level camera control for professional-grade video creation.
  • Seedance 2.0 Fast Reference to Video — ByteDance's most advanced reference-to-video model in its fast tier delivering lower latency and cost without compromising on cinematic output, native audio, multi-shot editing, and director-level camera control for professional-grade video creation.
  • Seedance 2.0 Fast Image to Video — ByteDance's most advanced image-to-video model in its fast tier delivering lower latency and cost without compromising on cinematic output, native audio, multi-shot editing, and director-level camera control for professional-grade video creation.
  • DreamActor 2.0 — DreamActor M2.0 by ByteDance generates videos by animating a reference image using motion from a driving video. It replicates motion, facial expressions, and lip movements from the template video while preserving the subject and background features of the input image.
  • Seedance v1.5 Pro Text to Video — ByteDance Seedance v1.5 Pro text-to-video model generating cinematic video from text prompts with native audio generation, camera control, and professional-grade output quality.
  • Seedance v1.5 Pro Image to Video — ByteDance Seedance v1.5 Pro image-to-video model transforming still images into cinematic video with native audio generation, camera control, and professional-grade output quality.
  • …and 9 more models in the sidebar.

Google ​

  • Gemini Omni Flash — Gemini Omni Flash by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Gemini Omni Flash — Gemini Omni Flash by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Gemini Omni Flash — Gemini Omni Flash Edit by Google - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Gemini Omni Flash — Gemini Omni Flash is Google's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Veo 3.1 Fast — Veo3.1 Fast by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Veo3.1 Lite FLF — Veo3.1 Lite First Last Frame To Video is Google's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Veo3.1 Lite Image to Video — Veo3.1 Lite by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Veo3.1 Lite Text to Video — Veo3.1 Lite is Google's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Veo 3.1 Fast — Veo3.1 Fast Extend Video by Google - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Veo 3.1 — Veo3.1 Extend Video is Google's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Veo 3.1 Fast — Veo3.1 Fast First Last Frame To Video is Google's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Veo 3.1 — Veo3.1 First Last Frame To Video by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • …and 11 more models in the sidebar.

Luma ​

  • Luma Ray Flash 2 Modify — Ray Flash 2 Modify by Luma - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Luma Ray 2 Modify — Ray 2 Modify by Luma - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Luma Ray Flash 2 Reframe — Ray Flash 2 Reframe is Luma's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Luma Ray 2 Reframe — Ray 2 Reframe is Luma's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Luma Ray Flash 2 Image to Video — Ray Flash 2 by Luma - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Luma Ray Flash 2 Text to Video — Ray Flash 2 is Luma's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Luma Ray 2 Image to Video — Ray 2 by Luma - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Luma Ray 2 Text to Video — Ray 2 is Luma's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

MiniMax ​

  • MiniMax H3 (Text to Video) — Generate native-stereo 2K video from text with MiniMax H3.
  • MiniMax H3 (Image to Video) — Generate native-stereo 2K video from a first frame and optional last frame with MiniMax H3.
  • MiniMax H3 (Reference to Video) — Generate native-stereo 2K video guided by image, video, and audio references with MiniMax H3.
  • MiniMax Hailuo 2.3 [Pro] (Image to Video) — Hailuo 2.3 Pro is MiniMax's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • MiniMax Hailuo 2.3 Fast [Standard] (Image to Video) — Hailuo 2.3 Fast Standard is MiniMax's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • MiniMax Hailuo 2.3 [Standard] (Image to Video) — Hailuo 2.3 Standard by MiniMax - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • MiniMax Hailuo 2.3 Fast [Pro] (Image to Video) — Hailuo 2.3 Fast Pro by MiniMax - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • MiniMax Hailuo 2.3 [Standard] (Text to Video) — Hailuo 2.3 Standard is MiniMax's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • MiniMax Hailuo 2.3 [Pro] (Text to Video) — Hailuo 2.3 Pro by MiniMax - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • MiniMax Hailuo 02 Fast (Image to Video) — Hailuo 02 Fast is MiniMax's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • MiniMax Hailuo 02 [Standard] (Image to Video) — Hailuo 02 Standard is MiniMax's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • MiniMax Hailuo 02 [Pro] (Image to Video) — Hailuo 02 Pro by MiniMax - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • …and 9 more models in the sidebar.

Pika ​

  • Pika V2.2 Frames — V2.2 Frames is Pika's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Pika V2.2 Image to Video — V2.2 is Pika's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Pika V2.2 Text to Video — V2.2 by Pika - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Pika V2 Image to Video Turbo — V2 Turbo is Pika's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Pika V1.5 Effects — V1.5 Effects by Pika - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Pika V2.2 Scenes — V2.2 Scenes is Pika's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Pika V2 Additions — V2 Additions by Pika - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Pika V2.1 Text to Video — V2.1 by Pika - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Pika V2 Text to Video Turbo — V2 Turbo by Pika - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Pika V2.1 Image to Video — V2.1 is Pika's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

VEED ​

  • VEED Fabric 1.0 Text to Video — Fabric 1.0 by VEED - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • VEED Video Background Removal Fast — Video Bg Removal Fast is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • VEED Video Background Removal — Video Bg Removal by VEED - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • VEED Video Background Removal Green Screen — Video Bg Removal Green Screen is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • VEED Fabric 1.0 Fast Image to Video — Fabric 1.0 Fast by VEED - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • VEED Fabric 1.0 Image to Video — Fabric 1.0 is VEED's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • VEED Lipsync — Lipsync is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • VEED Avatars Text to Video — Avatars is VEED's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • VEED Avatars Audio to Video — Avatars Audio To Video by VEED - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

Alibaba ​

  • Happy Horse Video Edit — Happy Horse Video Edit by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Happy Horse Reference to Video — Alibaba's #1-ranked Happy Horse 1.0 generates stunning 1080p videos with synchronized native audio and multilingual lip-sync, transforming text prompts or images into cinematic, true-to-life motion content for next-generation AIGC creation.
  • Happy Horse Image to Video — Alibaba's #1-ranked Happy Horse 1.0 generates stunning 1080p videos with synchronized native audio and multilingual lip-sync, transforming text prompts or images into cinematic, true-to-life motion content for next-generation AIGC creation.
  • Happy Horse Text to Video — Alibaba's #1-ranked Happy Horse 1.0 generates stunning 1080p videos with synchronized native audio and multilingual lip-sync, transforming text prompts or images into cinematic, true-to-life motion content for next-generation AIGC creation.
  • Wan 2.7 Text to Video — Alibaba Wan 2.7 text-to-video model with cinematic visuals, native audio generation, and configurable duration and resolution.
  • Wan 2.7 Reference to Video — Alibaba Wan 2.7 reference-to-video model generating video guided by reference content.
  • Wan 2.7 Edit Video — Alibaba Wan 2.7 video editing model.
  • Wan Motion — Wan Motion by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Wan 2.6 Reference to Video Flash — Wan 2.6 Flash by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Wan 2.6 Image to Video Flash — Wan 2.6 Flash by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Wan Move — Wan Move by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Wan 2.6 Text to Video — Alibaba Wan 2.6 text-to-video model with cinematic visuals, configurable duration (5 or 10 seconds) and resolution.
  • …and 45 more models in the sidebar.

Creatify ​

  • Creatify Aurora — Aurora by Creatify - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

ElevenLabs ​

  • ElevenLabs Dubbing — Dubbing by ElevenLabs - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

HeyGen ​

  • Heygen v5 Digital Twin — Avatar5 Digital Twin by HeyGen - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • HeyGen Video Agent V3 — Heygen Video Agent V3 is HeyGen's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • HeyGen Lipsync Precision — Heygen Lipsync Precision by HeyGen - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • HeyGen Lipsync Speed — Heygen Lipsync Speed by HeyGen - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • HeyGen Translate Speed — Heygen Translate Speed by HeyGen - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • HeyGen Translate Precision — Heygen Translate Precision by HeyGen - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • HeyGen Avatar 4 Image to Video — Heygen Avatar4 is HeyGen's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • HeyGen Avatar 4 Digital Twin — Heygen Avatar4 Digital Twin is HeyGen's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • HeyGen Avatar 3 Digital Twin — Heygen Avatar3 Digital Twin is HeyGen's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • HeyGen Video Agent V2 — Heygen Video Agent V2 is HeyGen's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

krea ​

  • Krea Wan 14b- Text to Video — Krea Wan by krea - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Krea Wan 14B — Krea Wan Video To Video is krea's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

KwaiVGI ​

  • Kling 3.0 Turbo Standard Image to Video — Kling Video 3.0 Turbo is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Kling 3.0 Turbo Pro Text to Video — Kling Video 3.0 Turbo is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Kling 3.0 Turbo Pro Image to Video — Kling Video 3.0 Turbo by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Kling 3.0 Turbo Standard Text to Video — Kling Video 3.0 Turbo by KwaiVGI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Kling Video O3 4k — Kling Video O3 4k by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Kling Video O3 4k — Kling Video O3 4k by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Kling Video O3 4k — Kling Video O3 4k is KwaiVGI's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Kling Video V3 4k — Kling Video V3 4k by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Kling Video V3 Pro — Kling Video V3 Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Kling Video V3 Standard — Kling Video V3 Standard by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Kling Video O3 Pro — Kling Video O3 Pro is KwaiVGI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Kling Video O3 Standard — Kling Video O3 Standard by KwaiVGI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • …and 48 more models in the sidebar.

Lightricks ​

  • LTX-2.3 22B Distilled — Ltx 2.3 22b Distilled Reference Video To Video is Lightricks's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • LTX 2.3 22B Distilled Reference to Video — Ltx 2.3 22b Distilled by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • LTX 2.3 22B — Ltx 2.3 22b Reference Video To Video Lora is Lightricks's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • LTX 2.3 22B Reference to Video — Ltx 2.3 22b by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • LTX-2.3 22B — Ltx 2.3 22b Extend Video Lora is Lightricks's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • LTX 2.3 22B Extend — Ltx 2.3 22b Extend by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • LTX-2.3 22B Distilled — Ltx 2.3 22b Distilled Video To Video is Lightricks's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • LTX-2.3 22B Distilled — Ltx 2.3 22b Distilled Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2.3 22B Distilled — Ltx 2.3 22b Distilled Lora is Lightricks's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • LTX-2.3 22B Distilled — Ltx 2.3 22b Distilled Lora by Lightricks - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • LTX 2.3 22B Distilled Video to Video — Ltx 2.3 22b Distilled Video To Video by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • LTX 2.3 22B Distilled Audio to Video — Ltx 2.3 22b Distilled Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • …and 50 more models in the sidebar.

meituan ​

  • LongCat Single Avatar — Longcat Single Avatar Image Audio To Video by sandbase-ai - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LongCat Video — Longcat Video 720p by meituan - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • LongCat Video — Longcat Video is meituan's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • LongCat Video — Longcat Video Image To Video 480p is meituan's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • LongCat Video — Longcat Video 480p by meituan - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • LongCat Video Distilled — Longcat Video Distilled is meituan's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • LongCat Video Distilled — Longcat Video Distilled Text To Video 720p by sandbase-ai - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • LongCat Video Distilled — Longcat Video Distilled Image To Video 480p is meituan's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • LongCat Video Distilled — Longcat Video Distilled 480p by meituan - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.

Meta ​

  • SAM 3.1 Video — Sam 3 1 Video Rle is Meta's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Sam 3 — Sam 3 Video is Meta's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

PixVerse ​

  • PixVerse C1 Reference to Video — C1 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse C1 Text to Video — C1 by PixVerse - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • PixVerse C1 Image to Video — C1 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse C1 Transition — C1 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse V6 Transition — V6 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse V6 Image to Video — V6 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse V6 Text to Video — V6 by PixVerse - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • PixVerse V5.6 Transition — V5.6 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse V5.6 Image to Video — V5.6 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • PixVerse V5.6 Text to Video — V5.6 by PixVerse - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • PixVerse V5.5 Effects — V5.5 Effects by PixVerse - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • PixVerse V5.5 Transition — V5.5 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • …and 28 more models in the sidebar.

Stability AI ​

  • Stable Avatar — Stable Avatar by stability-ai - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • High Quality Stable Video Diffusion — Stable Video by Stability AI - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

Tencent ​

  • Hunyuan Video 1.5 Image to Video — Hunyuan Video 1.5 by Tencent - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Hunyuan Video 1.5 Text to Video — Hunyuan Video 1.5 is Tencent's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Hunyuan Video Foley — Hunyuan Video Foley is Tencent's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Hunyuan Avatar — Hunyuan Avatar by Tencent - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Hunyuan Portrait — Hunyuan Portrait by Tencent - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Hunyuan Custom — Hunyuan Custom by Tencent - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Hunyuan Video Image to Video — Hunyuan Video by Tencent - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Hunyuan Video Video to Video — Hunyuan Video Video To Video by Tencent - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Hunyuan Video LoRA Inference — Hunyuan Video Lora by Tencent - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.
  • Hunyuan Video Text to Video — Hunyuan Video is Tencent's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

Topaz Labs ​

  • Topaz Video Upscale — Upscale Video is Topaz Labs's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

Vidu ​

  • Vidu Q3 Reference to Video Mix — Vidu Q3 Mix by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q3 Image to Video Turbo — Vidu Q3 Turbo by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q3 Text to Video Turbo — Vidu Q3 Turbo is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Vidu Q3 Image to Video — Vidu Q3 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q3 Text to Video — Vidu Q3 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Vidu Q2 Reference to Video Pro — Vidu Q2 Pro by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q2 Video Extension Pro — Vidu Q2 Video Extension is Shengshu's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Vidu Q2 Image to Video Turbo — Vidu Q2 Turbo by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q2 Image to Video Pro — Vidu Q2 Pro by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q2 Text to Video — Vidu Q2 is Shengshu's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.
  • Vidu Q1 Reference to Video — Vidu Q1 by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • Vidu Q1 Start-End to Video — Vidu Q1 Start End To Video by Shengshu - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.
  • …and 6 more models in the sidebar.

xAI ​

  • Grok Imagine Video 1.5 — Grok Imagine Video 1.5 is xAI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Grok Imagine Video Reference to Video — Grok Imagine Video is xAI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Grok Imagine Video — Grok Imagine Video is xAI's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.
  • Grok Imagine Video — Grok Imagine Video by xAI - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.

Capability coverage ​

audio-to-video, edit-video, first-last-frame-to-video, image-to-video, reference-to-video, text-to-video, video-to-video