SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsvideo generation api

alibaba/happy-horse/video-edit

Happy Horse Video Edit by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

Input
Text prompt describing the desired edit. Reference any supplied reference images using @Image1, @Image2, ... up to @Image5. Max 2500 characters.
URL of the source video to edit. Formats: MP4, MOV (H.264 recommended). Duration: 3-60 s. Longer side ≤ 2160 px, shorter side ≥ 320 px. Aspect ratio between 1:2.5 and 2.5:1. Frame rate > 8 fps. Max 100 MB. The output video preserves the source aspect ratio. Output duration matches the input video, capped at 15 s (longer inputs are truncated to the first 15 s). Max file size: 100.0MB, Min duration: 3s, Max duration: 60s, Timeout: 30.0s
Output video resolution tier. Allowed values: 720p, 1080p.
Random seed for reproducibility (0-2147483647).

PNG, JPEG, WebP, or GIF · 20 MiB maximum each

Optional reference images used to guide the edit (up to 5). Formats: JPEG, JPG, PNG, WEBP. Dimensions must be at least 300px. Aspect ratio between 1:2.5 and 2.5:1. Max 10 MB each.
Audio handling. 'auto': model decides whether to regenerate audio. 'origin': preserve the original audio from the input video. Allowed values: auto, origin.
Idle

Example output — click Run to generate your own

API README

Happy Horse Video Edit

Happy Horse Video Edit is a video editing workflow for turning supplied visual material and written direction into a new moving sequence. It starts from an existing clip and applies a directed transformation.

Use the route for rapid creative iteration when a key asset already exists and the goal is to extend, animate, or revise it. Describe subject action, camera behavior, timing, atmosphere, and the intended ending; local duration, resolution, aspect-ratio, and seed controls define the delivery envelope.

Highlights

  • Prompt-directed video revision. Reworks an existing clip according to a written creative change.
  • Directed subject movement. Uses the prompt to define action, pace, staging, and interaction within the shot.
  • Camera and atmosphere control. Lets the brief specify framing, camera movement, lighting, mood, and scene energy.
  • Production-oriented delivery. Supports explicit duration, resolution, framing, and repeatability controls where exposed by the route.

Pricing

ConfigurationModePrice
Any supported durationPer generated second$0.28

When to Use

✅ Good fit❌ Consider alternatives
The project needs revising an existing visual assetThe goal is a different media task or endpoint
The available inputs match the required local schemaRequired source media or permissions are unavailable
The brief can specify subject, composition, style, and deliveryThe result must be deterministic at pixel or sample level
The supported formats and controls match final placementDelivery requires unsupported dimensions, codecs, or duration
An asynchronous generated result fits the workflowA live, frame-synchronous, or real-time response is mandatory

Prompt Guide

Write the request as a production brief: identify the main subject or source material, state the intended transformation, describe composition or timing, and finish with style, atmosphere, and delivery constraints. Keep preservation requirements separate from requested changes, and use only fields exposed by this route.

{
  "prompt": "A cinematic, precisely directed scene with clear subject action, camera movement, lighting, and atmosphere",
  "seed": 1,
  "video": "https://example.com/source.mp4",
  "resolution": "720p"
}

Technical Specs

SpecValue
Model IDalibaba/happy-horse/video-edit
Input fieldsseed (integer)<br>video (string)<br>prompt (string)<br>resolution (string; 720p, 1080p)<br>audio_setting (string; auto, origin)<br>reference_image_urls (array)
Required inputprompt
Output fieldsurl, content_type
ExecutionAsynchronous job

Related Models

  • alibaba/qwen-image-3/edit
  • alibaba/wan/2.2/image-to-video/turbo
  • alibaba/wan/2.1/text-to-video

Related Models

alibaba/happy-horse/1.1/image-to-videoHappy Horse 1.1 by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.alibaba/happy-horse/1.1/reference-to-videoHappy Horse 1.1 by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.alibaba/happy-horse/1.1/text-to-videoHappy Horse 1.1 is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.alibaba/happy-horse/image-to-videoAlibaba's #1-ranked Happy Horse 1.0 generates stunning 1080p videos with synchronized native audio and multilingual lip-sync, transforming text prompts or images into cinematic, true-to-life motion content for next-generation AIGC creation.alibaba/happy-horse/reference-to-videoAlibaba's #1-ranked Happy Horse 1.0 generates stunning 1080p videos with synchronized native audio and multilingual lip-sync, transforming text prompts or images into cinematic, true-to-life motion content for next-generation AIGC creation.alibaba/happy-horse/text-to-videoAlibaba's #1-ranked Happy Horse 1.0 generates stunning 1080p videos with synchronized native audio and multilingual lip-sync, transforming text prompts or images into cinematic, true-to-life motion content for next-generation AIGC creation.alibaba/flashvsrFlashVSR is a video super-resolution model that upscales videos to higher resolutions (720p / 1080p / 2K / 4K) with fast inference.alibaba/wan-animateWan-Animate generates an animated video from an input image and a driving video, transferring motion onto the image subject.