MiniMax modelsvideo generation api

minimax/h3-max/image-to-video

H3 Max is MiniMax's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Input
Text prompt for video generation

PNG, JPEG, WebP, or GIF · 20 MiB maximum

Optional URL of the image to use as the first frame. When provided, the output aspect ratio follows this image. When omitted, the request is handled as text-to-video (16:9 by default).

PNG, JPEG, WebP, or GIF · 20 MiB maximum

Optional URL of the image to use as the last frame, for first-to-last keyframe generation.
The native generation resolution of the video. Allowed values: 480P, 768P.
515
The duration of the video in seconds. Range: 5 to 15.
How much effort to spend rewriting the prompt before generation. 'balanced' returns in about a second. 'quality' spends up to ~30s on a richer prompt.
Random seed. A random seed is selected when omitted.
Idle

Example output — click Run to generate your own

API README

MiniMax H3 Max Image to Video

Animate a starting image, with an optional ending image, using MiniMax H3 Max. The model supports prompt-directed motion, camera behavior, dialogue, and native audiovisual output.

Pricing

Current fal promotional rates are $0.0125 per second at 480P and $0.02 per second at 768P through September 7, 2026. Completed-task duration is used when available; otherwise the requested duration is used.

Technical Specs

  • Model ID: minimax/h3-max/image-to-video
  • Required inputs: prompt, starting image
  • Optional input: ending image
  • Duration: 5–15 seconds; default 5
  • Resolution: 480P or 768P; default 768P
  • Output: downloadable video URL
  • Execution: asynchronous

Related Models