MiniMax modelsvideo generation api

minimax/h3-max/image-to-video

H3 Max is MiniMax's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

Input
Text prompt for video generation
Optional URL of the image to use as the first frame. When provided, the output aspect ratio follows this image. When omitted, the request is handled as text-to-video (16:9 by default).
Optional URL of the image to use as the last frame, for first-to-last keyframe generation.
The native generation resolution of the video. Allowed values: 480P, 768P.
515
The duration of the video in seconds. Range: 5 to 15.
How much effort to spend rewriting the prompt before generation. 'balanced' returns in about a second. 'quality' spends up to ~30s on a richer prompt.
Random seed. A random seed is selected when omitted.
Idle

Example output — click Run to generate your own

API README

MiniMax H3 Max Image to Video

Animate a starting image, with an optional ending image, using MiniMax H3 Max. The model supports prompt-directed motion, camera behavior, dialogue, and native audiovisual output.

Pricing

Current fal promotional rates are $0.0125 per second at 480P and $0.02 per second at 768P through September 7, 2026. Completed-task duration is used when available; otherwise the requested duration is used.

Technical Specs

  • Model ID: minimax/h3-max/image-to-video
  • Required inputs: prompt, starting image
  • Optional input: ending image
  • Duration: 5–15 seconds; default 5
  • Resolution: 480P or 768P; default 768P
  • Output: downloadable video URL
  • Execution: asynchronous

Related Models