Seedance 2.0 Mini & Fast API at Lowest Prices Worldwide — up to 68% off official pricing
Home
Explore
MiniMax
MiniMax H3
minimax/h3-max/image-to-video
Atlas Cloud GeneratorUnlock your potential as a director.Go Create
MiniMax H3 Max Image-to-Video
image-to-video

MiniMax H3 Max Image-to-Video API by MiniMax

minimax/h3-max/image-to-video
Image-to-video

MiniMax H3 Max image-to-video: animate a first-frame image (optionally with a last frame) driven by a text prompt. Supports 480P、768P, 5-15s.

Compare models

MiniMax H3 Max Image-to-Video is developed by MiniMax. Atlas Cloud (operated by Atlas Cloud AI LLC) provides access to it and does not own it. All trademarks belong to their respective owners.

MiniMax H3 Max Image-to-Video

MiniMax H3 Max Image-to-Video brings a static image to life at speed. Provide a first-frame image — and, optionally, a last frame — plus a prompt describing the motion, and the model generates a complete 24fps clip with audio that starts (and optionally ends) exactly on your images. A 5-second 768P clip renders in under 3 seconds, so you can try several motion directions in the time a single render usually takes.

Why Choose This?

  • Near-instant generation A 5s 768P clip in under 3 seconds; 15s clips in about 15 seconds.

  • Image-driven generation Animate any image with natural, controllable motion.

  • First & last frame control Set the opening frame, and optionally pin the closing frame for a precise transition.

  • Audio included Every clip is generated as complete audio-video at 24fps — no separate scoring or sound pass.

  • Aspect ratio from your image Automatically determined by the input image (always adaptive).

  • Selectable duration Produce 5–15s clips at 480P or 768P.

Parameters

ParameterRequiredDescription
promptYesText description of the desired motion and action
imageYesFirst frame of the video (public URL or Base64)
end_imageNoLast frame of the video (public URL or Base64)
resolutionYesVideo resolution. Available options: 480P, 768P. Default 768P
durationYesDuration of the generated video in seconds. Available options: 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15. Default 8
ratioNoAspect ratio of the generated video. For image-to-video, the aspect ratio is determined by the input image and is always adaptive; passing any other value is ignored.
prompt_expansionNoEnable AI prompt expansion via H3 Context-IR. When true, your prompt is first expanded into a rich, structured description (shot, soundscape, music) before generation, and an additional per-token Context-IR fee applies. Default false: the prompt is used as-is with no extra charge.
callback_urlNoHTTPS URL notified whenever the task status changes, so you can react to completion without polling. Requires a one-time verification handshake — see Task Status Callbacks below.

Task Status Callbacks (callback_url)

When you supply a callback_url, the MiniMax server notifies it on every task status change, so you don't have to poll for the result.

  1. Verification handshake — right after the task is created, the server first sends a verification request to your URL containing a challenge field. Your endpoint must return the challenge value unchanged, within 3 seconds, to complete verification.
  2. Status pushes — once verification succeeds, the server sends a POST to your URL every time the task status changes. The push body has the same structure as the Query Task (task status) response.

Callback status values: queued, running, succeeded, failed, cancelled.

The callback_url must be a publicly reachable HTTPS endpoint.

How to Use

  1. Upload your first-frame image — the video will start from this image.
  2. (Optional) Upload a last-frame image — the video will end on this image.
  3. Write your prompt — describe the motion, camera movement, and action.
  4. Set resolution and duration — pick 480P or 768P, and any length from 5s to 15s.
  5. Run — submit and download your video. The aspect ratio follows your input image automatically.

Best Use Cases

  • Photo Animation — Bring portraits, landscapes, and product images to life.
  • Start/End Transitions — Morph smoothly from one image to another.
  • Marketing & Ads — Turn product photos into dynamic promotional videos.
  • Storytelling — Animate illustrations and artwork for narratives.

Pro Tips

  • Be specific about movement direction, speed, and camera angles.
  • When using a last frame, keep the two images' aspect ratios close for a clean transition.
  • Use high-quality source images for better video results.
  • Describe environmental effects (wind, smoke, dust) for more immersive motion.

Notes

  • Both prompt and first-frame image are required.
  • The last-frame image is optional; when provided, the clip ends on it.
  • Supported image formats: png, jpeg, jpg, webp. Ensure image URLs are publicly accessible.
  • Supported durations between 5s and 15s, at 24fps with audio.
  • Resolution is limited to 480P and 768P — 2K is not available on H3 Max.
  • Keyframe interpolation and multimodal references (reference image / video / audio) are not supported.
  • Generation is asynchronous — submit, then poll for the finished video.

Explore Similar Models

One API for All Media AI.

Explore all models