рджреБрдирд┐рдпрд╛ рднрд░ рдореЗрдВ рд╕рдмрд╕реЗ рдХрдо рдХреАрдорддреЛрдВ рдкрд░ Seedance 2.0 Mini & Fast API тАФ рдЖрдзрд┐рдХрд╛рд░рд┐рдХ рдХреАрдордд рдкрд░ 68% рддрдХ рдХреА рдЫреВрдЯ
Atlas Cloud AI рдХреНрд░рд┐рдПрд╢рди рд╕реНрдЯреВрдбрд┐рдпреЛрдЕрдкрдиреЗ рднреАрддрд░ рдХреЗ рдирд┐рд░реНрджреЗрд╢рдХ рдХреЛ рдкрд╣рдЪрд╛рдиреЗрдВредрдмрдирд╛рдирд╛ рд╢реБрд░реВ рдХрд░реЗрдВ
MiniMax H3 Reference-to-Video
рдЗрдореЗрдЬ-рд╕реЗ-рд╡реАрдбрд┐рдпреЛ

MiniMax H3 Reference-to-Video API by MiniMax

minimax/h3/reference-to-video
Reference-to-video

MiniMax H3 reference-to-video: generate a video that keeps the subject from a reference image, driven by a text prompt. Supports 2K, 5-15s.

MiniMax H3 Reference-to-Video рдХреЛ MiniMax рджреНрд╡рд╛рд░рд╛ рд╡рд┐рдХрд╕рд┐рдд рдХрд┐рдпрд╛ рдЧрдпрд╛ рд╣реИред Atlas Cloud (Atlas Cloud AI LLC рджреНрд╡рд╛рд░рд╛ рд╕рдВрдЪрд╛рд▓рд┐рдд) рдЗрд╕ рддрдХ рдкрд╣реБрдБрдЪ рдкреНрд░рджрд╛рди рдХрд░рддрд╛ рд╣реИ, рдЗрд╕рдХрд╛ рд╕реНрд╡рд╛рдореА рдирд╣реАрдВ рд╣реИред рд╕рднреА рдЯреНрд░реЗрдбрдорд╛рд░реНрдХ рдЙрдирдХреЗ рд╕рдВрдмрдВрдзрд┐рдд рд╕реНрд╡рд╛рдорд┐рдпреЛрдВ рдХреА рд╕рдВрдкрддреНрддрд┐ рд╣реИрдВред

MiniMax H3 Reference-to-Video

MiniMax H3 Reference-to-Video generates a video that keeps the subjects from your reference materials consistent throughout, driven by your text prompt. Instead of animating a fixed frame, it uses the references as identity anchors тАФ ideal for placing a character, product, or style into new scenes and motions. References can be any mix of images, videos, and an audio track (e.g. a music beat) to sync the motion to.

Why Choose This?

  • Subject consistency Preserve one or several people, characters, or objects' identities across the whole clip.

  • Mix reference types Combine reference images, reference videos, and reference audio in a single request.

  • Audio sync (optional) Provide a reference audio track and sync the action to the beat.

  • Prompt-driven scenes Put the reference subjects into any scene or action you describe.

  • High resolution output Generate videos in 2K quality.

  • Flexible aspect ratios adaptive, 21:9, 16:9, 4:3, 1:1, 3:4, or 9:16.

  • Selectable duration Produce 4тАУ15s clips.

Parameters

ParameterRequiredDescription
promptYesText description of the scene and action for the subjects
refersYesArray of reference materials, each { "url": "...", "type": "image|video|audio" }. type is optional (inferred from the URL). At least one image OR video is required; audio alone is not allowed.
resolutionYesVideo resolution. Available options: 480P, 768P, 2K
durationYesDuration of the generated video in seconds. Integer between 4 and 15. Available options: 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15
ratioNoAspect ratio of the generated video. Defaults to adaptive (the most suitable ratio is chosen automatically based on the input). You may also explicitly specify a concrete ratio. Available options: adaptive, 21:9, 16:9, 4:3, 1:1, 3:4, 9:16
prompt_expansionNoEnable AI prompt expansion via H3 Context-IR. When true, your prompt is first expanded into a rich, structured description (shot, soundscape, music) before generation, and an additional per-token Context-IR fee applies. Default false: the prompt is used as-is with no extra charge.

refers example

"refers": [ { "url": "https://.../subject1.png", "type": "image" }, { "url": "https://.../subject2.png" }, { "url": "https://.../clip.mp4", "type": "video" }, { "url": "https://.../beat.mp3", "type": "audio" } ]

How to Use

  1. Add your references тАФ any mix of images and/or videos (at least one), optionally plus an audio track.
  2. Write your prompt тАФ describe the scene, action, and camera movement for the subjects.
  3. Set resolution and duration тАФ balance quality against generation speed.
  4. Choose an aspect ratio тАФ match your target platform, or use adaptive.
  5. Run тАФ submit and download your video.

Best Use Cases

  • Character Consistency тАФ Keep the same character across multiple shots.
  • Product Placement тАФ Feature a specific product in generated scenes.
  • Branded Content тАФ Maintain a consistent mascot or spokesperson.
  • Creative Series тАФ Build multi-clip stories around a recurring subject.

Pro Tips

  • Use clean, well-lit reference images with each subject clearly visible.
  • Give one image per subject when combining several subjects in one video.
  • Describe the new scene and action, not the reference materials themselves.
  • Supply a reference audio track to drive the timing/beat of the motion.
  • Combine with different ratios and durations to repurpose the same subjects across platforms.

Notes

  • prompt and at least one reference image or video are required; a request with only audio is rejected.
  • Reference-to-video and image-to-video are mutually exclusive: this model uses reference materials, not first/last frames тАФ do not mix the two.
  • Ensure all reference URLs (image / video / audio) are publicly accessible.
  • Generation is asynchronous тАФ submit, then poll for the finished video.

рд╕рдорд╛рди рдореЙрдбрд▓ рджреЗрдЦреЗрдВ

рд╣рд░ рдореАрдбрд┐рдпрд╛ AI рдХреЗ рд▓рд┐рдП рдПрдХ рд╣реА APIред

рд╕рднреА рдореЙрдбрд▓ рдПрдХреНрд╕рдкреНрд▓реЛрд░ рдХрд░реЗрдВ