์ „ ์„ธ๊ณ„ ์ตœ์ €๊ฐ€๋กœ ๋งŒ๋‚˜๋Š” Seedance 2.0 Mini & Fast API โ€” ๊ณต์‹ ๊ฐ€๊ฒฉ ๋Œ€๋น„ ์ตœ๋Œ€ 68% ํ• ์ธ
ํ™ˆ
ํƒ์ƒ‰
MiniMax
MiniMax H3
minimax/h3/reference-to-video
Atlas Cloud AI ์ฐฝ์ž‘ ์ŠคํŠœ๋””์˜ค๊ฐ๋…์œผ๋กœ์„œ์˜ ์ž ์žฌ๋ ฅ์„ ํŽผ์ณ ๋ณด์„ธ์š”.์ œ์ž‘ ์‹œ์ž‘
MiniMax H3 Reference-to-Video
์ด๋ฏธ์ง€๋ฅผ ๋น„๋””์˜ค๋กœ

MiniMax H3 Reference-to-Video API by MiniMax

minimax/h3/reference-to-video
Reference-to-video

MiniMax H3 reference-to-video: generate a video that keeps the subject from a reference image, driven by a text prompt. Supports 2K, 5-15s.

MiniMax H3 Reference-to-Video์€(๋Š”) MiniMax์—์„œ ๊ฐœ๋ฐœํ•œ ๋ชจ๋ธ์ž…๋‹ˆ๋‹ค. Atlas Cloud(์šด์˜: Atlas Cloud AI LLC)๋Š” ํ•ด๋‹น ๋ชจ๋ธ์— ๋Œ€ํ•œ ์•ก์„ธ์Šค๋ฅผ ์ œ๊ณตํ•  ๋ฟ์ด๋ฉฐ ์ด๋ฅผ ์†Œ์œ ํ•˜์ง€ ์•Š์Šต๋‹ˆ๋‹ค. ๋ชจ๋“  ์ƒํ‘œ๋Š” ๊ฐ ์†Œ์œ ์ž์—๊ฒŒ ๊ท€์†๋ฉ๋‹ˆ๋‹ค.

MiniMax H3 Reference-to-Video

MiniMax H3 Reference-to-Video generates a video that keeps the subjects from your reference materials consistent throughout, driven by your text prompt. Instead of animating a fixed frame, it uses the references as identity anchors โ€” ideal for placing a character, product, or style into new scenes and motions. References can be any mix of images, videos, and an audio track (e.g. a music beat) to sync the motion to.

Why Choose This?

  • Subject consistency Preserve one or several people, characters, or objects' identities across the whole clip.

  • Mix reference types Combine reference images, reference videos, and reference audio in a single request.

  • Audio sync (optional) Provide a reference audio track and sync the action to the beat.

  • Prompt-driven scenes Put the reference subjects into any scene or action you describe.

  • High resolution output Generate videos in 2K quality.

  • Flexible aspect ratios adaptive, 21:9, 16:9, 4:3, 1:1, 3:4, or 9:16.

  • Selectable duration Produce 4โ€“15s clips.

Parameters

ParameterRequiredDescription
promptYesText description of the scene and action for the subjects
refersYesArray of reference materials, each { "url": "...", "type": "image|video|audio" }. type is optional (inferred from the URL). At least one image OR video is required; audio alone is not allowed.
resolutionYesVideo resolution. Available options: 480P, 768P, 2K
durationYesDuration of the generated video in seconds. Integer between 4 and 15. Available options: 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15
ratioNoAspect ratio of the generated video. Defaults to adaptive (the most suitable ratio is chosen automatically based on the input). You may also explicitly specify a concrete ratio. Available options: adaptive, 21:9, 16:9, 4:3, 1:1, 3:4, 9:16
prompt_expansionNoEnable AI prompt expansion via H3 Context-IR. When true, your prompt is first expanded into a rich, structured description (shot, soundscape, music) before generation, and an additional per-token Context-IR fee applies. Default false: the prompt is used as-is with no extra charge.

refers example

"refers": [ { "url": "https://.../subject1.png", "type": "image" }, { "url": "https://.../subject2.png" }, { "url": "https://.../clip.mp4", "type": "video" }, { "url": "https://.../beat.mp3", "type": "audio" } ]

How to Use

  1. Add your references โ€” any mix of images and/or videos (at least one), optionally plus an audio track.
  2. Write your prompt โ€” describe the scene, action, and camera movement for the subjects.
  3. Set resolution and duration โ€” balance quality against generation speed.
  4. Choose an aspect ratio โ€” match your target platform, or use adaptive.
  5. Run โ€” submit and download your video.

Best Use Cases

  • Character Consistency โ€” Keep the same character across multiple shots.
  • Product Placement โ€” Feature a specific product in generated scenes.
  • Branded Content โ€” Maintain a consistent mascot or spokesperson.
  • Creative Series โ€” Build multi-clip stories around a recurring subject.

Pro Tips

  • Use clean, well-lit reference images with each subject clearly visible.
  • Give one image per subject when combining several subjects in one video.
  • Describe the new scene and action, not the reference materials themselves.
  • Supply a reference audio track to drive the timing/beat of the motion.
  • Combine with different ratios and durations to repurpose the same subjects across platforms.

Notes

  • prompt and at least one reference image or video are required; a request with only audio is rejected.
  • Reference-to-video and image-to-video are mutually exclusive: this model uses reference materials, not first/last frames โ€” do not mix the two.
  • Ensure all reference URLs (image / video / audio) are publicly accessible.
  • Generation is asynchronous โ€” submit, then poll for the finished video.

์œ ์‚ฌํ•œ ๋ชจ๋ธ ํƒ์ƒ‰

ํ•˜๋‚˜์˜ API๋กœ ๋ชจ๋“  ๋ฏธ๋””์–ด AI๋ฅผ.

๋ชจ๋“  ๋ชจ๋ธ ํƒ์ƒ‰