Seedance 2.5 Now Live — First on Atlas Cloud
Home
Explore
MiniMax
MiniMax H3
minimax/h3/text-to-video
MiniMax H3 Text-to-Video
text-to-video

MiniMax H3 Text-to-Video API by MiniMax

minimax/h3/text-to-video
Text-to-video

MiniMax H3 text-to-video: generate a cinematic video from a text prompt. Supports 2K, 5-15s., and 16:9/9:16/1:1/adaptive aspect ratios.

Compare models

MiniMax H3 Text-to-Video is developed by MiniMax. Atlas Cloud (operated by Atlas Cloud AI LLC) provides access to it and does not own it. All trademarks belong to their respective owners.

MiniMax H3 Text-to-Video

MiniMax H3 Text-to-Video is a state-of-the-art AI video generation model that creates cinematic, high-fidelity videos directly from a text prompt. With crisp detail up to 2K, smooth natural motion, and flexible aspect ratios, it turns a single description into a polished clip.

Why Choose This?

  • High resolution output Generate videos in 2K quality.

  • Cinematic motion Fluid camera work and lifelike movement from a plain-text description.

  • Flexible aspect ratios 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 or adaptive to fit any platform.

  • Selectable duration Produce 4–15s clips.

Parameters

ParameterRequiredDescription
promptYesText description of the scene, subject, and action
resolutionYesVideo resolution. Available options: 768P, 2K
durationYesDuration of the generated video in seconds. Integer between 4 and 15. Available options: 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15
ratioYesAspect ratio of the generated video. For text-to-video, ratio is required and cannot be adaptive. Available options: 21:9, 16:9, 4:3, 1:1, 3:4, 9:16

How to Use

  1. Write your prompt — describe the scene, subject, lighting, and motion in detail.
  2. Set resolution — higher for quality, lower for faster generation.
  3. Choose an aspect ratio — pick one of the supported ratios; adaptive is not available for text-to-video.
  4. Adjust duration — pick 5s or 10s.
  5. Run — submit and download your video.

Pricing

Billed per second of generated video, by resolution:

ResolutionCost per second
2K$0.14
768p$0.10

Best Use Cases

  • Social Media Content — Short-form clips for TikTok, Reels, and Stories.
  • Concept Visualization — Bring ideas to life without filming.
  • Marketing Videos — Produce promotional content from text.
  • Storytelling — Create narrative scenes for creative projects.

Pro Tips

  • Be specific about camera angle, lighting, mood, and motion.
  • For text-to-video, ratio must be an explicit value (e.g. 16:9); adaptive is not supported for this mode.
  • Higher resolution (2K) suits hero shots; 768P is great for quick iteration.
  • Describe environmental effects (wind, smoke, golden-hour light) for richer results.

Notes

  • A prompt is required.
  • Supported durations between 4s and 15s.
  • Generation is asynchronous — submit, then poll for the finished video.

Explore Similar Models

One API for All Media AI.

Explore all models