
The MiniMax API opens up MiniMax's model family from the Shanghai lab behind Hailuo, spanning the M series of agentic reasoning LLMs and the Hailuo video generators. M3 powers long-context coding agents, while Hailuo ranks first on WorldModelBench for physics simulation with lifelike fluid dynamics and motion. On Atlas Cloud every model runs under one account with transparent pay-as-you-go pricing and Day-0 access to new releases. Start building today.
Generate cinematic, high-fidelity videos from text and images with the latest AI video generation models on Atlas Cloud.
Power chat, reasoning, and agents at scale with leading large language models, served fast and affordably on Atlas Cloud.
Compare standard vs. our pricing across every MiniMax model.
| Model | Standard Price (USD) | Our Price (USD) | Discount | |
|---|---|---|---|---|
| MiniMax M3 | $0.6/$2.4per 1M tokens524.3K context | $0.3/$1.2M in/outper 1M tokens524.3K context | -50% | View |
| MiniMax Speech 2.6 Turbo | $0.06/K chars | Start from$0.048/K chars | -20% | View |
| MiniMax Speech 2.6 HD | $0.1/K chars | Start from$0.08/K chars | -20% | View |
| MiniMax M2.7 | $0.3/$1.2per 1M tokens196.6K context | $0.3/$1.2M in/outper 1M tokens196.6K context | — | View |
| MiniMax M2.5 | $0.3/$1.2per 1M tokens196.6K context | $0.295/$1.2M in/outper 1M tokens196.6K context | — | View |
| MiniMax Music 2.6 | $0.15/K chars | Start from$0.15/K chars | — | View |
Instantly explore and experiment with 400+ production-ready models in the Atlas Playground. Start customizing with one click.
MiniMax's M series handles coding, agents, and long-context reasoning while Hailuo turns those outputs into video. Teams use both product lines together to run autonomous workflows and produce finished media without switching platforms.
Development teams use MiniMax M3 to run long-horizon coding agents that inspect large codebases, trace multi-file dependencies, and push working commits without human checkpoints. MiniMax demonstrated M3 running autonomously for nearly 12 hours to reproduce a research paper, generating 18 commits and 23 experimental figures from a single task description. Atlas Cloud's pay-as-you-go pricing makes it practical to run extended agentic sessions without committing to capacity upfront.
Legal and research teams use MiniMax M3's 1M token context window to process entire contracts, case files, or research corpora in a single call. The model traces cross-document references, flags inconsistencies, and produces structured summaries across inputs that would exceed the limits of shorter-context models. This removes the manual chunking step that typically breaks context-dependent analysis in document-heavy workflows.
Product and data teams use MiniMax M2.5 for automated web research pipelines, scoring 76.3% on BrowseComp for browser-based task completion. The MoE architecture keeps latency low and cost predictable at $0.295 per million input tokens, making it viable for large-scale agentic runs that call external tools repeatedly. Teams running known, repeatable production workflows benefit from M2.5's consistent speed advantage over heavier models.
Marketing teams use Hailuo 02 Pro to generate 1080p product videos and brand story clips from a single reference image or text prompt. The model's physics accuracy and cinematic camera handling produce output that matches directorial instructions at an 85% complex instruction response rate. At $0.49 per second on Atlas Cloud, a 10-second production-quality clip costs less than most stock video licensing fees.
Studios and content platforms use Hailuo 2.3 to generate anime-style and illustrated video content with accurate micro-expressions and fluid character body movement. The model's #1 ranking on WorldModelBench for physics simulation extends to stylized scenarios, keeping motion natural even in non-photorealistic outputs. Hailuo 2.3 Fast on Atlas Cloud at $0.19 per run keeps iteration costs low during the creative development phase.
Directors and animators use Hailuo 02 to generate motion reference clips and storyboard previsualisations before committing to full production. Complex action sequences, physics-driven scenes, and camera movements can be tested via the API at a fraction of the cost of live-action or animation studio work. The consistent prompt adherence across repeated calls means the same scene description reliably produces comparable output for client presentations.
The MiniMax API is a single interface to MiniMax's model family, covering the M-series large language models for text reasoning and agentic coding and the Hailuo models for video generation. On Atlas Cloud you reach all of them through one OpenAI-compatible endpoint, so a single integration handles both text and video workloads. Pricing is pay-as-you-go per call, with no subscription or commitment.
The lineup splits into two tracks. On the language side sit the M-series models, including M2, M2.5, and the newer M3, built for coding agents, tool use, and long-context reasoning. On the video side are the Hailuo models, such as Hailuo 02 and Hailuo 2.3, which produce short cinematic and stylized clips.
Yes. Every MiniMax model on Atlas Cloud, from the M-series LLMs to Hailuo 02 and Hailuo 2.3, runs under the same API key and base URL. You can move between a text reasoning call and a video generation call inside one pipeline without setting up separate authentication.
Create an Atlas Cloud account, generate one API key, and point your existing OpenAI-compatible client at the Atlas base URL. Because the MiniMax API follows the OpenAI request format, most integrations need only a base URL and model name change rather than a rewrite. Start building today.
Hailuo 02 Pro generates native 1080p video and supports both 6-second and 10-second clip lengths per request, while the Standard tier trades resolution for lower cost. Hailuo 2.3 continues 1080p output with stronger handling of stylized and character-driven motion. Video is billed per second of generated footage, so cost scales with the length you actually produce.
MiniMax M3 introduces the MSA (MiniMax Sparse Attention) architecture, native multimodal input for images and video, and a context window MiniMax documents at up to 1M tokens. The M2 series, including M2.5, targets high-throughput coding and agentic tasks and is text-only. Choose M3 for long-context, multimodal agents and the M2 series for cost-efficient, high-volume coding pipelines.
Hailuo 02 leans cinematic, built on MiniMax's NCR rendering architecture for photorealistic, live-action style output. Hailuo 2.3 keeps that physics foundation but broadens into anime, illustration, and game-CG styles with finer control over character motion and micro-expressions. Reach for 02 on product and live-action work, and 2.3 on stylized or animated content.
Hailuo 2.3 earned the Physics Champion title on WorldModelBench for its accuracy in simulating mass conservation, fluid dynamics, and spatial-temporal consistency. Rather than approximating motion visually, it models physical behavior, which holds up in demanding scenes like tipping liquids, drifting objects, and complex body movement. For action-heavy or technically precise prompts, that physical fidelity is often the deciding factor.
MiniMax documents M3 at up to 1M tokens of context, enabled by its sparse attention design. The earlier M-series models expose a large context window in the low hundreds of thousands of tokens, though the exact configured limit can vary by provider. Check the specific model page on Atlas Cloud for the context length in effect for your deployment.
Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.