TWO WEEKS ONLY | 20% OFF Seedream 5.0 Pro!

MiniMax API for Agentic LLMs and Video

The MiniMax API opens up MiniMax's model family from the Shanghai lab behind Hailuo, spanning the M series of agentic reasoning LLMs and the Hailuo video generators. M3 powers long-context coding agents, while Hailuo ranks first on WorldModelBench for physics simulation with lifelike fluid dynamics and motion. On Atlas Cloud every model runs under one account with transparent pay-as-you-go pricing and Day-0 access to new releases. Start building today.

AI Video Models by MiniMax

Generate cinematic, high-fidelity videos from text and images with the latest AI video generation models on Atlas Cloud.

View all models

Large Language Models by MiniMax

Power chat, reasoning, and agents at scale with leading large language models, served fast and affordably on Atlas Cloud.

View all models

MiniMax Models API Pricing Details

Compare standard vs. our pricing across every MiniMax model.

ModelStandard Price (USD)Our Price (USD)Discount
MiniMax M3
$0.6/$2.4per 1M tokens524.3K context
$0.3/$1.2M in/outper 1M tokens524.3K context
-50%View
MiniMax Speech 2.6 Turbo$0.06/K chars
Start from$0.048/K chars
-20%View
MiniMax Speech 2.6 HD$0.1/K chars
Start from$0.08/K chars
-20%View
MiniMax M2.7
$0.3/$1.2per 1M tokens196.6K context
$0.3/$1.2M in/outper 1M tokens196.6K context
View
MiniMax M2.5
$0.3/$1.2per 1M tokens196.6K context
$0.295/$1.2M in/outper 1M tokens196.6K context
View
MiniMax Music 2.6$0.15/K chars
Start from$0.15/K chars
View

Explore models from other providers

Instantly explore and experiment with 400+ production-ready models in the Atlas Playground. Start customizing with one click.

MiniMax API Use Cases You Can Build on Atlas Cloud

MiniMax's M series handles coding, agents, and long-context reasoning while Hailuo turns those outputs into video. Teams use both product lines together to run autonomous workflows and produce finished media without switching platforms.

Autonomous Software Engineering

Development teams use MiniMax M3 to run long-horizon coding agents that inspect large codebases, trace multi-file dependencies, and push working commits without human checkpoints. MiniMax demonstrated M3 running autonomously for nearly 12 hours to reproduce a research paper, generating 18 commits and 23 experimental figures from a single task description. Atlas Cloud's pay-as-you-go pricing makes it practical to run extended agentic sessions without committing to capacity upfront.

Long-context Document and Contract Analysis

Legal and research teams use MiniMax M3's 1M token context window to process entire contracts, case files, or research corpora in a single call. The model traces cross-document references, flags inconsistencies, and produces structured summaries across inputs that would exceed the limits of shorter-context models. This removes the manual chunking step that typically breaks context-dependent analysis in document-heavy workflows.

High-throughput Web Research and Tool Use

Product and data teams use MiniMax M2.5 for automated web research pipelines, scoring 76.3% on BrowseComp for browser-based task completion. The MoE architecture keeps latency low and cost predictable at $0.295 per million input tokens, making it viable for large-scale agentic runs that call external tools repeatedly. Teams running known, repeatable production workflows benefit from M2.5's consistent speed advantage over heavier models.

Brand Video and Product Ad Production

Marketing teams use Hailuo 02 Pro to generate 1080p product videos and brand story clips from a single reference image or text prompt. The model's physics accuracy and cinematic camera handling produce output that matches directorial instructions at an 85% complex instruction response rate. At $0.49 per second on Atlas Cloud, a 10-second production-quality clip costs less than most stock video licensing fees.

Anime and Stylized Content at Scale

Studios and content platforms use Hailuo 2.3 to generate anime-style and illustrated video content with accurate micro-expressions and fluid character body movement. The model's #1 ranking on WorldModelBench for physics simulation extends to stylized scenarios, keeping motion natural even in non-photorealistic outputs. Hailuo 2.3 Fast on Atlas Cloud at $0.19 per run keeps iteration costs low during the creative development phase.

Film Pre-production and Storyboard Previsualization

Directors and animators use Hailuo 02 to generate motion reference clips and storyboard previsualisations before committing to full production. Complex action sequences, physics-driven scenes, and camera movements can be tested via the API at a fraction of the cost of live-action or animation studio work. The consistent prompt adherence across repeated calls means the same scene description reliably produces comparable output for client presentations.

Render your enterprise vision into reality with Atlas Cloud AI.

Contact Sales

MiniMax API: Common Questions from Developers

The MiniMax API is a single interface to MiniMax's model family, covering the M-series large language models for text reasoning and agentic coding and the Hailuo models for video generation. On Atlas Cloud you reach all of them through one OpenAI-compatible endpoint, so a single integration handles both text and video workloads. Pricing is pay-as-you-go per call, with no subscription or commitment.

The lineup splits into two tracks. On the language side sit the M-series models, including M2, M2.5, and the newer M3, built for coding agents, tool use, and long-context reasoning. On the video side are the Hailuo models, such as Hailuo 02 and Hailuo 2.3, which produce short cinematic and stylized clips.

Yes. Every MiniMax model on Atlas Cloud, from the M-series LLMs to Hailuo 02 and Hailuo 2.3, runs under the same API key and base URL. You can move between a text reasoning call and a video generation call inside one pipeline without setting up separate authentication.

Create an Atlas Cloud account, generate one API key, and point your existing OpenAI-compatible client at the Atlas base URL. Because the MiniMax API follows the OpenAI request format, most integrations need only a base URL and model name change rather than a rewrite. Start building today.

Hailuo 02 Pro generates native 1080p video and supports both 6-second and 10-second clip lengths per request, while the Standard tier trades resolution for lower cost. Hailuo 2.3 continues 1080p output with stronger handling of stylized and character-driven motion. Video is billed per second of generated footage, so cost scales with the length you actually produce.

MiniMax M3 introduces the MSA (MiniMax Sparse Attention) architecture, native multimodal input for images and video, and a context window MiniMax documents at up to 1M tokens. The M2 series, including M2.5, targets high-throughput coding and agentic tasks and is text-only. Choose M3 for long-context, multimodal agents and the M2 series for cost-efficient, high-volume coding pipelines.

Hailuo 02 leans cinematic, built on MiniMax's NCR rendering architecture for photorealistic, live-action style output. Hailuo 2.3 keeps that physics foundation but broadens into anime, illustration, and game-CG styles with finer control over character motion and micro-expressions. Reach for 02 on product and live-action work, and 2.3 on stylized or animated content.

Hailuo 2.3 earned the Physics Champion title on WorldModelBench for its accuracy in simulating mass conservation, fluid dynamics, and spatial-temporal consistency. Rather than approximating motion visually, it models physical behavior, which holds up in demanding scenes like tipping liquids, drifting objects, and complex body movement. For action-heavy or technically precise prompts, that physical fidelity is often the deciding factor.

MiniMax documents M3 at up to 1M tokens of context, enabled by its sparse attention design. The earlier M-series models expose a large context window in the low hundreds of thousands of tokens, though the exact configured limit can vary by provider. Check the specific model page on Atlas Cloud for the context length in effect for your deployment.

Explore More Families

Seedance 2.0

The Seedance 2.0 API gives you production access to ByteDance's multimodal video model — quad-modal inputs (text, image, video, audio) and an industry-leading "Universal Reference" system that locks composition, camera movement, and character actions across shots. Integrate director-level control with one API call, a flat $0.09/s, instant key, and no waitlist — backed by enterprise-grade uptime and compliance. Seedance 2.0 Native 4K is now live!

View Family

Grok Imagine

The Grok Imagine API gives developers xAI's image, video, and audio generation in one suite. It produces up to 2K images with multilingual text rendering, plus video up to 15 seconds with native, synchronized audio and reference-based editing. On Atlas Cloud one key runs every Grok Imagine mode, so you move between image, video, and audio without separate setups, from $0.02 per image and $0.05 per second.

View Family

Gemini Omni Flash

The Gemini Omni API brings Google DeepMind's multimodal video generation and editing model, introduced at Google I/O 2026, to your stack. Gemini Omni fuses Gemini's reasoning engine with generative media, accepting any mix of text, images, video, and audio to produce consistent, knowledge-grounded output. Refine results through natural conversation, swapping objects, rewriting scenes, and shifting styles while physics, characters, and continuity stay intact. Atlas Cloud serves the full Gemini Omni Flash lineup, text-to-video, image-to-video with up to 7 reference images, and reference-to-video, through one unified API with transparent per-second pricing from $0.112 and no subscription. Start building today.

View Family

GPT Image 2

The GPT Image 2 API gives developers access to OpenAI's latest image model, the successor to GPT Image 1.5. It generates and edits images with accurate text rendering across Latin and CJK scripts, plus strong composition for posters, mockups, and infographics. On Atlas Cloud you reach it through one unified API alongside 300+ models, with free credits, 99.99% uptime, and no OpenAI organization verification required.

View Family

Google

Google's most powerful creative models are all available on Atlas Cloud. Veo 3.1 delivers cinematic video generation, Nano Banana 2 powers high-fidelity image creation, and Gemini brings multimodal intelligence to every workflow. Access the full Google model suite through one API key with Day-0 availability and pay-as-you-go pricing.

View Family

Seedance 2.0 Mini

The Seedance 2.0 Mini API is the lightest, lowest-cost tier of ByteDance's Seedance video line, built for teams where throughput and unit cost matter more than maximum polish. Use it for batch generation, rapid prototyping, and draft passes, all through one OpenAI-compatible key on Atlas Cloud.

View Family

ByteDance

From cinematic video generation to high-fidelity image creation, ByteDance's most powerful models are live on Atlas Cloud. Run Seedance and Seedream at scale with the lowest inference pricing and zero infrastructure overhead.

View Family

Alibaba

Atlas Cloud brings together Alibaba's full model lineup under one API: Qwen for language and image tasks, Wan for video generation up to 1080p. Access every model pay-as-you-go with no subscriptions. The Alibaba API is available via a single base URL using your existing OpenAI-compatible client.

View Family

OpenAI

Atlas Cloud gives you access to the full OpenAI API lineup, from GPT Image 2 for image generation to Sora 2 for video. Every model is available pay-as-you-go with no monthly commitment. Plug in with a single base URL swap using the OpenAI-compatible API.

View Family

xAI

Build complete image and video pipelines using the xAI API on Atlas Cloud. Generate at 2K, edit with reference images, and animate images into audio-synced clips.

View Family

Kwaivgi

The Kwaivgi API at 15% off standard rates. Day-0 access to every new Kling release, pay-as-you-go, no seat limits. One account covers the full Kling lineup.

View Family

Seedream 5.0 Pro

Seedream 5.0 Pro API gives developers ByteDance's controllable image editing model on Atlas Cloud. It places edits precisely with anchors and coordinates, separates images into editable layers, fuses multiple references, and matches exact colors and materials, with multilingual text at 2K and 3K. On Atlas Cloud you reach it through one key!

View Family

Recommended Articles

Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.