Seedance 2.5 Now Live — First on Atlas Cloud

Alibaba Models on Atlas Cloud | Wan & Qwen

Atlas Cloud brings together Alibaba's full model lineup under one API: Qwen for language and image tasks, Wan for video generation up to 1080p. Access every model pay-as-you-go with no subscriptions. The Alibaba API is available via a single base URL using your existing OpenAI-compatible client.

Alibaba is developed by Alibaba. Atlas Cloud (operated by Atlas Cloud AI LLC) provides access to it and does not own it. All trademarks belong to their respective owners.

AI Video Models by Alibaba

Generate cinematic, high-fidelity videos from text and images with the latest AI video generation models on Atlas Cloud.

View all models
Wan 3.0
text-to-video
image-to-video
video-to-video

Wan 3.0

Wan 3.0 API is the next generation of Alibaba's Wan video family, built to push long-form generation, multi-reference control, and audiovisual quality to new heights. Atlas Cloud already hosts Wan 2.7, 2.6, and 2.5, and Wan 3.0 runs on the same unified key with no separate setup. Start building today. Scroll down to the showcase to see what Wan 3.0 can create.

3 modelsExplore Wan 3.0
Wan 2.7
image-to-video
text-to-video
video-to-video
image-to-image

Wan 2.7

The Wan 2.7 API gives developers Alibaba's all-in-one video suite, covering text to video, image to video, reference to video, and video editing, plus image generation. It produces native 1080p clips up to 15 seconds with synced audio, first to last frame control, and up to 5 character references. (See the latest [Wan 3.0 API page](https://www.atlascloud.ai/models/wan-3.0) for native 30-second video and up to 20 mixed references.)

8 modelsExplore Wan 2.7
Happy Horse
LLM
image-to-video
text-to-video
video-to-video

Happy Horse

The HappyHorse API on Atlas Cloud connects your application to Alibaba's HappyHorse 1.0 and 1.1 video generation models. Produce clips running anywhere from 3 to 15 seconds at 720p or 1080p, steer results with up to nine reference images, and pick the aspect ratio that fits your product. Atlas Cloud adds Day-0 model availability, reliable uptime, and a simple asynchronous REST integration. Start building today.

7 modelsExplore Happy Horse
Wan 2.6
image-to-video
image-to-image
video-to-video
text-to-video

Wan 2.6

The Wan 2.6 API builds a multi-shot story from one prompt instead of a single clip. Alibaba's storytelling engine cuts between shots with coherent transitions while holding character identity, voice, and style steady across the scene. It spans text, image, and reference to video at 1080p, with synchronized audio and multilingual lip-sync. Reach it through one key on Atlas Cloud, next to 300+ models.

6 modelsExplore Wan 2.6
Van
text-to-video
image-to-video

Van

Van is a flagship AI video series built on the Wan 2.5 and 2.6 frameworks, and the Van API unifies its full text-to-video and image-to-video lineup on Atlas Cloud. Render cinematic clips up to 1080p, then pick speed-optimized 2.6 models for fast iteration or cost-efficient 2.5 variants for large batches. Every model ships on Day-0 behind one OpenAI-compatible key with transparent pay-as-you-go pricing. Start building today.

4 modelsExplore Van
Wan 2.5
text-to-video
image-to-video
image-to-image
text-to-image

Wan 2.5

The Wan 2.5 API produces video and audio in one pass, so voice, sound, and lip-sync line up without a separate step. Built on Alibaba's Diffusion Transformer architecture, it covers text and image to video from 480p to 1080p at 5 or 10 seconds, with reliable sync even for Chinese prompts. Reach it through one key on Atlas Cloud, alongside 300+ models.

6 modelsExplore Wan 2.5
Wan
image-to-video

Wan

The Wan API brings Alibaba's open Wan video models to Atlas Cloud through one unified key. Wan 2.2 pioneered a Mixture-of-Experts architecture for video diffusion, lifting capacity and motion control at the same inference cost. It handles text to video, image to video, and video to video, with first-to-last frame control, extend, and upscaling, all reachable alongside 300+ models.

2 modelsExplore Wan

AI Image Models by Alibaba

Create stunning, production-ready visuals from prompts and references using state-of-the-art AI image generation models on Atlas Cloud.

View all models
Wan 2.7
image-to-video
text-to-video
video-to-video
image-to-image

Wan 2.7

The Wan 2.7 API gives developers Alibaba's all-in-one video suite, covering text to video, image to video, reference to video, and video editing, plus image generation. It produces native 1080p clips up to 15 seconds with synced audio, first to last frame control, and up to 5 character references. (See the latest [Wan 3.0 API page](https://www.atlascloud.ai/models/wan-3.0) for native 30-second video and up to 20 mixed references.)

8 modelsExplore Wan 2.7
Qwen Image
image-to-image
text-to-image

Qwen Image

The Qwen Image API brings Alibaba's Tongyi Qianwen image family into your product, from Standard and Professional text-to-image models to instruction-driven editing. Expect faithful English and Chinese text, reliable composition across dense layouts, and high-resolution output up to 2K. Atlas Cloud serves the full family through one unified API with transparent pay-as-you-go pricing from $0.035 per image and Day-0 access to new releases. Start building today.

16 modelsExplore Qwen Image
Wan 2.6
image-to-video
image-to-image
video-to-video
text-to-video

Wan 2.6

The Wan 2.6 API builds a multi-shot story from one prompt instead of a single clip. Alibaba's storytelling engine cuts between shots with coherent transitions while holding character identity, voice, and style steady across the scene. It spans text, image, and reference to video at 1080p, with synchronized audio and multilingual lip-sync. Reach it through one key on Atlas Cloud, next to 300+ models.

6 modelsExplore Wan 2.6
Wan 2.5
text-to-video
image-to-video
image-to-image
text-to-image

Wan 2.5

The Wan 2.5 API produces video and audio in one pass, so voice, sound, and lip-sync line up without a separate step. Built on Alibaba's Diffusion Transformer architecture, it covers text and image to video from 480p to 1080p at 5 or 10 seconds, with reliable sync even for Chinese prompts. Reach it through one key on Atlas Cloud, alongside 300+ models.

6 modelsExplore Wan 2.5

Large Language Models by Alibaba

Power chat, reasoning, and agents at scale with leading large language models, served fast and affordably on Atlas Cloud.

View all models

Alibaba Models API Pricing Details

Compare standard vs. our pricing across every Alibaba model.

ModelStandard Price (USD)Our Price (USD)Discount
Qwen3.6 Plus
$0.5/$3per 1M tokens1000K context
$0.325/$1.95M in/outper 1M tokens1000K context
-35%View
Qwen-Image Edit Plus 20251215$0.03/pic
Start from$0.021/pic
-30%View
Qwen Image Edit$0.045/pic
Start from$0.032/pic
-30%View
Qwen-Image Text-to-image Max$0.075/pic
Start from$0.052/pic
-30%View
Qwen-Image Text-to-image Plus$0.03/pic
Start from$0.021/pic
-30%View
Qwen-Image Edit$0.045/pic
Start from$0.032/pic
-30%View

Explore models from other providers

Instantly explore and experiment with 400+ production-ready models in the Atlas Playground. Start customizing with one click.

Alibaba API Use Cases You Can Build on Atlas Cloud

Qwen and Wan cover the full range of AI media production: language tasks, image generation, and video creation from a single API. Developers and teams use them together to build pipelines that go from text prompt to finished video without switching providers.

E-commerce Product Videos

E-commerce teams animate static product images into short video clips for product pages and social ads. Wan 2.7's image-to-video endpoint takes a product photo and a text prompt describing the motion, then outputs a 1080p clip with synchronized audio. At $0.10 per run on Atlas Cloud, teams can generate video variants for each SKU without a dedicated production budget.

Image-to-Video Production Pipeline

Developers build end-to-end pipelines that generate a scene image with Qwen Image 2.0, then pass it directly into Wan 2.7's image-to-video endpoint for animation. Both models run under the same Atlas Cloud API key, so no separate authentication or provider switching is needed. First-and-last frame control in Wan 2.7 lets you define exactly how each animated clip opens and closes.

Multilingual Document Processing

Enterprise teams use Qwen LLM to process documents, contracts, and scanned notes across 32 languages from a single API call. The model handles non-Latin scripts accurately, which makes it practical for Asian market operations where other LLMs fall short. Atlas Cloud delivers Qwen LLM access pay-as-you-go with no per-seat licensing.

Agentic Code Generation

Development teams use Qwen 3.6-Plus for complex coding tasks: breaking down repository-level problems, writing code, running tests, and iterating until completion. The model handles multi-step programming workflows without manual intervention between steps. Access via the OpenAI-compatible API on Atlas Cloud means it plugs into existing agent frameworks without custom integration work.

Scene-Directed Short Video

Wan 2.7's first-and-last frame control lets creators set the opening and closing frame of a clip, with the model handling motion in between. This gives video teams repeatable control over pacing and composition that text-only prompts can't match. Output reaches 1080p at up to 15 seconds per clip, suitable for social media and advertising formats.

Automated Video News Digest

Media teams build pipelines that convert daily text articles into short video segments using Wan 2.7's text-to-video endpoint. Each article becomes a standalone clip with generated visuals and audio, assembled into a digest without manual editing. The flat per-run pricing on Atlas Cloud keeps costs predictable when generating high volumes of clips on a daily schedule.

Render your enterprise vision into reality with Atlas Cloud AI.

Contact Sales

Frequently Asked Questions about Alibaba Models API

Atlas Cloud hosts the full Qwen and Wan model lineup under one API. This covers Wan 2.7, Wan 2.6, Wan 2.5, and Wan 2.2 for video generation, Qwen Image 2.0 for image generation, and Qwen LLM models for language tasks. All models are available pay-as-you-go with a single API key.

Yes. Atlas Cloud adds new Qwen and Wan versions as they launch. Wan 2.7, released in March 2026, was available on Atlas Cloud from day one. Check the Qwen and Wan collection page for the latest model versions as they go live.

Wan 2.7 accepts text prompts for text-to-video and reference images for image-to-video. First-and-last frame input is available when you need precise control over how a scene starts and ends. Instruction-based video editing is also supported for modifying existing clips through the API.

Wan 2.7 generates video at up to 1080p resolution with clip lengths from 2 to 15 seconds. Aspect ratio and duration are configurable per API call. Output includes synchronized audio generated alongside the video in a single request.

Yes. Qwen Image 2.0 outputs images up to 2K resolution, which you can pass directly into Wan 2.7's image-to-video endpoint. Both models run under the same Atlas Cloud API key. This makes it practical to build a full image-to-video pipeline without switching providers.

Yes. Atlas Cloud provides an OpenAI-compatible API at api.atlascloud.ai/v1. Swap the base URL in your existing OpenAI SDK setup and all Qwen LLM calls work without other changes. Video and image model endpoints follow the same authentication pattern.

Wan 2.7 supports 1080p output, multi-reference inputs, and instruction-based editing in a single API call, which covers most production requirements. The flat per-run pricing on Atlas Cloud makes cost predictable at volume. Pay-as-you-go access means no capacity commitments are needed to scale up.

Wan 2.7 video generation is charged at a flat $0.10 per run with no duration-based fees. Qwen Image models are priced at up to 30% below the official Alibaba API rate. All models are available pay-as-you-go with no monthly minimums.

Explore More Families

Seedance 2.5

Seedance 2.5 API is now available on Atlas Cloud! It gives developers ByteDance's newest video model. It generates up to 30 seconds of native video in a single pass from text, a single image, or as many as 50 multimodal references, with synchronized audio and in-frame multilingual text. On Atlas Cloud you reach it through one key, with subject consistency and improved physics keeping long shots coherent. (Update: Seedance 2.5 1080P API Is Available NOW!)

View Family

Wan 3.0

Wan 3.0 API is the next generation of Alibaba's Wan video family, built to push long-form generation, multi-reference control, and audiovisual quality to new heights. Atlas Cloud already hosts Wan 2.7, 2.6, and 2.5, and Wan 3.0 runs on the same unified key with no separate setup. Start building today. Scroll down to the showcase to see what Wan 3.0 can create.

View Family

MiniMax H3

The MiniMax H3 API opens MiniMax's general purpose multimodal video model, which reads text, images, video and audio as one context instead of one task at a time. Clips run 5 to 15 seconds at 24 FPS across aspect ratios from 21:9 to 9:16, and one prompt can swap characters, replace backgrounds, rewrite dialogue or clone a voice from a reference clip. Atlas Cloud serves it all through one OpenAI-compatible endpoint. Start building today.

View Family

Seedream 5.0 Pro

Seedream 5.0 Pro API gives developers ByteDance's controllable image editing model on Atlas Cloud. It places edits precisely with anchors and coordinates, separates images into editable layers, fuses multiple references, and matches exact colors and materials, with multilingual text at 2K and 3K. On Atlas Cloud you reach it through one key!

View Family

Seedance 2.0

The Seedance 2.0 API gives you production access to ByteDance's multimodal video model — quad-modal inputs (text, image, video, audio) and an industry-leading "Universal Reference" system that locks composition, camera movement, and character actions across shots. Integrate director-level control with one API call, a flat $0.09/s, instant key, and no waitlist — backed by enterprise-grade uptime and compliance. Seedance 2.0 Native 4K is now live!

View Family

GPT Image 2

The GPT Image 2 API gives developers access to OpenAI's latest image model, the successor to GPT Image 1.5. It generates and edits images with accurate text rendering across Latin and CJK scripts, plus strong composition for posters, mockups, and infographics. On Atlas Cloud you reach it through one unified API alongside 300+ models, with free credits, 99.99% uptime, and no OpenAI organization verification required.

View Family

Gemini Omni Flash

The Gemini Omni API brings Google DeepMind's multimodal video generation and editing model, introduced at Google I/O 2026, to your stack. Gemini Omni fuses Gemini's reasoning engine with generative media, accepting any mix of text, images, video, and audio to produce consistent, knowledge-grounded output. Refine results through natural conversation, swapping objects, rewriting scenes, and shifting styles while physics, characters, and continuity stay intact. Atlas Cloud serves the full Gemini Omni Flash lineup, text-to-video, image-to-video with up to 7 reference images, and reference-to-video, through one unified API with transparent per-second pricing from $0.112 and no subscription. Start building today.

View Family

Grok Imagine

The Grok Imagine API covers xAI's image, video, and speech models, from Image 2.0 to Video 1.5 and xAI TTS v1. Render 1K or 2K stills across 14 aspect ratios, push a scene to 15 seconds of 1080p motion, steer shots with up to 7 reference images, or narrate them in 20 languages. Atlas Cloud runs every mode on one endpoint, priced pay-as-you-go from $0.02 per image and $0.05 per second. Start building today.

View Family

Google

Google's most powerful creative models are all available on Atlas Cloud. Veo 3.1 delivers cinematic video generation, Nano Banana 2 powers high-fidelity image creation, and Gemini brings multimodal intelligence to every workflow. Access the full Google model suite through one API key with Day-0 availability and pay-as-you-go pricing.

View Family

Seedance 2.0 Mini

The Seedance 2.0 Mini API is the lightest, lowest-cost tier of ByteDance's Seedance video line, built for teams where throughput and unit cost matter more than maximum polish. Use it for batch generation, rapid prototyping, and draft passes, all through one OpenAI-compatible key on Atlas Cloud.

View Family

ByteDance

From cinematic video generation to high-fidelity image creation, ByteDance's most powerful models are live on Atlas Cloud. Run Seedance and Seedream at scale with the lowest inference pricing and zero infrastructure overhead.

View Family

Alibaba

Atlas Cloud brings together Alibaba's full model lineup under one API: Qwen for language and image tasks, Wan for video generation up to 1080p. Access every model pay-as-you-go with no subscriptions. The Alibaba API is available via a single base URL using your existing OpenAI-compatible client.

View Family

Recommended Articles

Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.