Alibaba Models on Atlas Cloud | Wan & Qwen

Atlas Cloud brings together Alibaba's full model lineup under one API: Qwen for language and image tasks, Wan for video generation up to 1080p. Access every model pay-as-you-go with no subscriptions. The Alibaba API is available via a single base URL using your existing OpenAI-compatible client.

Alibaba is developed by Alibaba. Atlas Cloud (operated by Atlas Cloud AI LLC) provides access to it and does not own it. All trademarks belong to their respective owners.

AI Video Models by Alibaba

Generate cinematic, high-fidelity videos from text and images with the latest AI video generation models on Atlas Cloud.

View all models
Wan 3.0
text-to-video
image-to-video
video-to-video

Wan 3.0

Wan 3.0 API is the next generation of Alibaba's Wan video family, built to push long-form generation, multi-reference control, and audiovisual quality to new heights. Atlas Cloud already hosts Wan 2.7, 2.6, and 2.5, and Wan 3.0 runs on the same unified key with no separate setup. Start building today. Scroll down to the showcase to see what Wan 3.0 can create.

6 modelsExplore Wan 3.0
Wan 2.7
image-to-video
text-to-video
video-to-video
image-to-image

Wan 2.7

The Wan 2.7 API gives developers Alibaba's all-in-one video suite, covering text to video, image to video, reference to video, and video editing, plus image generation. It produces native 1080p clips up to 15 seconds with synced audio, first to last frame control, and up to 5 character references. (See the latest [Wan 3.0 API page](https://www.atlascloud.ai/models/wan-3.0) for native 30-second video and up to 20 mixed references.)

8 modelsExplore Wan 2.7
Happy Horse
LLM
image-to-video
text-to-video
video-to-video

Happy Horse

The HappyHorse API on Atlas Cloud connects your application to Alibaba's HappyHorse 1.0 and 1.1 video generation models. Produce clips running anywhere from 3 to 15 seconds at 720p or 1080p, steer results with up to nine reference images, and pick the aspect ratio that fits your product. Atlas Cloud adds Day-0 model availability, reliable uptime, and a simple asynchronous REST integration. Start building today.

7 modelsExplore Happy Horse
Wan 2.6
image-to-video
image-to-image
video-to-video
text-to-video

Wan 2.6

The Wan 2.6 API builds a multi-shot story from one prompt instead of a single clip. Alibaba's storytelling engine cuts between shots with coherent transitions while holding character identity, voice, and style steady across the scene. It spans text, image, and reference to video at 1080p, with synchronized audio and multilingual lip-sync. Reach it through one key on Atlas Cloud, next to 300+ models.

6 modelsExplore Wan 2.6
Van
text-to-video
image-to-video

Van

Van is a flagship AI video series built on the Wan 2.5 and 2.6 frameworks, and the Van API unifies its full text-to-video and image-to-video lineup on Atlas Cloud. Render cinematic clips up to 1080p, then pick speed-optimized 2.6 models for fast iteration or cost-efficient 2.5 variants for large batches. Every model ships on Day-0 behind one OpenAI-compatible key with transparent pay-as-you-go pricing. Start building today.

4 modelsExplore Van
Wan 2.5
text-to-video
image-to-video
image-to-image
text-to-image

Wan 2.5

The Wan 2.5 API produces video and audio in one pass, so voice, sound, and lip-sync line up without a separate step. Built on Alibaba's Diffusion Transformer architecture, it covers text and image to video from 480p to 1080p at 5 or 10 seconds, with reliable sync even for Chinese prompts. Reach it through one key on Atlas Cloud, alongside 300+ models.

6 modelsExplore Wan 2.5
Wan
image-to-video

Wan

The Wan API brings Alibaba's open Wan video models to Atlas Cloud through one unified key. Wan 2.2 pioneered a Mixture-of-Experts architecture for video diffusion, lifting capacity and motion control at the same inference cost. It handles text to video, image to video, and video to video, with first-to-last frame control, extend, and upscaling, all reachable alongside 300+ models.

7 modelsExplore Wan

AI Image Models by Alibaba

Create stunning, production-ready visuals from prompts and references using state-of-the-art AI image generation models on Atlas Cloud.

View all models
Wan 2.7
image-to-video
text-to-video
video-to-video
image-to-image

Wan 2.7

The Wan 2.7 API gives developers Alibaba's all-in-one video suite, covering text to video, image to video, reference to video, and video editing, plus image generation. It produces native 1080p clips up to 15 seconds with synced audio, first to last frame control, and up to 5 character references. (See the latest [Wan 3.0 API page](https://www.atlascloud.ai/models/wan-3.0) for native 30-second video and up to 20 mixed references.)

8 modelsExplore Wan 2.7
Qwen Image 3.0 Pro
text-to-image
image-to-image

Qwen Image 3.0 Pro

Qwen Image 3.0 pro is Qwen's image generation and editing family for developers building production visual workflows. It creates images up to 2048×2048, rewrites prompts automatically, selects resolution from prompt intent, and follows natural language editing instructions while preserving facial features and identity. Access both models through Atlas Cloud with one OpenAI compatible key and standard pay as you go pricing of $0.04 per call. Start building today.

2 modelsExplore Qwen Image 3.0 Pro
Qwen Image 3.0
text-to-image
image-to-image

Qwen Image 3.0

Qwen Image 3.0 is Qwen's image generation and editing family for developers building visual products. It combines precise prompt adherence and complex text rendering with automatic prompt rewriting and prompt guided resolution selection. Through Atlas Cloud, developers can access both generation and editing models with standard pay as you go pricing of $0.03 per call. Start building today.

2 modelsExplore Qwen Image 3.0
Qwen Image 2.0 Pro
text-to-image
image-to-image

Qwen Image 2.0 Pro

Qwen Image 2.0 pro comes from the Qwen model family and serves professional image generation and editing workflows. Advanced prompt and instruction understanding turns detailed creative direction into high quality new images or precise edits. Atlas Cloud provides ready-to-use REST inference endpoints for both models at the standard price of $0.075 per call, simplifying integration and cost planning. Start building today.

2 modelsExplore Qwen Image 2.0 Pro
Qwen Image 2.0
text-to-image
image-to-image

Qwen Image 2.0

Qwen Image 2.0 is the next generation image model from the Qwen team at Alibaba Cloud. Built for text and image inputs, it delivers native 2K output, stronger semantic adherence, realistic detail, and precise instruction following across creation and revision workflows. Atlas Cloud provides ready to use REST endpoints for both modes with no cold starts and a standard price of $0.035 per image. Start building today.

2 modelsExplore Qwen Image 2.0
Qwen Image
image-to-image
text-to-image

Qwen Image

Qwen Image is Alibaba's Tongyi Qianwen family for image generation and editing, built to render complex Chinese and English text across varied visual styles. Its 20B MMDiT models support detailed creation, in image text changes, object manipulation, action changes, style transfer, and detail enhancement. On Atlas Cloud, access generation and editing endpoints at standard rates from $0.03 per image. Start today.

8 modelsExplore Qwen Image
Wan 2.6
image-to-video
image-to-image
video-to-video
text-to-video

Wan 2.6

The Wan 2.6 API builds a multi-shot story from one prompt instead of a single clip. Alibaba's storytelling engine cuts between shots with coherent transitions while holding character identity, voice, and style steady across the scene. It spans text, image, and reference to video at 1080p, with synchronized audio and multilingual lip-sync. Reach it through one key on Atlas Cloud, next to 300+ models.

6 modelsExplore Wan 2.6
Wan 2.5
text-to-video
image-to-video
image-to-image
text-to-image

Wan 2.5

The Wan 2.5 API produces video and audio in one pass, so voice, sound, and lip-sync line up without a separate step. Built on Alibaba's Diffusion Transformer architecture, it covers text and image to video from 480p to 1080p at 5 or 10 seconds, with reliable sync even for Chinese prompts. Reach it through one key on Atlas Cloud, alongside 300+ models.

6 modelsExplore Wan 2.5

Large Language Models by Alibaba

Power chat, reasoning, and agents at scale with leading large language models, served fast and affordably on Atlas Cloud.

View all models

Alibaba Models API Pricing Details

Compare standard vs. our pricing across every Alibaba model.

ModelStandard Price (USD)Our Price (USD)Discount
Qwen3.6 Plus
$0.5/$3per 1M tokens1000K context
$0.325/$1.95M in/outper 1M tokens1000K context
-35%View
Qwen-Image Edit Plus 20251215$0.03/pic
Start from$0.021/pic
-30%View
Qwen Image Edit$0.045/pic
Start from$0.032/pic
-30%View
Qwen-Image Text-to-image Max$0.075/pic
Start from$0.052/pic
-30%View
Qwen-Image Text-to-image Plus$0.03/pic
Start from$0.021/pic
-30%View
Qwen-Image Edit$0.045/pic
Start from$0.032/pic
-30%View

Explore models from other providers

Instantly explore and experiment with 400+ production-ready models in the Atlas Playground. Start customizing with one click.

Alibaba API Use Cases You Can Build on Atlas Cloud

Qwen and Wan cover the full range of AI media production: language tasks, image generation, and video creation from a single API. Developers and teams use them together to build pipelines that go from text prompt to finished video without switching providers.

E-commerce Product Videos

E-commerce teams animate static product images into short video clips for product pages and social ads. Wan 2.7's image-to-video endpoint takes a product photo and a text prompt describing the motion, then outputs a 1080p clip with synchronized audio. At $0.10 per run on Atlas Cloud, teams can generate video variants for each SKU without a dedicated production budget.

Image-to-Video Production Pipeline

Developers build end-to-end pipelines that generate a scene image with Qwen Image 2.0, then pass it directly into Wan 2.7's image-to-video endpoint for animation. Both models run under the same Atlas Cloud API key, so no separate authentication or provider switching is needed. First-and-last frame control in Wan 2.7 lets you define exactly how each animated clip opens and closes.

Multilingual Document Processing

Enterprise teams use Qwen LLM to process documents, contracts, and scanned notes across 32 languages from a single API call. The model handles non-Latin scripts accurately, which makes it practical for Asian market operations where other LLMs fall short. Atlas Cloud delivers Qwen LLM access pay-as-you-go with no per-seat licensing.

Agentic Code Generation

Development teams use Qwen 3.6-Plus for complex coding tasks: breaking down repository-level problems, writing code, running tests, and iterating until completion. The model handles multi-step programming workflows without manual intervention between steps. Access via the OpenAI-compatible API on Atlas Cloud means it plugs into existing agent frameworks without custom integration work.

Scene-Directed Short Video

Wan 2.7's first-and-last frame control lets creators set the opening and closing frame of a clip, with the model handling motion in between. This gives video teams repeatable control over pacing and composition that text-only prompts can't match. Output reaches 1080p at up to 15 seconds per clip, suitable for social media and advertising formats.

Automated Video News Digest

Media teams build pipelines that convert daily text articles into short video segments using Wan 2.7's text-to-video endpoint. Each article becomes a standalone clip with generated visuals and audio, assembled into a digest without manual editing. The flat per-run pricing on Atlas Cloud keeps costs predictable when generating high volumes of clips on a daily schedule.

Render your enterprise vision into reality with Atlas Cloud AI.

Contact Sales

Frequently Asked Questions about Alibaba Models API

Atlas Cloud hosts the full Qwen and Wan model lineup under one API. This covers Wan 2.7, Wan 2.6, Wan 2.5, and Wan 2.2 for video generation, Qwen Image 2.0 for image generation, and Qwen LLM models for language tasks. All models are available pay-as-you-go with a single API key.

Yes. Atlas Cloud adds new Qwen and Wan versions as they launch. Wan 2.7, released in March 2026, was available on Atlas Cloud from day one. Check the Qwen and Wan collection page for the latest model versions as they go live.

Wan 2.7 accepts text prompts for text-to-video and reference images for image-to-video. First-and-last frame input is available when you need precise control over how a scene starts and ends. Instruction-based video editing is also supported for modifying existing clips through the API.

Wan 2.7 generates video at up to 1080p resolution with clip lengths from 2 to 15 seconds. Aspect ratio and duration are configurable per API call. Output includes synchronized audio generated alongside the video in a single request.

Yes. Qwen Image 2.0 outputs images up to 2K resolution, which you can pass directly into Wan 2.7's image-to-video endpoint. Both models run under the same Atlas Cloud API key. This makes it practical to build a full image-to-video pipeline without switching providers.

Yes. Atlas Cloud provides an OpenAI-compatible API at api.atlascloud.ai/v1. Swap the base URL in your existing OpenAI SDK setup and all Qwen LLM calls work without other changes. Video and image model endpoints follow the same authentication pattern.

Wan 2.7 supports 1080p output, multi-reference inputs, and instruction-based editing in a single API call, which covers most production requirements. The flat per-run pricing on Atlas Cloud makes cost predictable at volume. Pay-as-you-go access means no capacity commitments are needed to scale up.

Wan 2.7 video generation is charged at a flat $0.10 per run with no duration-based fees. Qwen Image models are priced at up to 30% below the official Alibaba API rate. All models are available pay-as-you-go with no monthly minimums.

Explore More Families

Seedance 2.5

Seedance 2.5 API is now available on Atlas Cloud! It gives developers ByteDance's newest video model. It generates up to 30 seconds of native video in a single pass from text, a single image, or as many as 50 multimodal references, with synchronized audio and in-frame multilingual text. On Atlas Cloud you reach it through one key, with subject consistency and improved physics keeping long shots coherent. (Update: Seedance 2.5 1080P API Is Available NOW!)

View Family

Wan 3.0

Wan 3.0 API is the next generation of Alibaba's Wan video family, built to push long-form generation, multi-reference control, and audiovisual quality to new heights. Atlas Cloud already hosts Wan 2.7, 2.6, and 2.5, and Wan 3.0 runs on the same unified key with no separate setup. Start building today. Scroll down to the showcase to see what Wan 3.0 can create.

View Family

MiniMax H3

MiniMax H3 is MiniMax's multimodal video family for text, image, and reference guided creation. Across supported routes, it preserves subjects from reference media, offers flexible aspect ratios, and pairs generated sound with visuals through H3 Developer, with output profiles selected by endpoint. Atlas Cloud unifies the family behind one OpenAI-compatible key with transparent pay-as-you-go pricing from the standard rate of $0.038 per second. Start building today.

View Family

Seedream 5.0 Pro

Seedream 5.0 Pro API gives developers ByteDance's controllable image editing model on Atlas Cloud. It places edits precisely with anchors and coordinates, separates images into editable layers, fuses multiple references, and matches exact colors and materials, with multilingual text at 2K and 3K. On Atlas Cloud you reach it through one key!

View Family

Seedance 2.0

Seedance 2.0 is ByteDance’s production video model for precise shot creation. Turn prompts into video, animate a first-frame image with optional last-frame guidance, or shape results with reference media and optional web search. Atlas Cloud brings these workflows into one unified API with transparent pay-as-you-go pricing and one OpenAI-compatible key. Start building today.

View Family

GPT Image 2.5

The gpt-image-2.5 family from OpenAI gives developers a choice of Flare and Sunburst for production image workflows. Render at arbitrary resolutions up to 3840x2160 and select from five quality tiers, including xhigh and max, to match specific output requirements. Atlas Cloud provides ready-to-use REST inference with no cold starts and standard pricing from $0.004 per generation. Start building today.

View Family

GPT Image 2

The GPT Image 2 API gives developers access to OpenAI's latest image model, the successor to GPT Image 1.5. It generates and edits images with accurate text rendering across Latin and CJK scripts, plus strong composition for posters, mockups, and infographics. On Atlas Cloud you reach it through one unified API alongside 300+ models, with free credits, 99.99% uptime, and no OpenAI organization verification required.

View Family

Gemini Omni Flash

The gemini omni API brings Google DeepMind's natively multimodal Gemini Omni Flash family, including Gemini Omni 1.1 Flash, to developers. Create cinematic video with synchronized native audio, animate still images with precise start and end frame control, or revise existing footage through text guided edits that preserve untouched content. Atlas Cloud provides one OpenAI-compatible key, unified access, and transparent pay-as-you-go pricing. Start building today.

View Family

Grok Imagine

Grok Imagine Image is xAI's family for generating polished visuals and revising one or more reference images through natural language instructions. Its standard and quality endpoints cover text to image creation, single image changes, and indexed multi-image composition. Atlas Cloud brings these workflows into one API, with standard generation and editing priced at $0.02 per image. Start building today.

View Family

Google

Google's most powerful creative models are all available on Atlas Cloud. Veo 3.1 delivers cinematic video generation, Nano Banana 2 powers high-fidelity image creation, and Gemini brings multimodal intelligence to every workflow. Access the full Google model suite through one API key with Day-0 availability and pay-as-you-go pricing.

View Family

Seedance 2.0 Mini

The Seedance 2.0 Mini API is the lightest, lowest-cost tier of ByteDance's Seedance video line, built for teams where throughput and unit cost matter more than maximum polish. Use it for batch generation, rapid prototyping, and draft passes, all through one OpenAI-compatible key on Atlas Cloud.

View Family

ByteDance

From cinematic video generation to high-fidelity image creation, ByteDance's most powerful models are live on Atlas Cloud. Run Seedance and Seedream at scale with the lowest inference pricing and zero infrastructure overhead.

View Family

Recommended Articles

Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.