alibaba/wan-2.5/text-to-video

texte-vers-vidéo

Wan 2.5 Text-to-Video API by Alibaba

alibaba/wan-2.5/text-to-video

Text-to-video

A speed-optimized text-to-video option that prioritizes lower latency while retaining strong visual fidelity. Ideal for iteration, batch generation, and prompt testing.

Entrée

Prompt *

Prompt Négatif

Audio

Vous pouvez glisser-déposer le fichier ou cliquer pour télécharger

MAX:1

Taille

Durée

Expansion du Prompt

Générer l'Audio

Graine

Sortie

Inactif

Les vidéos générées apparaîtront ici

Configurez vos paramètres et cliquez sur exécuter pour commencer

Votre requête coûtera $0.035 par exécution. Avec $10, vous pouvez exécuter ce modèle environ 285 fois.

Vous pouvez continuer avec :

Seedance 2.0 Kling v3 Vidu Wan2.7

Paramètres

Exemple de code
import requests
import time

# Step 1: Start video generation
generate_url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "alibaba/wan-2.5/text-to-video",  # Required. model name
    "audio": "example_value",  # Audio URL to guide generation (optional)
    "duration": 5,  # The duration of the generated media in seconds. options: 5 | 10
    "enable_prompt_expansion": False,  # If set to true, the prompt optimizer will be enabled
    "negative_prompt": "example_value",  # Negative prompt for the generation
    "prompt": "A beautiful sunset over the ocean with gentle waves",  # Required. The prompt for generating the output
    "seed": -1,  # The random seed to use for the generation
    "size": "1920*1080",  # Required. The size of the generated media in pixels (width*height)
    "generate_audio": True,  # Whether to automatically add audio to the generated video
}

generate_response = requests.post(generate_url, headers=headers, json=data)
generate_result = generate_response.json()
prediction_id = generate_result["data"]["id"]

# Step 2: Poll for result
poll_url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"

def check_status():
    while True:
        response = requests.get(poll_url, headers={"Authorization": "Bearer $ATLASCLOUD_API_KEY"})
        result = response.json()

        if result["data"]["status"] in ["completed", "succeeded"]:
            print("Generated video:", result["data"]["outputs"][0])
            return result["data"]["outputs"][0]
        elif result["data"]["status"] == "failed":
            raise Exception(result["data"]["error"] or "Generation failed")
        else:
            # Still processing, wait 2 seconds
            time.sleep(2)

video_url = check_status()

Installer

Installez le package requis pour votre langage.

pip install requests

Authentification

Toutes les requêtes API nécessitent une authentification via une clé API. Vous pouvez obtenir votre clé API depuis le tableau de bord Atlas Cloud.

export ATLASCLOUD_API_KEY="your-api-key-here"

En-têtes HTTP

import os

API_KEY = os.environ.get("ATLASCLOUD_API_KEY")
headers = {
    "Content-Type": "application/json",
    "Authorization": f"Bearer {API_KEY}"
}

Protégez votre clé API

N'exposez jamais votre clé API dans du code côté client ou dans des dépôts publics. Utilisez plutôt des variables d'environnement ou un proxy backend.

Soumettre une requête

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "your-model",
    "prompt": "A beautiful landscape"
}

response = requests.post(url, headers=headers, json=data)
print(response.json())

Soumettre une requête

Soumettez une requête de génération asynchrone. L'API renvoie un identifiant de prédiction que vous pouvez utiliser pour vérifier le statut et récupérer le résultat.

POST/api/v1/model/generateVideo

Corps de la requête

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}

data = {
    "model": "alibaba/wan-2.5/text-to-video",
    "prompt": "A beautiful sunset over the ocean with gentle waves"
}

response = requests.post(url, headers=headers, json=data)
result = response.json()

print(f"Prediction ID: {result['data']['id']}")
print(f"Status: {result['data']['status']}")

Réponse

{
  "code": 200,
  "data": {
    "id": "pred_abc123",
    "status": "processing",
    "model": "model-name",
    "created_at": "2025-01-01T00:00:00Z"
  }
}

Vérifier le statut

Interrogez le point de terminaison de prédiction pour vérifier le statut actuel de votre requête.

GET/api/v1/model/prediction/{prediction_id}

Exemple d'interrogation

import requests
import time

prediction_id = "pred_abc123"
url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

while True:
    response = requests.get(url, headers=headers)
    result = response.json()
    status = result["data"]["status"]
    print(f"Status: {status}")

    if status in ["completed", "succeeded"]:
        output_url = result["data"]["outputs"][0]
        print(f"Output URL: {output_url}")
        break
    elif status == "failed":
        print(f"Error: {result['data'].get('error', 'Unknown')}")
        break

    time.sleep(3)

Valeurs de statut

processingLa requête est encore en cours de traitement.

completedLa génération est terminée. Les résultats sont disponibles.

succeededLa génération a réussi. Les résultats sont disponibles.

failedLa génération a échoué. Vérifiez le champ d'erreur.

Réponse terminée

{
  "data": {
    "id": "pred_abc123",
    "status": "completed",
    "outputs": [
      "https://storage.atlascloud.ai/outputs/result.mp4"
    ],
    "metrics": {
      "predict_time": 45.2
    },
    "created_at": "2025-01-01T00:00:00Z",
    "completed_at": "2025-01-01T00:00:10Z"
  }
}

Téléverser des fichiers

Téléversez des fichiers vers le stockage Atlas Cloud et obtenez une URL utilisable dans vos requêtes API. Utilisez multipart/form-data pour le téléversement.

POST/api/v1/model/uploadMedia

Exemple de téléversement

import requests

url = "https://api.atlascloud.ai/api/v1/model/uploadMedia"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

with open("image.png", "rb") as f:
    files = {"file": ("image.png", f, "image/png")}
    response = requests.post(url, headers=headers, files=files)

result = response.json()
download_url = result["data"]["download_url"]
print(f"File URL: {download_url}")

Réponse

{
  "data": {
    "download_url": "https://storage.atlascloud.ai/uploads/abc123/image.png",
    "file_name": "image.png",
    "content_type": "image/png",
    "size": 1024000
  }
}

Schema d'entrée

Les paramètres suivants sont acceptés dans le corps de la requête.

Total: 9Requis: 3Optionnel: 6

modelstringrequired

model name

Default: "alibaba/wan-2.5/text-to-video"

audiostring

Audio URL to guide generation (optional).

durationinteger

The duration of the generated media in seconds.

Default: 5

510

enable_prompt_expansionboolean

If set to true, the prompt optimizer will be enabled.

Default: false

negative_promptstring

Negative prompt for the generation.

promptstringrequired

The prompt for generating the output.

seedinteger

The random seed to use for the generation. -1 means a random seed will be used.

Default: -1

sizestringrequired

The size of the generated media in pixels (width*height).

Default: "1920*1080"

832*480480*832624*6241280*720720*1280960*9601088*832832*10881920*10801080*19201440*14401632*12481248*1632

generate_audioboolean

Whether to automatically add audio to the generated video.

Default: true

Exemple de corps de requête

{
  "model": "alibaba/wan-2.5/text-to-video",
  "duration": 5,
  "enable_prompt_expansion": false,
  "prompt": "A beautiful landscape",
  "seed": -1,
  "size": "1920*1080",
  "generate_audio": true
}

Schema de sortie

L'API renvoie une réponse de prédiction avec les URL des résultats générés.

created_atstring

ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”).

idstring

Unique identifier for the prediction, the ID of the prediction to get.

modelstring

Model ID used for the prediction.

outputsarray

Array of URLs to the generated content (empty when status is not completed).

statusstring

Status of the task: created, processing, completed, or failed.

Exemple de réponse

{
  "id": "pred_abc123",
  "status": "completed",
  "model": "model-name",
  "outputs": [
    "https://storage.atlascloud.ai/outputs/result.mp4"
  ],
  "metrics": {
    "predict_time": 45.2
  },
  "created_at": "2025-01-01T00:00:00Z",
  "completed_at": "2025-01-01T00:00:10Z"
}

Atlas Cloud Skills

Atlas Cloud Skills intègre plus de 400 modèles d'IA directement dans votre assistant de codage IA. Une seule commande pour installer, puis utilisez le langage naturel pour générer des images, des vidéos et discuter avec des LLM.

Clients pris en charge

Claude Code

OpenAI Codex

Gemini CLI

Cursor

Windsurf

VS Code

Trae

GitHub Copilot

Cline

Roo Code

Amp

Goose

Replit

40+ clients pris en charge

Installer

npx skills add AtlasCloudAI/atlas-cloud-skills

Configurer la clé API

Obtenez votre clé API depuis le tableau de bord Atlas Cloud et définissez-la comme variable d'environnement.

export ATLASCLOUD_API_KEY="your-api-key-here"

Fonctionnalités

Une fois installé, vous pouvez utiliser le langage naturel dans votre assistant IA pour accéder à tous les modèles Atlas Cloud.

Génération d'imagesGénérez des images avec des modèles comme Nano Banana 2, Z-Image, et plus encore.

Création de vidéosCréez des vidéos à partir de texte ou d'images avec Kling, Vidu, Veo, etc.

Chat LLMDiscutez avec Qwen, DeepSeek et d'autres grands modèles de langage.

Téléversement de médiasTéléversez des fichiers locaux pour l'édition d'images et les workflows image-vers-vidéo.

En savoir plus

github.com/AtlasCloudAI/atlas-cloud-skills

Serveur MCP

Le serveur MCP Atlas Cloud connecte votre IDE avec plus de 400 modèles d'IA via le Model Context Protocol. Compatible avec tout client compatible MCP.

Clients pris en charge

Cursor

VS Code

Windsurf

Claude Code

OpenAI Codex

Gemini CLI

Cline

Roo Code

100+ clients pris en charge

Installer

npx -y atlascloud-mcp

Configuration

Ajoutez la configuration suivante au fichier de paramètres MCP de votre IDE.

{
  "mcpServers": {
    "atlascloud": {
      "command": "npx",
      "args": [
        "-y",
        "atlascloud-mcp"
      ],
      "env": {
        "ATLASCLOUD_API_KEY": "your-api-key-here"
      }
    }
  }
}

Outils disponibles

atlas_generate_imageGénérez des images à partir de prompts textuels.

atlas_generate_videoCréez des vidéos à partir de texte ou d'images.

atlas_chatDiscutez avec de grands modèles de langage.

atlas_list_modelsParcourez plus de 400 modèles d'IA disponibles.

atlas_quick_generateCréation de contenu en une étape avec sélection automatique du modèle.

atlas_upload_mediaTéléversez des fichiers locaux pour les workflows API.

En savoir plus

github.com/AtlasCloudAI/mcp-server

Schéma API

{
  "info": {
    "title": "AtlasCloud API",
    "version": "1.0.0",
    "description": "The AtlasCloud API."
  },
  "openapi": "3.0.0",
  "paths": {
    "/api/v1/model/generateVideo": {
      "post": {
        "requestBody": {
          "content": {
            "application/json": {
              "schema": {
                "$ref": "#/components/schemas/Input"
              }
            }
          },
          "required": true
        },
        "responses": {
          "200": {
            "content": {
              "application/json": {
                "schema": {
                  "$ref": "#/components/schemas/PredictionResponse"
                }
              }
            },
            "description": "The request status."
          }
        }
      },
      "x-api-name": "model_run"
    },
    "/api/v1/model/result/{request_id}": {
      "get": {
        "parameters": [
          {
            "in": "path",
            "name": "request_id",
            "required": true,
            "schema": {
              "description": "Request ID",
              "type": "string"
            }
          }
        ],
        "responses": {
          "200": {
            "content": {
              "application/json": {
                "schema": {
                  "$ref": "#/components/schemas/PredictionResponse"
                }
              }
            },
            "description": "Result of the request."
          }
        }
      },
      "x-api-name": "model_result"
    }
  },
  "components": {
    "schemas": {
      "Input": {
        "properties": {
          "model": {
            "type": "string",
            "description": "model name",
            "default": "alibaba/wan-2.5/text-to-video"
          },
          "audio": {
            "description": "Audio URL to guide generation (optional).",
            "type": "string"
          },
          "duration": {
            "default": 5,
            "description": "The duration of the generated media in seconds.",
            "enum": [
              5,
              10
            ],
            "type": "integer",
            "x-ui-component": "select"
          },
          "enable_prompt_expansion": {
            "default": false,
            "description": "If set to true, the prompt optimizer will be enabled.",
            "type": "boolean"
          },
          "negative_prompt": {
            "description": "Negative prompt for the generation.",
            "type": "string"
          },
          "prompt": {
            "description": "The prompt for generating the output.",
            "type": "string",
            "x-rows": 10,
            "x-ui-component": "textarea"
          },
          "seed": {
            "default": -1,
            "description": "The random seed to use for the generation. -1 means a random seed will be used.",
            "type": "integer"
          },
          "size": {
            "default": "1920*1080",
            "description": "The size of the generated media in pixels (width*height).",
            "enum": [
              "832*480",
              "480*832",
              "624*624",
              "1280*720",
              "720*1280",
              "960*960",
              "1088*832",
              "832*1088",
              "1920*1080",
              "1080*1920",
              "1440*1440",
              "1632*1248",
              "1248*1632"
            ],
            "type": "string"
          },
          "generate_audio": {
            "default": true,
            "description": "Whether to automatically add audio to the generated video.",
            "type": "boolean"
          }
        },
        "required": [
          "model",
          "prompt",
          "size"
        ],
        "type": "object",
        "x-order-properties": [
          "model",
          "prompt",
          "negative_prompt",
          "audio",
          "size",
          "duration",
          "enable_prompt_expansion",
          "generate_audio",
          "seed"
        ]
      },
      "PredictionResponse": {
        "properties": {
          "created_at": {
            "description": "ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”).",
            "format": "date-time",
            "type": "string"
          },
          "has_nsfw_contents": {
            "description": "Array of boolean values indicating NSFW detection for each output.",
            "items": {
              "type": "boolean"
            },
            "type": "array"
          },
          "id": {
            "description": "Unique identifier for the prediction, the ID of the prediction to get.",
            "type": "string"
          },
          "model": {
            "description": "Model ID used for the prediction.",
            "type": "string"
          },
          "outputs": {
            "description": "Array of URLs to the generated content (empty when status is not completed).",
            "items": {
              "type": "object"
            },
            "type": "array"
          },
          "status": {
            "description": "Status of the task: created, processing, completed, or failed.",
            "type": "string"
          },
          "urls": {
            "description": "Object containing related API endpoints.",
            "type": "object"
          }
        },
        "type": "object"
      }
    }
  },
  "servers": [
    {
      "url": "https://api.atlascloud.ai"
    }
  ]
}

Template de Prompt pour LLM

# alibaba/wan-2.5/text-to-video

> A speed-optimized text-to-video option that prioritizes lower latency while retaining strong visual fidelity. Ideal for iteration, batch generation, and prompt testing.


## Overview

- **Submit endpoint (POST)**: `https://api.atlascloud.ai/api/v1/model/generateVideo` — start an async generation; returns a `prediction_id`
- **Poll endpoint (GET)**: `https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}` — poll this until the prediction finishes
- **Model ID**: `alibaba/wan-2.5/text-to-video`


## API Information

This model can be used via our HTTP API or more conveniently via our client libraries.
See the input and output schema below, as well as the usage examples.


### Input Schema

The API accepts the following input parameters:

- **`model`** (`string`, _required_):
  model name
  - Default: `"alibaba/wan-2.5/text-to-video"`

- **`prompt`** (`string`, _required_):
  The prompt for generating the output.

- **`negative_prompt`** (`string`, _optional_):
  Negative prompt for the generation.

- **`audio`** (`string`, _optional_):
  Audio URL to guide generation (optional).

- **`size`** (`string`, _required_):
  The size of the generated media in pixels (width*height).
  - Default: `"1920*1080"`
  - Options: "832*480", "480*832", "624*624", "1280*720", "720*1280", "960*960", "1088*832", "832*1088", "1920*1080", "1080*1920", "1440*1440", "1632*1248", "1248*1632"

- **`duration`** (`integer`, _optional_):
  The duration of the generated media in seconds.
  - Default: `5`
  - Options: 5, 10

- **`enable_prompt_expansion`** (`boolean`, _optional_):
  If set to true, the prompt optimizer will be enabled.
  - Default: `false`

- **`generate_audio`** (`boolean`, _optional_):
  Whether to automatically add audio to the generated video.
  - Default: `true`

- **`seed`** (`integer`, _optional_):
  The random seed to use for the generation. -1 means a random seed will be used.
  - Default: `-1`



**Required Parameters Example**:

```json
{
  "model": "alibaba/wan-2.5/text-to-video",
  "prompt": "",
  "size": "1920*1080"
}
```


**Full Example**:

```json
{
  "model": "alibaba/wan-2.5/text-to-video",
  "prompt": "",
  "negative_prompt": "",
  "audio": "",
  "size": "1920*1080",
  "duration": 5,
  "enable_prompt_expansion": false,
  "generate_audio": true,
  "seed": -1
}
```


### Output Schema

The API returns the following output format:


- **`created_at`** (`string`, _optional_):
  ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”).

- **`has_nsfw_contents`** (`array[boolean]`, _optional_):
  Array of boolean values indicating NSFW detection for each output.

- **`id`** (`string`, _optional_):
  Unique identifier for the prediction, the ID of the prediction to get.

- **`model`** (`string`, _optional_):
  Model ID used for the prediction.

- **`outputs`** (`array[object]`, _optional_):
  Array of URLs to the generated content (empty when status is not completed).

- **`status`** (`string`, _optional_):
  Status of the task: created, processing, completed, or failed.

- **`urls`** (`object`, _optional_):
  Object containing related API endpoints.



**Example Response**:

```json
{
  "created_at": "",
  "has_nsfw_contents": [],
  "id": "",
  "model": "",
  "outputs": [],
  "status": "",
  "urls": {}
}
```


## Usage Examples

### cURL

```bash
# Step 1: Start generation (async)
curl -X POST "https://api.atlascloud.ai/api/v1/model/generateVideo" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "alibaba/wan-2.5/text-to-video",
  "prompt": "",
  "negative_prompt": "",
  "audio": "",
  "size": "1920*1080",
  "duration": 5,
  "enable_prompt_expansion": false,
  "generate_audio": true,
  "seed": -1
}'

# Response will contain: {"code": 200, "data": {"id": "prediction_id", "status": "processing"}}

# Step 2: Poll for result (replace {prediction_id} with the id returned above)
curl -X GET "https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY"

# Keep polling until status is "completed", "succeeded" or "failed"
# When completed, outputs will contain the generated content URL(s)
```

## Additional Resources

### Documentation

- [Model Playground](https://www.atlascloud.ai/models/alibaba/wan-2.5/text-to-video)

A middle-aged man sitting at a wooden desk in a cozy study room, surrounded by bookshelves and a warm lamp glow. He opens an old book and reads aloud with a calm, deep voice: 'History teaches us more than just facts… it shows us who we are.' The room has subtle background sounds: pages turning, the faint ticking of a clock, and distant rain against the window.

A young man in his early 30s sits in a modern studio, wearing a navy blazer and white shirt. Soft lighting illuminates his face. He speaks directly to the camera, his lips moving naturally as he says: “Welcome to today’s interview. We’re going to explore how AI is changing our daily lives.” His gestures are subtle, occasionally raising his hands for emphasis, creating a professional and engaging tone.

A cinematic opening sequence of a sci-fi movie: a spaceship travels across the galaxy, and the movie title “星河远征 · Galactic Odyssey” emerges in golden 3D letters, with flawless kerning and no distortion, floating stably in space as the camera rotates.

A handsome, muscular man with well-defined abs is catching his breath after an intense workout. Sweat drips down his torso. He is shirtless, wearing only black athletic shorts, and is leaning against gym equipment. The lighting comes from the upper side, highlighting the contours of his chest and arms. The scene is filled with a raw, masculine energy, hyper-realistic, high-contrast lighting.

A graceful ballerina with her hair in a messy bun, performing a powerful and emotional contemporary ballet routine. She is in a minimalist, dark art studio. Abstract patterns of light and shadow, projected from a hidden source, dance across her body and the surrounding walls, constantly shifting with her movements. The camera focuses on the tension in her muscles and the expressive gestures of her hands. A single, dramatic slow-motion shot captures her mid-air leap, with the light patterns swirling around her like a galaxy. Moody, artistic, high contrast.

A young couple sitting on a park bench during sunset. The woman leans her head on the man’s shoulder. He whispers softly: 'No matter where we go, I’ll always be here with you.' The sound includes the rustling of leaves, distant laughter of children playing, and the gentle hum of cicadas in the evening air.

A low-angle panning shot of a concrete wall under a highway overpass at night. Graffiti of a young man comes to life and starts rapping. The style is a dynamic blend of 2D street art animation on a realistic, dark, cinematic background. Cityscape is visible in the distance.

A 3D animated, anthropomorphic badger wearing a brown leather vest is angrily sweeping yellow autumn leaves from the doorway of his rustic wooden cabin. The style is reminiscent of a Pixar film, with detailed fur and expressive animation. Sunny day, lush green meadow with a forest in the background.

Chargement...

Pourquoi choisir Wan 2.5 ?

Plus abordable

Malgré les récentes baisses de prix de Google, Veo 3 reste globalement onéreux. Wan 2.5 est léger et économique, offrant aux créateurs davantage d'options tout en réduisant considérablement les coûts de production.

Génération en une étape, synchronisation de bout en bout

Avec Wan 2.5, inutile d'enregistrer la voix séparément ou d'aligner manuellement les lèvres. Fournissez simplement un prompt clair et structuré pour générer en une seule fois des vidéos complètes avec audio/voix off et synchronisation labiale — plus rapide et plus simple.

Multilingue et accessible

Lorsque les prompts sont en chinois, Wan 2.5 génère de manière fiable des vidéos audio-visuelles synchronisées. En revanche, Veo 3 affiche souvent « langue inconnue » pour les prompts en chinois.

Recréation précise des personnages

Wan 2.5 excelle dans la restitution des traits des personnages, en reproduisant fidèlement leur apparence, leurs expressions et leurs styles de mouvement, rendant les personnages générés plus reconnaissables et personnalisés pour une narration et une immersion renforcées.

Rendu de style artistique

Prend en charge le rendu de style Studio Ghibli, créant des textures aquarelles peintes à la main et des effets d'animation. Offre des expériences visuelles chaleureuses et oniriques qui renforcent l'attrait artistique et la profondeur narrative.

Qui peut en bénéficier ?

Équipes marketing

Qu'il s'agisse de lancements de produits, de campagnes promotionnelles ou de marketing de marque, Wan 2.5 vous aide à générer rapidement des vidéos de haute qualité, rendant la création simple et efficace.

Démonstrations produits et tutoriels sans tracas de coordination
Marketing sur les réseaux sociaux avec sous-titres multilingues et synchronisation labiale
Le contenu généré par IA libère les équipes pour se concentrer sur la stratégie et la créativité

Bottom line: En résumé : La création n'a jamais été aussi simple, rapide et intelligente — Wan 2.5 est votre arme secrète pour le marketing !

Entreprises mondiales

Fournit des solutions idéales de localisation de contenu pour les multinationales, rendant la création plus facile et plus efficace.

Support vidéo multilingue avec reconnaissance des prompts
Génération en un clic de sous-titres et voix off synchronisés
Localisation rapide du contenu pour les marchés mondiaux

Bottom line: En résumé : La création de contenu transfrontalier n'a jamais été aussi simple, rapide et intelligente.

Créateurs de contenu / YouTubers

Les créateurs peuvent tirer parti de Wan 2.5 pour améliorer leur efficacité de production vidéo tout en garantissant une sortie de haute qualité.

Narration immersive avec des actions et expressions précises des personnages
Meilleure efficacité de publication avec moins de montage et de post-production
Contenus variés, des courtes vidéos aux segments de récits animés

Équipes de formation en entreprise

Wan 2.5 rend la formation en entreprise plus efficace et engageante.

Des vidéos professionnelles remplacent les documents texte ennuyeux
Création rapide de démonstrations opérationnelles et de tutoriels de formation
Style cohérent et production standardisée pour un déploiement mondial

Créatifs indépendants / Petits studios

Wan 2.5 libère la créativité sans équipement coûteux ni acteurs — l'IA génère tout efficacement.

Explorez des œuvres variées, des courts métrages aux contenus pour les réseaux sociaux
De l'inspiration à l'achèvement grâce à la « génération en un clic »
Contenu de haute qualité sans équipement coûteux ni acteurs professionnels

Bottom line: En résumé : Wan 2.5 rend la création plus facile, plus libre et plus excitante à chaque tentative !

Établissements d'enseignement / Créateurs de cours en ligne

Transformez la créativité en réalité sans coûts élevés — Wan 2.5 rend la production de contenu de qualité simple et économique.

Explorez divers styles, des courts métrages aux vidéos promotionnelles
Meilleure efficacité de production, du concept au produit fini
Contenu de qualité sans équipement coûteux ni talents professionnels

Bottom line: En résumé : Wan 2.5 rend la création sans effort, efficace et libre — chaque tentative est spectaculaire !

Capacités Principales

Génération A/V en une étape

Générez des vidéos complètes avec audio synchronisé, voix off et synchronisation labiale en un seul processus

Synchronisation de deux personnages

Prend en charge la génération simultanée de deux personnages avec actions, expressions et synchronisation labiale coordonnées pour des interactions naturelles

Qualité professionnelle

Sortie vidéo haute qualité avec des expressions réalistes des personnages et une synchronisation labiale précise

Support multilingue

Excellent support des prompts en chinois et génération fiable de contenu multilingue

Rapport qualité-prix

Coûts considérablement inférieurs à ceux des concurrents tout en maintenant une qualité professionnelle

Restitution des traits des personnages

Recrée avec précision l'apparence, les expressions et les styles de mouvement des personnages avec une haute fidélité et personnalité

Rendu de style artistique

Prend en charge divers styles artistiques, notamment les textures aquarelles peintes à la main d'inspiration Studio Ghibli

Scènes immersives

Idéal pour les scènes de dialogue, interviews ou courts métrages à deux personnages avec une cohérence audio-visuelle naturelle

Digital Human Sync

Study Room Scholar

Middle-aged man reading with perfect lip-sync in a warm study environment

Lip-sync with audioEnvironmental soundsCharacter emotion

Prompt

Dual Character Scene

Park Sunset Romance

Couple interaction with synchronized dual character actions and expressions

Dual character syncNatural interactionAmbient soundscape

Prompt

A young couple sitting on a park bench during sunset. The woman leans her head on the man's shoulder. He whispers softly: 'No matter where we go, I'll always be here with you.' The sound includes the rustling of leaves, distant laughter of children playing, and the gentle hum of cicadas in the evening air.

Character Restoration

Ballet Performance Art

Precise character trait restoration with artistic movement and expression

Character trait restorationMovement precisionArtistic lighting

Prompt

Artistic Style Rendering

Ghibli Forest Magic

Studio Ghibli-inspired animation with hand-painted watercolor texture

Ghibli art styleHand-painted textureMagical atmosphere

Prompt

Studio Ghibli-inspired anime style. A young girl with a straw hat lies peacefully in a sun-dappled magical forest, surrounded by friendly, glowing forest spirits (Kodama). A gentle breeze rustles the leaves of the giant, ancient trees. The air is filled with sparkling dust motes, illuminated by shafts of sunlight. The art style is soft, with a hand-painted watercolor texture. The scene feels serene, magical, and heartwarming.

Parfait Pour

🎬

Production vidéo

📢

Contenu marketing

🎓

Vidéos éducatives

📱

Réseaux sociaux

🌐

Contenu multilingue

💼

Formation en entreprise

🎭

Divertissement

💃

Arts du spectacle

🎨

Animation et anime

📚

Narration

👥

Vidéos à deux personnages

🎙️

Entretiens

📺

Médias audiovisuels

Spécifications Techniques

Type de modèle :Génération audio-visuelle synchronisée

Fonctionnalités clés :Sync A/V, Restitution des personnages, Rendu artistique, Multilingue

Langues prises en charge :Chinois, anglais et plus

Qualité de sortie :Vidéo HD professionnelle avec audio

Vitesse de génération :Génération rapide en une étape

Intégration API :API RESTful avec documentation complète

Découvrez Wan 2.5 - Votre révolution dans la création vidéo

Rejoignez les milliers de créateurs et d'entreprises qui transforment leur création de contenu vidéo grâce à la génération audio-vidéo synchronisée.

🎬Sync A/V en une étape

🌍Support multilingue

⚡Rapport qualité-prix

Wan 2.5: A next-generation AI video generation model developed by Alibaba Wanxiang.

Model Card Overview

Field	Description
Model Name	Wan 2.5
Developed By	Alibaba Group
Release Date	September 24, 2025
Model Type	Generative AI, Video Foundation Model
Related Links	Official Website: https://wan.video/, Hugging Face: https://huggingface.co/Wan-AI, Technical Paper (Wan Series): https://arxiv.org/abs/2503.20314

Introduction

Wan 2.5 is a state-of-the-art, open-source video foundation model developed by Alibaba's Wan AI team. It is designed to generate high-quality, cinematic videos complete with synchronized audio directly from text or image prompts. The model represents a significant advancement in the field of generative AI, aiming to lower the barrier for creative video production. Its core contribution lies in its ability to produce coherent, dynamic, and narratively consistent video clips with a high degree of realism and integrated audio-visual elements, such as lip-sync and sound effects, in a single, streamlined process.

Key Features & Innovations

Wan 2.5 introduces several key features that distinguish it from previous models and competitors:

Unified Audio-Visual Synthesis: Unlike many models that require separate steps for video and audio generation, Wan 2.5 creates video with natively synchronized audio, including voice, sound effects, and lip-sync, in one step.
High-Fidelity, High-Resolution Output: The model is capable of generating videos in multiple resolutions, including 480p, 720p, and full 1080p HD, with significant improvements in visual quality and frame-to-frame stability over its predecessors.
Extended Video Duration: Wan 2.5 can generate video clips up to 10 seconds in length, offering more creative flexibility for storytelling compared to other models in its class.
Advanced Cinematic Control: The model demonstrates a sophisticated understanding of cinematic language, allowing for precise control over camera movement, shot composition, and character consistency within scenes.
Open-Source Commitment: Following the precedent set by earlier versions, the Wan series of models, including Wan 2.5, are open-sourced to encourage research, development, and innovation within the broader AI community.

Model Architecture & Technical Details

Wan 2.5 is built upon the Diffusion Transformer (DiT) paradigm, which has become a mainstream approach for high-quality generative tasks. The technical report for the Wan model series outlines a suite of innovations that contribute to its performance.

The architecture includes a novel Variational Autoencoder (VAE) designed for high-efficiency video compression, enabling the model to handle high-resolution video data effectively. The Wan series is available in multiple sizes to balance performance and computational requirements, such as the 1.3B and 14B parameter models detailed for Wan 2.2. The model was trained on a massive, curated dataset comprising billions of images and videos, which enhances its ability to generalize across a wide range of motions, semantics, and aesthetic styles.

Intended Use & Applications

Wan 2.5 is designed for a wide array of applications in creative and commercial fields. Its intended uses include:

Content Creation: Generating short-form videos for social media, marketing campaigns, and digital advertising.
Storytelling and Filmmaking: Creating cinematic scenes, character animations, and narrative sequences for short films and conceptual art.
Prototyping: Rapidly visualizing scripts and storyboards for film, television, and game development.
Personalized Media: Enabling users to create unique, personalized video content from their own ideas and images.

Performance

Wan 2.5 has demonstrated significant performance improvements over previous versions and holds a competitive position against other leading video generation models. Independent reviews and benchmarks provide insight into its capabilities.

Benchmark Scores

A review conducted by Curious Refuge Labs™ evaluated the model's visual generation capabilities across several metrics.

Metric	Score (out of 10)
Prompt Adherence	7.0
Temporal Consistency	6.6
Visual Fidelity	6.5
Motion Quality	5.9
Style & Cinematic Realism	5.7
Overall Score	6.3

These scores indicate strong prompt understanding and a notable improvement in visual quality from Wan 2.2, although it still shows limitations in complex motion and realism compared to top-tier commercial models.

Découvrir des modèles similaires

NEW

HOT

texte-vers-vidéo

Van-2.5 Text-to-video

Convert prompts into cinematic video clips with synchronized sound. Van 2.5 generates 720p/1080p outputs with stable motion, native audio sync, and prompt-faithful visual storytelling.

Van-2.5 Image-to-video

Get animated visuals from your images faster without major quality sacrifice. Perfect for preview workflows, previews at scale, or mass production of animated assets.

HappyHorse-1.1 Reference-to-video

Generates videos from one to nine reference images and a text prompt, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

HappyHorse-1.1 Image-to-video

Animates a first-frame image into video with optional prompt guidance, 720P or 1080P output, and durations from 3 to 15 seconds.

HappyHorse-1.1 Text-to-video

Generates videos from text prompts with HappyHorse 1.1, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

HappyHorse-1.0 Image-to-video

Animates a first-frame image into video with optional prompt guidance, 720P or 1080P output, and durations from 3 to 15 seconds.

HappyHorse-1.0 Text-to-video

Generates videos from text prompts with HappyHorse 1.0, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

HappyHorse-1.0 Video-edit

Edits an input video with text instructions and optional reference images, supporting 720P or 1080P output.

HappyHorse-1.0 Reference-to-video

Generates videos from one to nine reference images and a text prompt, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

From

$0.14/SEC