alibaba/wan-2.5/video-extend

ข้อความเป็นวิดีโอ

Wan 2.5 Video Extend API by Alibaba

alibaba/wan-2.5/video-extend

Video-extend

Extend your videos with Alibaba WAN 2.5 video extender model with audio.

อินพุต

กำลังโหลดการตั้งค่าพารามิเตอร์...

เอาต์พุต

รอดำเนินการ

วิดีโอที่สร้างจะแสดงที่นี่

ตั้งค่าพารามิเตอร์แล้วคลิกรันเพื่อเริ่มสร้าง

แต่ละครั้งจะใช้ $0.052 ด้วย $10 คุณสามารถรันได้ประมาณ 192 ครั้ง

คุณสามารถทำต่อได้:

Seedance 2.0 Kling v3 Vidu Wan2.7

พารามิเตอร์

ตัวอย่างโค้ด
import requests
import time

# Step 1: Start video generation
generate_url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "alibaba/wan-2.5/video-extend",
    "prompt": "A beautiful sunset over the ocean with gentle waves",
    "width": 512,
    "height": 512,
    "duration": 3,
    "fps": 24,
}

generate_response = requests.post(generate_url, headers=headers, json=data)
generate_result = generate_response.json()
prediction_id = generate_result["data"]["id"]

# Step 2: Poll for result
poll_url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"

def check_status():
    while True:
        response = requests.get(poll_url, headers={"Authorization": "Bearer $ATLASCLOUD_API_KEY"})
        result = response.json()

        if result["data"]["status"] in ["completed", "succeeded"]:
            print("Generated video:", result["data"]["outputs"][0])
            return result["data"]["outputs"][0]
        elif result["data"]["status"] == "failed":
            raise Exception(result["data"]["error"] or "Generation failed")
        else:
            # Still processing, wait 2 seconds
            time.sleep(2)

video_url = check_status()

ติดตั้ง

ติดตั้งแพ็กเกจที่จำเป็น

pip install requests

การยืนยันตัวตน

คำขอ API ทั้งหมดต้องมีการยืนยันตัวตนผ่าน API key คุณสามารถรับ API key ได้จากแดชบอร์ด Atlas Cloud

export ATLASCLOUD_API_KEY="your-api-key-here"

HTTP Headers

import os

API_KEY = os.environ.get("ATLASCLOUD_API_KEY")
headers = {
    "Content-Type": "application/json",
    "Authorization": f"Bearer {API_KEY}"
}

รักษา API key ของคุณให้ปลอดภัย

อย่าเปิดเผย API key ของคุณในโค้ดฝั่งไคลเอนต์หรือที่เก็บข้อมูลสาธารณะ ให้ใช้ตัวแปรสภาพแวดล้อมหรือพร็อกซีฝั่งเซิร์ฟเวอร์แทน

ส่งคำขอ

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "your-model",
    "prompt": "A beautiful landscape"
}

response = requests.post(url, headers=headers, json=data)
print(response.json())

ส่งคำขอ

ส่งคำขอสร้างแบบอะซิงโครนัส API จะส่งคืน prediction ID ที่คุณสามารถใช้ตรวจสอบสถานะและดึงผลลัพธ์ได้

POST/api/v1/model/generateVideo

เนื้อหาคำขอ

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}

data = {
    "model": "alibaba/wan-2.5/video-extend",
    "input": {
        "prompt": "A beautiful sunset over the ocean with gentle waves"
    }
}

response = requests.post(url, headers=headers, json=data)
result = response.json()

print(f"Prediction ID: {result['id']}")
print(f"Status: {result['status']}")

การตอบกลับ

{
  "id": "pred_abc123",
  "status": "processing",
  "model": "model-name",
  "created_at": "2025-01-01T00:00:00Z"
}

ตรวจสอบสถานะ

ตรวจสอบสถานะปัจจุบันของคำขอด้วยการเรียก prediction endpoint เป็นระยะ

GET/api/v1/model/prediction/{prediction_id}

ตัวอย่างการตรวจสอบสถานะเป็นระยะ

import requests
import time

prediction_id = "pred_abc123"
url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

while True:
    response = requests.get(url, headers=headers)
    result = response.json()
    status = result["data"]["status"]
    print(f"Status: {status}")

    if status in ["completed", "succeeded"]:
        output_url = result["data"]["outputs"][0]
        print(f"Output URL: {output_url}")
        break
    elif status == "failed":
        print(f"Error: {result['data'].get('error', 'Unknown')}")
        break

    time.sleep(3)

ค่าสถานะ

processingคำขอยังอยู่ระหว่างการประมวลผล

completedการสร้างเสร็จสมบูรณ์แล้ว ผลลัพธ์พร้อมใช้งาน

succeededการสร้างสำเร็จแล้ว ผลลัพธ์พร้อมใช้งาน

failedการสร้างล้มเหลว ตรวจสอบฟิลด์ error

การตอบกลับที่เสร็จสมบูรณ์

{
  "data": {
    "id": "pred_abc123",
    "status": "completed",
    "outputs": [
      "https://storage.atlascloud.ai/outputs/result.mp4"
    ],
    "metrics": {
      "predict_time": 45.2
    },
    "created_at": "2025-01-01T00:00:00Z",
    "completed_at": "2025-01-01T00:00:10Z"
  }
}

อัปโหลดไฟล์

อัปโหลดไฟล์ไปยังที่เก็บข้อมูล Atlas Cloud และรับ URL ที่คุณสามารถใช้ในคำขอ API ของคุณ ใช้ multipart/form-data ในการอัปโหลด

POST/api/v1/model/uploadMedia

ตัวอย่างการอัปโหลด

import requests

url = "https://api.atlascloud.ai/api/v1/model/uploadMedia"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

with open("image.png", "rb") as f:
    files = {"file": ("image.png", f, "image/png")}
    response = requests.post(url, headers=headers, files=files)

result = response.json()
download_url = result["data"]["download_url"]
print(f"File URL: {download_url}")

การตอบกลับ

{
  "data": {
    "download_url": "https://storage.atlascloud.ai/uploads/abc123/image.png",
    "file_name": "image.png",
    "content_type": "image/png",
    "size": 1024000
  }
}

Input Schema

พารามิเตอร์ต่อไปนี้ยอมรับในเนื้อหาคำขอ

ทั้งหมด: 0จำเป็น: 0ไม่บังคับ: 0

ไม่มีพารามิเตอร์ที่ใช้ได้

ตัวอย่างเนื้อหาคำขอ

{
  "model": "alibaba/wan-2.5/video-extend"
}

Output Schema

API จะส่งคืนการตอบกลับ prediction พร้อม URL ของผลลัพธ์ที่สร้างขึ้น

idstringrequired

Unique identifier for the prediction.

statusstringrequired

Current status of the prediction.

processingcompletedsucceededfailed

modelstringrequired

The model used for generation.

outputsarray[string]

Array of output URLs. Available when status is "completed".

errorstring

Error message if status is "failed".

metricsobject

Performance metrics.

predict_timenumber

Time taken for video generation in seconds.

created_atstringrequired

ISO 8601 timestamp when the prediction was created.

Format: date-time

completed_atstring

ISO 8601 timestamp when the prediction was completed.

Format: date-time

ตัวอย่างการตอบกลับ

{
  "id": "pred_abc123",
  "status": "completed",
  "model": "model-name",
  "outputs": [
    "https://storage.atlascloud.ai/outputs/result.mp4"
  ],
  "metrics": {
    "predict_time": 45.2
  },
  "created_at": "2025-01-01T00:00:00Z",
  "completed_at": "2025-01-01T00:00:10Z"
}

Atlas Cloud Skills

Atlas Cloud Skills เชื่อมต่อโมเดล AI กว่า 300+ เข้ากับผู้ช่วยเขียนโค้ด AI ของคุณโดยตรง ติดตั้งด้วยคำสั่งเดียว จากนั้นใช้ภาษาธรรมชาติเพื่อสร้างรูปภาพ วิดีโอ และสนทนากับ LLM

ไคลเอนต์ที่รองรับ

Claude Code

OpenAI Codex

Gemini CLI

Cursor

Windsurf

VS Code

Trae

GitHub Copilot

Cline

Roo Code

Amp

Goose

Replit

40+ ไคลเอนต์ที่รองรับ

ติดตั้ง

npx skills add AtlasCloudAI/atlas-cloud-skills

ตั้งค่า API Key

รับ API key จากแดชบอร์ด Atlas Cloud และตั้งค่าเป็นตัวแปรสภาพแวดล้อม

export ATLASCLOUD_API_KEY="your-api-key-here"

ความสามารถ

เมื่อติดตั้งแล้ว คุณสามารถใช้ภาษาธรรมชาติในผู้ช่วย AI ของคุณเพื่อเข้าถึงโมเดล Atlas Cloud ทั้งหมด

สร้างรูปภาพสร้างรูปภาพด้วยโมเดลเช่น Nano Banana 2, Z-Image และอื่นๆ

สร้างวิดีโอสร้างวิดีโอจากข้อความหรือรูปภาพด้วย Kling, Vidu, Veo เป็นต้น

สนทนา LLMสนทนากับ Qwen, DeepSeek และโมเดลภาษาขนาดใหญ่อื่นๆ

อัปโหลดสื่ออัปโหลดไฟล์จากเครื่องสำหรับการแก้ไขรูปภาพและเวิร์กโฟลว์รูปภาพเป็นวิดีโอ

เรียนรู้เพิ่มเติม

github.com/AtlasCloudAI/atlas-cloud-skills

MCP Server

Atlas Cloud MCP Server เชื่อมต่อ IDE ของคุณกับโมเดล AI กว่า 300+ ผ่าน Model Context Protocol ใช้งานได้กับไคลเอนต์ที่รองรับ MCP ทุกตัว

ไคลเอนต์ที่รองรับ

Cursor

VS Code

Windsurf

Claude Code

OpenAI Codex

Gemini CLI

Cline

Roo Code

100+ ไคลเอนต์ที่รองรับ

ติดตั้ง

npx -y atlascloud-mcp

การกำหนดค่า

เพิ่มการกำหนดค่าต่อไปนี้ลงในไฟล์ตั้งค่า MCP ของ IDE ของคุณ

{
  "mcpServers": {
    "atlascloud": {
      "command": "npx",
      "args": [
        "-y",
        "atlascloud-mcp"
      ],
      "env": {
        "ATLASCLOUD_API_KEY": "your-api-key-here"
      }
    }
  }
}

เครื่องมือที่ใช้ได้

atlas_generate_imageสร้างรูปภาพจากข้อความ prompt

atlas_generate_videoสร้างวิดีโอจากข้อความหรือรูปภาพ

atlas_chatสนทนากับโมเดลภาษาขนาดใหญ่

atlas_list_modelsเรียกดูโมเดล AI กว่า 300+ ที่ใช้ได้

atlas_quick_generateสร้างเนื้อหาขั้นตอนเดียวพร้อมเลือกโมเดลอัตโนมัติ

atlas_upload_mediaอัปโหลดไฟล์จากเครื่องสำหรับเวิร์กโฟลว์ API

เรียนรู้เพิ่มเติม

github.com/AtlasCloudAI/mcp-server

API Schema

ไม่มี Schema

ไม่มีตัวอย่าง

กรุณาเข้าสู่ระบบเพื่อดูประวัติคำขอ

คุณต้องเข้าสู่ระบบเพื่อเข้าถึงประวัติคำขอโมเดล

เข้าสู่ระบบ

Wan 2.5 - ตัวเลือกของผู้สร้างวิดีโออัจฉริยะ

ยอดนิยม

การสร้างเสียงและวิดีโอแบบซิงก์ครบในที่เดียว

Wan 2.5 คือโมเดลสร้างวิดีโอ AI ที่ปฏิวัติวงการ สร้างเนื้อหาเสียงและภาพแบบซิงก์ได้ในขั้นตอนเดียว ไม่ต้องบันทึกเสียงแยกหรือปรับ Lip Sync ด้วยตนเอง เพียงใส่ Prompt ที่ชัดเจนและมีโครงสร้าง ก็สร้างวิดีโอสมบูรณ์พร้อมเสียง/พากย์เสียงและ Lip Sync ได้ทันที

ทำไมต้องเลือก Wan 2.5?

คุ้มค่ากว่า

แม้ Google จะลดราคาเมื่อไม่นานมานี้ แต่ Veo 3 โดยรวมยังคงมีราคาสูง Wan 2.5 มีน้ำหนักเบาและคุ้มค่า มอบตัวเลือกเพิ่มเติมให้กับนักสร้างสรรค์ พร้อมลดต้นทุนการผลิตได้อย่างมีนัยสำคัญ

สร้างในขั้นตอนเดียว ซิงก์แบบครบวงจร

ด้วย Wan 2.5 ไม่จำเป็นต้องบันทึกเสียงแยกต่างหากหรือปรับแนว Lip Sync ด้วยตนเอง เพียงใส่ Prompt ที่ชัดเจนและมีโครงสร้าง ก็สร้างวิดีโอสมบูรณ์พร้อมเสียง/พากย์และ Lip Sync ได้ในครั้งเดียว — รวดเร็วและง่ายกว่า

รองรับหลายภาษา

เมื่อใช้ Prompt ภาษาจีน Wan 2.5 สร้างวิดีโอที่ซิงก์เสียงและภาพได้อย่างน่าเชื่อถือ ในขณะที่ Veo 3 มักแสดง ภาษาไม่รู้จัก สำหรับ Prompt ภาษาจีน

สร้างตัวละครได้แม่นยำ

Wan 2.5 โดดเด่นด้านการฟื้นฟูลักษณะตัวละคร นำเสนอรูปลักษณ์ สีหน้า และรูปแบบการเคลื่อนไหวของตัวละครได้อย่างแม่นยำ ทำให้ตัวละครในวิดีโอมีเอกลักษณ์และบุคลิกที่โดดเด่น เพิ่มความลึกในการเล่าเรื่องและความดื่มด่ำ

แสดงผลสไตล์ศิลปะ

รองรับการแสดงผลสไตล์ Studio Ghibli สร้างพื้นผิวสีน้ำที่วาดด้วยมือและเอฟเฟกต์แอนิเมชัน มอบประสบการณ์ภาพที่อบอุ่นและฝันหวาน เพิ่มเสน่ห์ทางศิลปะและความลึกของการเล่าเรื่อง

ใครได้รับประโยชน์?

ทีมการตลาด

ไม่ว่าจะเป็นการเปิดตัวผลิตภัณฑ์ แคมเปญโปรโมชัน หรือการตลาดแบรนด์ Wan 2.5 ช่วยให้คุณสร้างวิดีโอคุณภาพสูงได้อย่างรวดเร็ว ทำให้การสร้างสรรค์ง่ายและมีประสิทธิภาพ

Demo ผลิตภัณฑ์และบทเรียน โดยไม่ต้องยุ่งยากกับการประสานงาน
การตลาดโซเชียลมีเดียพร้อมซับไตเติ้ลหลายภาษาและ Lip Sync
เนื้อหาที่สร้างโดย AI ให้ทีมมุ่งเน้นกลยุทธ์และความคิดสร้างสรรค์

Bottom line: สรุป: การสร้างสรรค์ไม่เคยง่าย รวดเร็ว และฉลาดเท่านี้มาก่อน — Wan 2.5 คือเครื่องมือลับสำหรับการตลาดของคุณ!

องค์กรระดับโลก

มอบโซลูชันการโลคัลไลซ์เนื้อหาที่เหมาะสมสำหรับบริษัทข้ามชาติ ทำให้การสร้างสรรค์ง่ายและมีประสิทธิภาพยิ่งขึ้น

รองรับวิดีโอหลายภาษาพร้อมการรู้จำ Prompt
สร้างซับไตเติ้ลและพากย์เสียงพร้อม Lip Sync ด้วยคลิกเดียว
โลคัลไลซ์เนื้อหาได้รวดเร็วสำหรับตลาดทั่วโลก

Bottom line: สรุป: การสร้างเนื้อหาข้ามพรมแดนไม่เคยง่าย รวดเร็ว และฉลาดเท่านี้มาก่อน

นักสร้างเรื่องราว / YouTubers

นักสร้างสรรค์สามารถใช้ประโยชน์จาก Wan 2.5 เพื่อเพิ่มประสิทธิภาพการผลิตวิดีโอพร้อมรับประกันคุณภาพสูง

การเล่าเรื่องแบบดื่มด่ำด้วยการเคลื่อนไหวและสีหน้าตัวละครที่แม่นยำ
ประสิทธิภาพการเผยแพร่สูงขึ้นด้วยการลดเวลาตัดต่อและหลังการผลิต
เนื้อหาที่หลากหลายตั้งแต่วิดีโอสั้นไปจนถึงเรื่องราวแอนิเมชัน

ทีมฝึกอบรมองค์กร

Wan 2.5 ทำให้การฝึกอบรมองค์กรมีประสิทธิภาพและน่าสนใจยิ่งขึ้น

วิดีโอระดับมืออาชีพแทนที่เอกสารข้อความที่น่าเบื่อ
สร้าง Demo การปฏิบัติงานและบทเรียนฝึกอบรมได้อย่างรวดเร็ว
สไตล์ที่สม่ำเสมอและผลลัพธ์มาตรฐานสำหรับการขยายทั่วโลก

ฟรีแลนซ์สร้างสรรค์ / สตูดิโอขนาดเล็ก

Wan 2.5 ปล่อยให้ความคิดสร้างสรรค์ไหลลื่นโดยไม่ต้องใช้อุปกรณ์ราคาแพงหรือนักแสดง — AI สร้างทุกอย่างได้อย่างมีประสิทธิภาพ

ทดลองสร้างผลงานที่หลากหลาย ตั้งแต่หนังสั้นไปจนถึงเนื้อหาโซเชียลมีเดีย
จากแรงบันดาลใจสู่ผลงานสำเร็จด้วย การสร้างคลิกเดียว
เนื้อหาคุณภาพสูงโดยไม่ต้องใช้อุปกรณ์ราคาแพงหรือนักแสดงมืออาชีพ

Bottom line: สรุป: Wan 2.5 ทำให้การสร้างสรรค์ง่ายขึ้น อิสระขึ้น และน่าตื่นเต้นขึ้นในทุกครั้ง!

สถาบันการศึกษา / นักสร้างคอร์สออนไลน์

แปลงความคิดสร้างสรรค์เป็นความจริงโดยไม่ต้องใช้ต้นทุนสูง — Wan 2.5 ทำให้การผลิตเนื้อหาคุณภาพง่ายและประหยัด

ทดลองสไตล์ต่างๆ ตั้งแต่หนังสั้นไปจนถึงวิดีโอโปรโมชัน
ประสิทธิภาพการผลิตสูงขึ้นตั้งแต่แนวคิดจนถึงผลิตภัณฑ์สำเร็จรูป
เนื้อหาคุณภาพดีโดยไม่ต้องใช้อุปกรณ์ราคาแพงหรือผู้เชี่ยวชาญ

Bottom line: สรุป: Wan 2.5 ทำให้การสร้างสรรค์เป็นเรื่องง่าย มีประสิทธิภาพ และอิสระ — ทุกครั้งที่ลองล้วนน่าตื่นตาตื่นใจ!

คุณสมบัติหลัก

สร้างเสียงและภาพในขั้นตอนเดียว

สร้างวิดีโอสมบูรณ์พร้อมเสียงซิงก์ พากย์เสียง และ Lip Sync ในกระบวนการเดียว

ซิงก์ตัวละครคู่

รองรับการสร้างตัวละครสองตัวพร้อมกัน ซิงก์การเคลื่อนไหว สีหน้า และ Lip Sync เพื่อปฏิสัมพันธ์ที่เป็นธรรมชาติ

คุณภาพระดับมืออาชีพ

วิดีโอคุณภาพสูงพร้อมสีหน้าตัวละครที่สมจริงและการซิงก์ Lip Sync ที่แม่นยำ

รองรับหลายภาษา

รองรับ Prompt ภาษาจีนได้ยอดเยี่ยม และสร้างเนื้อหาหลายภาษาได้อย่างน่าเชื่อถือ

คุ้มค่าคุ้มราคา

ต้นทุนต่ำกว่าคู่แข่งอย่างมีนัยสำคัญ พร้อมรักษาคุณภาพระดับมืออาชีพ

ฟื้นฟูลักษณะตัวละคร

สร้างรูปลักษณ์ สีหน้า และรูปแบบการเคลื่อนไหวของตัวละครได้อย่างแม่นยำด้วยความเที่ยงตรงและบุคลิกที่สูง

แสดงผลสไตล์ศิลปะ

รองรับสไตล์ศิลปะหลากหลาย รวมถึงพื้นผิวสีน้ำที่วาดด้วยมือแบบ Studio Ghibli

ฉากดื่มด่ำ

เหมาะอย่างยิ่งสำหรับฉากบทสนทนา การสัมภาษณ์ หรือหนังสั้นคู่ ด้วยความสอดคล้องของเสียงและภาพที่เป็นธรรมชาติ

Digital Human Sync

Study Room Scholar

Middle-aged man reading with perfect lip-sync in a warm study environment

Lip-sync with audioEnvironmental soundsCharacter emotion

Prompt

A middle-aged man sitting at a wooden desk in a cozy study room, surrounded by bookshelves and a warm lamp glow. He opens an old book and reads aloud with a calm, deep voice: 'History teaches us more than just facts… it shows us who we are.' The room has subtle background sounds: pages turning, the faint ticking of a clock, and distant rain against the window.

Dual Character Scene

Park Sunset Romance

Couple interaction with synchronized dual character actions and expressions

Dual character syncNatural interactionAmbient soundscape

Prompt

A young couple sitting on a park bench during sunset. The woman leans her head on the man's shoulder. He whispers softly: 'No matter where we go, I'll always be here with you.' The sound includes the rustling of leaves, distant laughter of children playing, and the gentle hum of cicadas in the evening air.

Character Restoration

Ballet Performance Art

Precise character trait restoration with artistic movement and expression

Character trait restorationMovement precisionArtistic lighting

Prompt

A graceful ballerina with her hair in a messy bun, performing a powerful and emotional contemporary ballet routine. She is in a minimalist, dark art studio. Abstract patterns of light and shadow, projected from a hidden source, dance across her body and the surrounding walls, constantly shifting with her movements. The camera focuses on the tension in her muscles and the expressive gestures of her hands. A single, dramatic slow-motion shot captures her mid-air leap, with the light patterns swirling around her like a galaxy. Moody, artistic, high contrast.

Artistic Style Rendering

Ghibli Forest Magic

Studio Ghibli-inspired animation with hand-painted watercolor texture

Ghibli art styleHand-painted textureMagical atmosphere

Prompt

Studio Ghibli-inspired anime style. A young girl with a straw hat lies peacefully in a sun-dappled magical forest, surrounded by friendly, glowing forest spirits (Kodama). A gentle breeze rustles the leaves of the giant, ancient trees. The air is filled with sparkling dust motes, illuminated by shafts of sunlight. The art style is soft, with a hand-painted watercolor texture. The scene feels serene, magical, and heartwarming.

เหมาะสำหรับ

🎬

การผลิตวิดีโอ

📢

เนื้อหาการตลาด

🎓

วิดีโอการศึกษา

📱

โซเชียลมีเดีย

🌐

เนื้อหาหลายภาษา

💼

การฝึกอบรมองค์กร

🎭

ความบันเทิง

💃

ศิลปะการแสดง

🎨

แอนิเมชันและอนิเมะ

📚

การเล่าเรื่อง

👥

วิดีโอตัวละครคู่

🎙️

การสัมภาษณ์

📺

สื่อกระจายเสียง

ข้อมูลจำเพาะทางเทคนิค

ประเภทโมเดล:การสร้างเสียงและภาพแบบซิงก์

คุณสมบัติหลัก:ซิงก์เสียง/ภาพ, ฟื้นฟูตัวละคร, แสดงผลศิลปะ, หลายภาษา

รองรับภาษา:จีน, อังกฤษ และอื่นๆ

คุณภาพผลลัพธ์:วิดีโอ HD ระดับมืออาชีพพร้อมเสียง

ความเร็วในการสร้าง:สร้างเร็วในขั้นตอนเดียว

การผสาน API:RESTful API พร้อมเอกสารประกอบครบถ้วน

สัมผัส Wan 2.5 - การปฏิวัติการสร้างวิดีโอของคุณ

ร่วมกับผู้สร้างสรรค์และองค์กรนับพัน ที่กำลังพลิกโฉมการสร้างเนื้อหาวิดีโอด้วยเทคโนโลยีการสร้างเสียงและภาพแบบซิงก์

🎬ซิงก์เสียง/ภาพในขั้นตอนเดียว

🌍รองรับหลายภาษา

⚡คุ้มค่าคุ้มราคา

Wan 2.5: A next-generation AI video generation model developed by Alibaba Wanxiang.

Model Card Overview

Field	Description
Model Name	Wan 2.5
Developed By	Alibaba Group
Release Date	September 24, 2025
Model Type	Generative AI, Video Foundation Model
Related Links	Official Website: https://wan.video/, Hugging Face: https://huggingface.co/Wan-AI, Technical Paper (Wan Series): https://arxiv.org/abs/2503.20314

Introduction

Wan 2.5 is a state-of-the-art, open-source video foundation model developed by Alibaba's Wan AI team. It is designed to generate high-quality, cinematic videos complete with synchronized audio directly from text or image prompts. The model represents a significant advancement in the field of generative AI, aiming to lower the barrier for creative video production. Its core contribution lies in its ability to produce coherent, dynamic, and narratively consistent video clips with a high degree of realism and integrated audio-visual elements, such as lip-sync and sound effects, in a single, streamlined process.

Key Features & Innovations

Wan 2.5 introduces several key features that distinguish it from previous models and competitors:

Unified Audio-Visual Synthesis: Unlike many models that require separate steps for video and audio generation, Wan 2.5 creates video with natively synchronized audio, including voice, sound effects, and lip-sync, in one step.
High-Fidelity, High-Resolution Output: The model is capable of generating videos in multiple resolutions, including 480p, 720p, and full 1080p HD, with significant improvements in visual quality and frame-to-frame stability over its predecessors.
Extended Video Duration: Wan 2.5 can generate video clips up to 10 seconds in length, offering more creative flexibility for storytelling compared to other models in its class.
Advanced Cinematic Control: The model demonstrates a sophisticated understanding of cinematic language, allowing for precise control over camera movement, shot composition, and character consistency within scenes.
Open-Source Commitment: Following the precedent set by earlier versions, the Wan series of models, including Wan 2.5, are open-sourced to encourage research, development, and innovation within the broader AI community.

Model Architecture & Technical Details

Wan 2.5 is built upon the Diffusion Transformer (DiT) paradigm, which has become a mainstream approach for high-quality generative tasks. The technical report for the Wan model series outlines a suite of innovations that contribute to its performance.

The architecture includes a novel Variational Autoencoder (VAE) designed for high-efficiency video compression, enabling the model to handle high-resolution video data effectively. The Wan series is available in multiple sizes to balance performance and computational requirements, such as the 1.3B and 14B parameter models detailed for Wan 2.2. The model was trained on a massive, curated dataset comprising billions of images and videos, which enhances its ability to generalize across a wide range of motions, semantics, and aesthetic styles.

Intended Use & Applications

Wan 2.5 is designed for a wide array of applications in creative and commercial fields. Its intended uses include:

Content Creation: Generating short-form videos for social media, marketing campaigns, and digital advertising.
Storytelling and Filmmaking: Creating cinematic scenes, character animations, and narrative sequences for short films and conceptual art.
Prototyping: Rapidly visualizing scripts and storyboards for film, television, and game development.
Personalized Media: Enabling users to create unique, personalized video content from their own ideas and images.

Performance

Wan 2.5 has demonstrated significant performance improvements over previous versions and holds a competitive position against other leading video generation models. Independent reviews and benchmarks provide insight into its capabilities.

Benchmark Scores

A review conducted by Curious Refuge Labs™ evaluated the model's visual generation capabilities across several metrics.

Metric	Score (out of 10)
Prompt Adherence	7.0
Temporal Consistency	6.6
Visual Fidelity	6.5
Motion Quality	5.9
Style & Cinematic Realism	5.7
Overall Score	6.3

These scores indicate strong prompt understanding and a notable improvement in visual quality from Wan 2.2, although it still shows limitations in complex motion and realism compared to top-tier commercial models.

สำรวจโมเดลที่คล้ายกัน

NEW

HOT

ข้อความเป็นวิดีโอ

Van-2.5 Text-to-video

Convert prompts into cinematic video clips with synchronized sound. Van 2.5 generates 720p/1080p outputs with stable motion, native audio sync, and prompt-faithful visual storytelling.

Van-2.5 Image-to-video

Get animated visuals from your images faster without major quality sacrifice. Perfect for preview workflows, previews at scale, or mass production of animated assets.

HappyHorse-1.0 Image-to-video

Animates a first-frame image into video with optional prompt guidance, 720P or 1080P output, and durations from 3 to 15 seconds.

HappyHorse-1.0 Text-to-video

Generates videos from text prompts with HappyHorse 1.0, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

HappyHorse-1.0 Video-edit

Edits an input video with text instructions and optional reference images, supporting 720P or 1080P output.

HappyHorse-1.0 Reference-to-video

Generates videos from one to nine reference images and a text prompt, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

Wan-2.7 Image-to-video

Animates images into videos with first-frame, first-and-last-frame, video continuation, and audio-driven modes.

Wan-2.7 Text-to-video

Generates videos from text prompts with multi-shot narrative, audio generation, and sound-image synchronization.

Wan-2.7 Video-edit

Edits videos using text instructions, reference images, and style transfer with multi-modal input support.

Wan-2.7 Reference-to-video

Generates character-driven videos from reference images and videos, with multi-subject and voice-cloning support.

Wan-2.2 Image-to-video

Open and Advanced Large-Scale Video Generative Models.

Wan-2.2 Image-to-video Lora

Open and Advanced Large-Scale Video Generative Models.

From

$0.04/วินาที

ภาพเป็นวิดีโอ

Wan-2.2-spicy Image-to-video

Open and Advanced Large-Scale Video Generative Models.

From

$0.03/วินาที

ภาพเป็นวิดีโอ

Wan-2.2-spicy Image-to-video Lora

Open and Advanced Large-Scale Video Generative Models.

Wan-2.6 Image-to-video Flash

Wan2.6 image to video flash, faster and more cost-effective generation. Intelligent shot scheduling enables multi‑camera storytelling, supports stable multi‑speaker dialogue with more natural and realistic vocal timbres.

Wan-2.6 Video-to-video

A speed-optimized video-to-video option that prioritizes lower latency while retaining strong visual fidelity. Ideal for iteration, batch generation, and prompt testing.

From$0.1/วินาที

$0.07/วินาที

-30%

API เดียวสำหรับ AI สื่อทุกประเภท

สำรวจโมเดลทั้งหมด

Wan 2.5 Video Extend API by Alibaba

อินพุต

เอาต์พุต

พารามิเตอร์

ตัวอย่างโค้ด

ติดตั้ง

การยืนยันตัวตน

HTTP Headers

ส่งคำขอ

ส่งคำขอ

เนื้อหาคำขอ

การตอบกลับ

ตรวจสอบสถานะ

ตัวอย่างการตรวจสอบสถานะเป็นระยะ

ค่าสถานะ

การตอบกลับที่เสร็จสมบูรณ์

อัปโหลดไฟล์

ตัวอย่างการอัปโหลด

การตอบกลับ

Input Schema

ตัวอย่างเนื้อหาคำขอ

Output Schema

ตัวอย่างการตอบกลับ

Atlas Cloud Skills

ไคลเอนต์ที่รองรับ

ติดตั้ง

ตั้งค่า API Key

ความสามารถ

MCP Server

ไคลเอนต์ที่รองรับ

ติดตั้ง

การกำหนดค่า

เครื่องมือที่ใช้ได้

API Schema

กรุณาเข้าสู่ระบบเพื่อดูประวัติคำขอ

Wan 2.5 - ตัวเลือกของผู้สร้างวิดีโออัจฉริยะ

ทำไมต้องเลือก Wan 2.5?

คุ้มค่ากว่า

สร้างในขั้นตอนเดียว ซิงก์แบบครบวงจร

รองรับหลายภาษา

สร้างตัวละครได้แม่นยำ

แสดงผลสไตล์ศิลปะ

ใครได้รับประโยชน์?

ทีมการตลาด

องค์กรระดับโลก

นักสร้างเรื่องราว / YouTubers

ทีมฝึกอบรมองค์กร

ฟรีแลนซ์สร้างสรรค์ / สตูดิโอขนาดเล็ก

สถาบันการศึกษา / นักสร้างคอร์สออนไลน์

คุณสมบัติหลัก

สร้างเสียงและภาพในขั้นตอนเดียว

ซิงก์ตัวละครคู่

คุณภาพระดับมืออาชีพ

รองรับหลายภาษา

คุ้มค่าคุ้มราคา

ฟื้นฟูลักษณะตัวละคร

แสดงผลสไตล์ศิลปะ

ฉากดื่มด่ำ

Wan 2.5 Prompt Showcase

Study Room Scholar

Park Sunset Romance

Ballet Performance Art

Ghibli Forest Magic

เหมาะสำหรับ

ข้อมูลจำเพาะทางเทคนิค

สัมผัส Wan 2.5 - การปฏิวัติการสร้างวิดีโอของคุณ

Wan 2.5: A next-generation AI video generation model developed by Alibaba Wanxiang.

Model Card Overview

Introduction

Key Features & Innovations

Model Architecture & Technical Details

Intended Use & Applications

Performance

Benchmark Scores

สำรวจโมเดลที่คล้ายกัน

Van-2.5 Text-to-video

Van-2.5 Image-to-video

HappyHorse-1.0 Image-to-video

HappyHorse-1.0 Text-to-video

HappyHorse-1.0 Video-edit