bytedance/seedance-2.0/image-to-video

ภาพเป็นวิดีโอ

Seedance 2.0 Image-to-Video API by ByteDance

bytedance/seedance-2.0/image-to-video

Image-to-video

Generate videos from a first-frame image (and optional last-frame) with native audio.

อินพุต

พรอมต์

รูปภาพ *

ลากไฟล์มาวางที่นี่ หรือคลิกเพื่ออัปโหลด

MAX:1

เฟรมท้ายสุด

ลากไฟล์มาวางที่นี่ หรือคลิกเพื่ออัปโหลด

MAX:1

ระยะเวลา

ความละเอียด

อัตราส่วนภาพ

โหมดอัตราบิต

การตั้งค่าขั้นสูง

เอาต์พุต

รอดำเนินการ

วิดีโอที่สร้างจะแสดงที่นี่

ตั้งค่าพารามิเตอร์แล้วคลิกรันเพื่อเริ่มสร้าง

วิดีโอ 720p ที่คุณสร้างจะถูกเรียกเก็บ $0.2419/วินาที ต่อทุก 1 วินาที คำขอของคุณคิด $0.0112 ต่อ 1000 tokens จำนวน tokens คำนวณจาก (ความสูงของวิดีโอที่ส่งออก × ความกว้างของวิดีโอที่ส่งออก ×(ระยะเวลาอินพุต + ระยะเวลาเอาต์พุต)× 24) / 1024 หากใส่วิดีโอเป็นอินพุต อัตราค่าบริการจะลดลงเหลือ $0.00688 ต่อ 1000 tokens เมื่อใช้วิดีโออินพุตและความละเอียด 720p ราคาจะอยู่ที่ $0.1486 ต่อวินาที

คุณสามารถทำต่อได้:

Seedance 2.0 Kling v3 Vidu Wan2.7

พารามิเตอร์

ตัวอย่างโค้ด
import requests
import time

# Step 1: Start video generation
generate_url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "bytedance/seedance-2.0/image-to-video",  # Required. Model name
    "prompt": "The scene comes alive with gentle motion and cinematic lighting",  # Text prompt describing the desired video motion
    "image": "example_value",  # Required. First-frame image URL, Base64, or asset reference (asset://<ASSET_ID>)
    "last_image": "example_value",  # Last-frame image URL, Base64, or asset reference
    "duration": 5,  # Video duration in seconds (4-15), or -1 for model to choose automatically
    "resolution": "720p",  # Video resolution
    "ratio": "adaptive",  # Aspect ratio
    "bitrate_mode": "standard",  # Output video bitrate mode. options: standard | high
    "generate_audio": True,  # Whether to generate synchronized audio
    "seed": -1,  # Seed integer used to control the randomness of generated content. (min: -1, max: 4294967295)
    "watermark": False,  # Whether to add a watermark
    "return_last_frame": False,  # Whether to return the last frame as a separate image
}

generate_response = requests.post(generate_url, headers=headers, json=data)
generate_result = generate_response.json()
prediction_id = generate_result["data"]["id"]

# Step 2: Poll for result
poll_url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"

def check_status():
    while True:
        response = requests.get(poll_url, headers={"Authorization": "Bearer $ATLASCLOUD_API_KEY"})
        result = response.json()

        if result["data"]["status"] in ["completed", "succeeded"]:
            print("Generated video:", result["data"]["outputs"][0])
            return result["data"]["outputs"][0]
        elif result["data"]["status"] == "failed":
            raise Exception(result["data"]["error"] or "Generation failed")
        else:
            # Still processing, wait 2 seconds
            time.sleep(2)

video_url = check_status()

ติดตั้ง

ติดตั้งแพ็กเกจที่จำเป็น

pip install requests

การยืนยันตัวตน

คำขอ API ทั้งหมดต้องมีการยืนยันตัวตนผ่าน API key คุณสามารถรับ API key ได้จากแดชบอร์ด Atlas Cloud

export ATLASCLOUD_API_KEY="your-api-key-here"

HTTP Headers

import os

API_KEY = os.environ.get("ATLASCLOUD_API_KEY")
headers = {
    "Content-Type": "application/json",
    "Authorization": f"Bearer {API_KEY}"
}

รักษา API key ของคุณให้ปลอดภัย

อย่าเปิดเผย API key ของคุณในโค้ดฝั่งไคลเอนต์หรือที่เก็บข้อมูลสาธารณะ ให้ใช้ตัวแปรสภาพแวดล้อมหรือพร็อกซีฝั่งเซิร์ฟเวอร์แทน

ส่งคำขอ

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "your-model",
    "prompt": "A beautiful landscape"
}

response = requests.post(url, headers=headers, json=data)
print(response.json())

ส่งคำขอ

ส่งคำขอสร้างแบบอะซิงโครนัส API จะส่งคืน prediction ID ที่คุณสามารถใช้ตรวจสอบสถานะและดึงผลลัพธ์ได้

POST/api/v1/model/generateVideo

เนื้อหาคำขอ

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}

data = {
    "model": "bytedance/seedance-2.0/image-to-video",
    "prompt": "A beautiful sunset over the ocean with gentle waves"
}

response = requests.post(url, headers=headers, json=data)
result = response.json()

print(f"Prediction ID: {result['data']['id']}")
print(f"Status: {result['data']['status']}")

การตอบกลับ

{
  "code": 200,
  "data": {
    "id": "pred_abc123",
    "status": "processing",
    "model": "model-name",
    "created_at": "2025-01-01T00:00:00Z"
  }
}

ตรวจสอบสถานะ

ตรวจสอบสถานะปัจจุบันของคำขอด้วยการเรียก prediction endpoint เป็นระยะ

GET/api/v1/model/prediction/{prediction_id}

ตัวอย่างการตรวจสอบสถานะเป็นระยะ

import requests
import time

prediction_id = "pred_abc123"
url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

while True:
    response = requests.get(url, headers=headers)
    result = response.json()
    status = result["data"]["status"]
    print(f"Status: {status}")

    if status in ["completed", "succeeded"]:
        output_url = result["data"]["outputs"][0]
        print(f"Output URL: {output_url}")
        break
    elif status == "failed":
        print(f"Error: {result['data'].get('error', 'Unknown')}")
        break

    time.sleep(3)

ค่าสถานะ

processingคำขอยังอยู่ระหว่างการประมวลผล

completedการสร้างเสร็จสมบูรณ์แล้ว ผลลัพธ์พร้อมใช้งาน

succeededการสร้างสำเร็จแล้ว ผลลัพธ์พร้อมใช้งาน

failedการสร้างล้มเหลว ตรวจสอบฟิลด์ error

การตอบกลับที่เสร็จสมบูรณ์

{
  "data": {
    "id": "pred_abc123",
    "status": "completed",
    "outputs": [
      "https://storage.atlascloud.ai/outputs/result.mp4"
    ],
    "metrics": {
      "predict_time": 45.2
    },
    "created_at": "2025-01-01T00:00:00Z",
    "completed_at": "2025-01-01T00:00:10Z"
  }
}

อัปโหลดไฟล์

อัปโหลดไฟล์ไปยังที่เก็บข้อมูล Atlas Cloud และรับ URL ที่คุณสามารถใช้ในคำขอ API ของคุณ ใช้ multipart/form-data ในการอัปโหลด

POST/api/v1/model/uploadMedia

ตัวอย่างการอัปโหลด

import requests

url = "https://api.atlascloud.ai/api/v1/model/uploadMedia"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

with open("image.png", "rb") as f:
    files = {"file": ("image.png", f, "image/png")}
    response = requests.post(url, headers=headers, files=files)

result = response.json()
download_url = result["data"]["download_url"]
print(f"File URL: {download_url}")

การตอบกลับ

{
  "data": {
    "download_url": "https://storage.atlascloud.ai/uploads/abc123/image.png",
    "file_name": "image.png",
    "content_type": "image/png",
    "size": 1024000
  }
}

Input Schema

พารามิเตอร์ต่อไปนี้ยอมรับในเนื้อหาคำขอ

ทั้งหมด: 12จำเป็น: 2ไม่บังคับ: 10

modelstringrequired

Model name.

Default: "bytedance/seedance-2.0/image-to-video"

promptstring

Text prompt describing the desired video motion. Optional but recommended.

Default: "The scene comes alive with gentle motion and cinematic lighting"

imagestringrequired

First-frame image URL, Base64, or asset reference (asset://<ASSET_ID>). The video starts from this image. Formats: jpeg/png/webp/bmp/tiff/gif, dimensions (300,6000)px, aspect ratio (0.4,2.5), max 30MB.

last_imagestring

Last-frame image URL, Base64, or asset reference. The video transitions from the first frame to this last frame. Same format requirements as image.

durationinteger

Video duration in seconds (4-15), or -1 for model to choose automatically.

Default: 5

-1456789101112131415

resolutionstring

Video resolution. "4k" is native UHD (3840×2160, 16:9) available on the full Seedance 2.0 models only (not Fast/Mini); 4K output is 10-bit, H.265/HEVC encoded and may not play in all browsers.

Default: "720p"

480p720p720p-SR1080p1080p-SR1440p-SR4k

ratiostring

Aspect ratio. 'adaptive' matches the first frame image aspect ratio.

Default: "adaptive"

16:94:31:13:49:1621:9adaptive

bitrate_modestring

Output video bitrate mode. 'high' encodes at a higher bitrate for a crisper, larger file; 'standard' uses the normal bitrate. Does not affect token cost.

Default: "standard"

standardhigh

generate_audioboolean

Whether to generate synchronized audio.

Default: true

seedinteger

Seed integer used to control the randomness of generated content. Value range: [-1, 2^32-1]. The default -1 means a random seed is used. The same seed with the same request produces similar results, but complete consistency is not guaranteed.

Default: -1Min: -1Max: 4294967295

watermarkboolean

Whether to add a watermark.

Default: false

return_last_frameboolean

Whether to return the last frame as a separate image.

Default: false

ตัวอย่างเนื้อหาคำขอ

{
  "model": "bytedance/seedance-2.0/image-to-video",
  "prompt": "The scene comes alive with gentle motion and cinematic lighting",
  "image": "example_image",
  "duration": 5,
  "resolution": "720p",
  "ratio": "adaptive",
  "bitrate_mode": "standard",
  "generate_audio": true,
  "seed": -1,
  "watermark": false,
  "return_last_frame": false
}

Output Schema

API จะส่งคืนการตอบกลับ prediction พร้อม URL ของผลลัพธ์ที่สร้างขึ้น

idstring

Unique identifier for the prediction.

modelstring

Model ID used for the prediction.

statusstring

Status: processing, completed, failed, or timeout.

outputsarray

URLs to generated content (video + optional last frame).

created_atstring

ISO timestamp of creation.

completion_tokensinteger

Tokens consumed for billing.

total_tokensinteger

Total tokens consumed.

ตัวอย่างการตอบกลับ

{
  "id": "pred_abc123",
  "status": "completed",
  "model": "model-name",
  "outputs": [
    "https://storage.atlascloud.ai/outputs/result.mp4"
  ],
  "metrics": {
    "predict_time": 45.2
  },
  "created_at": "2025-01-01T00:00:00Z",
  "completed_at": "2025-01-01T00:00:10Z"
}

Atlas Cloud Skills

Atlas Cloud Skills เชื่อมต่อโมเดล AI กว่า 400+ เข้ากับผู้ช่วยเขียนโค้ด AI ของคุณโดยตรง ติดตั้งด้วยคำสั่งเดียว จากนั้นใช้ภาษาธรรมชาติเพื่อสร้างรูปภาพ วิดีโอ และสนทนากับ LLM

ไคลเอนต์ที่รองรับ

Claude Code

OpenAI Codex

Gemini CLI

Cursor

Windsurf

VS Code

Trae

GitHub Copilot

Cline

Roo Code

Amp

Goose

Replit

40+ ไคลเอนต์ที่รองรับ

ติดตั้ง

npx skills add AtlasCloudAI/atlas-cloud-skills

ตั้งค่า API Key

รับ API key จากแดชบอร์ด Atlas Cloud และตั้งค่าเป็นตัวแปรสภาพแวดล้อม

export ATLASCLOUD_API_KEY="your-api-key-here"

ความสามารถ

เมื่อติดตั้งแล้ว คุณสามารถใช้ภาษาธรรมชาติในผู้ช่วย AI ของคุณเพื่อเข้าถึงโมเดล Atlas Cloud ทั้งหมด

สร้างรูปภาพสร้างรูปภาพด้วยโมเดลเช่น Nano Banana 2, Z-Image และอื่นๆ

สร้างวิดีโอสร้างวิดีโอจากข้อความหรือรูปภาพด้วย Kling, Vidu, Veo เป็นต้น

สนทนา LLMสนทนากับ Qwen, DeepSeek และโมเดลภาษาขนาดใหญ่อื่นๆ

อัปโหลดสื่ออัปโหลดไฟล์จากเครื่องสำหรับการแก้ไขรูปภาพและเวิร์กโฟลว์รูปภาพเป็นวิดีโอ

เรียนรู้เพิ่มเติม

github.com/AtlasCloudAI/atlas-cloud-skills

MCP Server

Atlas Cloud MCP Server เชื่อมต่อ IDE ของคุณกับโมเดล AI กว่า 400+ ผ่าน Model Context Protocol ใช้งานได้กับไคลเอนต์ที่รองรับ MCP ทุกตัว

ไคลเอนต์ที่รองรับ

Cursor

VS Code

Windsurf

Claude Code

OpenAI Codex

Gemini CLI

Cline

Roo Code

100+ ไคลเอนต์ที่รองรับ

ติดตั้ง

npx -y atlascloud-mcp

การกำหนดค่า

เพิ่มการกำหนดค่าต่อไปนี้ลงในไฟล์ตั้งค่า MCP ของ IDE ของคุณ

{
  "mcpServers": {
    "atlascloud": {
      "command": "npx",
      "args": [
        "-y",
        "atlascloud-mcp"
      ],
      "env": {
        "ATLASCLOUD_API_KEY": "your-api-key-here"
      }
    }
  }
}

เครื่องมือที่ใช้ได้

atlas_generate_imageสร้างรูปภาพจากข้อความ prompt

atlas_generate_videoสร้างวิดีโอจากข้อความหรือรูปภาพ

atlas_chatสนทนากับโมเดลภาษาขนาดใหญ่

atlas_list_modelsเรียกดูโมเดล AI กว่า 400+ ที่ใช้ได้

atlas_quick_generateสร้างเนื้อหาขั้นตอนเดียวพร้อมเลือกโมเดลอัตโนมัติ

atlas_upload_mediaอัปโหลดไฟล์จากเครื่องสำหรับเวิร์กโฟลว์ API

เรียนรู้เพิ่มเติม

github.com/AtlasCloudAI/mcp-server

API Schema

{
  "info": {
    "title": "AtlasCloud API",
    "version": "1.0.0",
    "description": "The AtlasCloud API."
  },
  "paths": {
    "/api/v1/model/prediction/{request_id}": {
      "get": {
        "parameters": [
          {
            "in": "path",
            "name": "request_id",
            "required": true,
            "schema": {
              "description": "Request ID",
              "type": "string"
            }
          }
        ],
        "responses": {
          "200": {
            "content": {
              "application/json": {
                "schema": {
                  "$ref": "#/components/schemas/PredictionResponse"
                }
              }
            },
            "description": "Result of the request."
          }
        }
      },
      "x-api-name": "model_result"
    },
    "/api/v1/model/generateVideo": {
      "post": {
        "requestBody": {
          "content": {
            "application/json": {
              "schema": {
                "$ref": "#/components/schemas/Input"
              }
            }
          },
          "required": true
        },
        "responses": {
          "200": {
            "content": {
              "application/json": {
                "schema": {
                  "$ref": "#/components/schemas/PredictionResponse"
                }
              }
            },
            "description": "The request status."
          }
        }
      },
      "x-api-name": "model_run"
    }
  },
  "openapi": "3.0.0",
  "servers": [
    {
      "url": "https://api.atlascloud.ai"
    }
  ],
  "components": {
    "schemas": {
      "Input": {
        "type": "object",
        "required": [
          "model",
          "image"
        ],
        "properties": {
          "model": {
            "type": "string",
            "description": "Model name.",
            "default": "bytedance/seedance-2.0/image-to-video"
          },
          "prompt": {
            "type": "string",
            "default": "The scene comes alive with gentle motion and cinematic lighting",
            "description": "Text prompt describing the desired video motion. Optional but recommended."
          },
          "image": {
            "type": "string",
            "description": "First-frame image URL, Base64, or asset reference (asset://<ASSET_ID>). The video starts from this image. Formats: jpeg/png/webp/bmp/tiff/gif, dimensions (300,6000)px, aspect ratio (0.4,2.5), max 30MB."
          },
          "last_image": {
            "type": "string",
            "description": "Last-frame image URL, Base64, or asset reference. The video transitions from the first frame to this last frame. Same format requirements as image."
          },
          "duration": {
            "type": "integer",
            "default": 5,
            "enum": [
              -1,
              4,
              5,
              6,
              7,
              8,
              9,
              10,
              11,
              12,
              13,
              14,
              15
            ],
            "description": "Video duration in seconds (4-15), or -1 for model to choose automatically.",
            "x-ui-component": "slider"
          },
          "resolution": {
            "type": "string",
            "default": "720p",
            "enum": [
              "480p",
              "720p",
              "720p-SR",
              "1080p",
              "1080p-SR",
              "1440p-SR",
              "4k"
            ],
            "description": "Video resolution. \"4k\" is native UHD (3840×2160, 16:9) available on the full Seedance 2.0 models only (not Fast/Mini); 4K output is 10-bit, H.265/HEVC encoded and may not play in all browsers."
          },
          "ratio": {
            "type": "string",
            "default": "adaptive",
            "enum": [
              "16:9",
              "4:3",
              "1:1",
              "3:4",
              "9:16",
              "21:9",
              "adaptive"
            ],
            "description": "Aspect ratio. 'adaptive' matches the first frame image aspect ratio."
          },
          "bitrate_mode": {
            "type": "string",
            "default": "standard",
            "enum": [
              "standard",
              "high"
            ],
            "description": "Output video bitrate mode. 'high' encodes at a higher bitrate for a crisper, larger file; 'standard' uses the normal bitrate. Does not affect token cost."
          },
          "generate_audio": {
            "type": "boolean",
            "default": true,
            "description": "Whether to generate synchronized audio."
          },
          "seed": {
            "type": "integer",
            "default": -1,
            "minimum": -1,
            "maximum": 4294967295,
            "description": "Seed integer used to control the randomness of generated content. Value range: [-1, 2^32-1]. The default -1 means a random seed is used. The same seed with the same request produces similar results, but complete consistency is not guaranteed."
          },
          "watermark": {
            "type": "boolean",
            "default": false,
            "description": "Whether to add a watermark."
          },
          "return_last_frame": {
            "type": "boolean",
            "default": false,
            "description": "Whether to return the last frame as a separate image."
          }
        },
        "x-order-properties": [
          "model",
          "image",
          "last_image",
          "prompt",
          "duration",
          "resolution",
          "ratio",
          "bitrate_mode",
          "generate_audio",
          "watermark",
          "return_last_frame"
        ]
      },
      "PredictionResponse": {
        "type": "object",
        "properties": {
          "id": {
            "type": "string",
            "description": "Unique identifier for the prediction."
          },
          "urls": {
            "type": "object",
            "description": "Object containing related API endpoints."
          },
          "model": {
            "type": "string",
            "description": "Model ID used for the prediction."
          },
          "status": {
            "type": "string",
            "description": "Status: processing, completed, failed, or timeout."
          },
          "outputs": {
            "type": "array",
            "items": {
              "type": "string"
            },
            "description": "URLs to generated content (video + optional last frame)."
          },
          "created_at": {
            "type": "string",
            "format": "date-time",
            "description": "ISO timestamp of creation."
          },
          "completion_tokens": {
            "type": "integer",
            "description": "Tokens consumed for billing."
          },
          "total_tokens": {
            "type": "integer",
            "description": "Total tokens consumed."
          },
          "has_nsfw_contents": {
            "type": "array",
            "items": {
              "type": "boolean"
            },
            "description": "NSFW detection per output."
          }
        }
      }
    },
    "securitySchemes": {
      "apiKeyAuth": {
        "in": "header",
        "name": "Authorization",
        "type": "apiKey"
      }
    }
  }
}

เทมเพลต Prompt สำหรับ LLM

# bytedance/seedance-2.0/image-to-video

> Generate videos from a first-frame image (and optional last-frame) with native audio.


## Overview

- **Submit endpoint (POST)**: `https://api.atlascloud.ai/api/v1/model/generateVideo` — start an async generation; returns a `prediction_id`
- **Poll endpoint (GET)**: `https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}` — poll this until the prediction finishes
- **Model ID**: `bytedance/seedance-2.0/image-to-video`


## API Information

This model can be used via our HTTP API or more conveniently via our client libraries.
See the input and output schema below, as well as the usage examples.


### Input Schema

The API accepts the following input parameters:

- **`model`** (`string`, _required_):
  Model name.
  - Default: `"bytedance/seedance-2.0/image-to-video"`

- **`image`** (`string`, _required_):
  First-frame image URL, Base64, or asset reference (asset://<ASSET_ID>). The video starts from this image. Formats: jpeg/png/webp/bmp/tiff/gif, dimensions (300,6000)px, aspect ratio (0.4,2.5), max 30MB.

- **`last_image`** (`string`, _optional_):
  Last-frame image URL, Base64, or asset reference. The video transitions from the first frame to this last frame. Same format requirements as image.

- **`prompt`** (`string`, _optional_):
  Text prompt describing the desired video motion. Optional but recommended.
  - Default: `"The scene comes alive with gentle motion and cinematic lighting"`

- **`duration`** (`integer`, _optional_):
  Video duration in seconds (4-15), or -1 for model to choose automatically.
  - Default: `5`
  - Options: -1, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15

- **`resolution`** (`string`, _optional_):
  Video resolution. "4k" is native UHD (3840×2160, 16:9) available on the full Seedance 2.0 models only (not Fast/Mini); 4K output is 10-bit, H.265/HEVC encoded and may not play in all browsers.
  - Default: `"720p"`
  - Options: "480p", "720p", "720p-SR", "1080p", "1080p-SR", "1440p-SR", "4k"

- **`ratio`** (`string`, _optional_):
  Aspect ratio. 'adaptive' matches the first frame image aspect ratio.
  - Default: `"adaptive"`
  - Options: "16:9", "4:3", "1:1", "3:4", "9:16", "21:9", "adaptive"

- **`bitrate_mode`** (`string`, _optional_):
  Output video bitrate mode. 'high' encodes at a higher bitrate for a crisper, larger file; 'standard' uses the normal bitrate. Does not affect token cost.
  - Default: `"standard"`
  - Options: "standard", "high"

- **`generate_audio`** (`boolean`, _optional_):
  Whether to generate synchronized audio.
  - Default: `true`

- **`watermark`** (`boolean`, _optional_):
  Whether to add a watermark.
  - Default: `false`

- **`return_last_frame`** (`boolean`, _optional_):
  Whether to return the last frame as a separate image.
  - Default: `false`



**Required Parameters Example**:

```json
{
  "model": "bytedance/seedance-2.0/image-to-video",
  "image": ""
}
```


**Full Example**:

```json
{
  "model": "bytedance/seedance-2.0/image-to-video",
  "image": "",
  "last_image": "",
  "prompt": "The scene comes alive with gentle motion and cinematic lighting",
  "duration": 5,
  "resolution": "720p",
  "ratio": "adaptive",
  "bitrate_mode": "standard",
  "generate_audio": true,
  "watermark": false,
  "return_last_frame": false
}
```


### Output Schema

The API returns the following output format:


- **`id`** (`string`, _optional_):
  Unique identifier for the prediction.

- **`urls`** (`object`, _optional_):
  Object containing related API endpoints.

- **`model`** (`string`, _optional_):
  Model ID used for the prediction.

- **`status`** (`string`, _optional_):
  Status: processing, completed, failed, or timeout.

- **`outputs`** (`array[string]`, _optional_):
  URLs to generated content (video + optional last frame).

- **`created_at`** (`string`, _optional_):
  ISO timestamp of creation.

- **`completion_tokens`** (`integer`, _optional_):
  Tokens consumed for billing.

- **`total_tokens`** (`integer`, _optional_):
  Total tokens consumed.

- **`has_nsfw_contents`** (`array[boolean]`, _optional_):
  NSFW detection per output.



**Example Response**:

```json
{
  "id": "",
  "urls": {},
  "model": "",
  "status": "",
  "outputs": [
    ""
  ],
  "created_at": "",
  "completion_tokens": 0,
  "total_tokens": 0,
  "has_nsfw_contents": []
}
```


## Usage Examples

### cURL

```bash
# Step 1: Start generation (async)
curl -X POST "https://api.atlascloud.ai/api/v1/model/generateVideo" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "bytedance/seedance-2.0/image-to-video",
  "image": "",
  "last_image": "",
  "prompt": "The scene comes alive with gentle motion and cinematic lighting",
  "duration": 5,
  "resolution": "720p",
  "ratio": "adaptive",
  "bitrate_mode": "standard",
  "generate_audio": true,
  "watermark": false,
  "return_last_frame": false
}'

# Response will contain: {"code": 200, "data": {"id": "prediction_id", "status": "processing"}}

# Step 2: Poll for result (replace {prediction_id} with the id returned above)
curl -X GET "https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY"

# Keep polling until status is "completed", "succeeded" or "failed"
# When completed, outputs will contain the generated content URL(s)
```


## Asset Library (subject reference assets)

This model supports **subject reference inputs** — upload a subject reference image to the Asset Library first, then reference it from image fields (e.g. `reference_images`, `image`, `last_image`). Instead of a regular image URL, pass `asset://<atlas_asset_id>` once the subject reference has been registered. Once an asset reaches `Active` status, it can be reused across many generation calls.

### ⚠️ Important: Different Base URL

Asset Library endpoints live on the **console** host, NOT on the inference host where generation endpoints live.

- Asset Library API: `https://console.atlascloud.ai/api/v1/sd/assets/...`
- Generation API:       `https://api.atlascloud.ai/api/v1/model/...`

Do not call `https://api.atlascloud.ai/api/v1/sd/assets` — it will return 404. Both hosts use the same `Authorization: Bearer <api-key>` header.

### Base URL

`https://console.atlascloud.ai/api/v1`

All requests require `Authorization: Bearer <api-key>`.

### Endpoints

| Method | Path | Description |
|--------|------|-------------|
| POST   | `/sd/assets`               | Create a new portrait asset from a public image URL |
| GET    | `/sd/assets/{id}`          | Get a single asset (auto-polls upstream when Processing) |
| GET    | `/sd/assets`               | List assets (paginated, supports filters) |
| PUT    | `/sd/assets/{id}`          | Rename an asset |
| DELETE | `/sd/assets/{id}`          | Move asset to trash (recoverable) |
| GET    | `/sd/assets/trash`         | List deleted assets |
| POST   | `/sd/assets/{id}/restore`  | Restore a deleted asset from trash |

### Create Asset

`POST /sd/assets`

**Body fields**:
- `url` (string, **required**): Publicly accessible image URL.
- `name` (string, optional): Display name. Max 64 characters (UTF-8 bytes).
- `asset_type` (string, optional): Default `Image`.

**Image requirements**: JPEG / PNG / WebP / BMP / TIFF / GIF / HEIC, side length 300–6000px, aspect ratio 0.4–2.5, max 30 MB.

**Example response**:
```json
{
  "code": "200",
  "data": {
    "id": 1,
    "atlas_asset_id": "asset-2026xxxx-xxxxx",
    "name": "My Portrait",
    "status": "Processing",
    "created_at": 1712467200000
  }
}
```

### Get Asset

`GET /sd/assets/{id}` — returns the asset details. If the asset is still `Processing`, the API automatically checks the upstream status and returns the latest state, so clients only need to poll this single endpoint.

### List Assets

`GET /sd/assets`

**Query parameters**:
- `page_number` (int): Page number, default `1`.
- `page_size` (int): Results per page, default `20`, max `100`.
- `status` (string): Filter — one of `Processing`, `Active`, `Failed`.
- `name` (string): Fuzzy search by name.
- `created_from` (int64): Created at or after this unix-millis timestamp.
- `created_to` (int64): Created at or before this unix-millis timestamp.

### Update Asset

`PUT /sd/assets/{id}` — body: `{ "name": "new name" }` (max 64 chars).

### Delete & Restore

- `DELETE /sd/assets/{id}` — moves the asset to trash. The asset is hidden from list endpoints but can be restored later.
- `GET /sd/assets/trash` — lists deleted assets. Same `page_number` / `page_size` / `name` filters as the regular list endpoint.
- `POST /sd/assets/{id}/restore` — restores a deleted asset from trash.

### Asset Lifecycle

```
Create → Processing → Active   (ready for video generation)
                    → Failed   (check error_code / error_message)
```

- **Processing** — Image is being preprocessed. Poll `GET /sd/assets/{id}` until it transitions.
- **Active** — Ready to use. Reference it as `asset://<atlas_asset_id>` in any image field of a video generation request.
- **Failed** — Preprocessing failed. Inspect `error_code` and `error_message` on the asset (e.g. unsupported format, oversized image, no face detected).

### Quick Start (cURL)

```bash
# 1. Create an asset
curl -X POST https://console.atlascloud.ai/api/v1/sd/assets \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"url": "https://example.com/portrait.jpg", "name": "My Portrait"}'

# 2. Poll until status is "Active"
curl https://console.atlascloud.ai/api/v1/sd/assets/1 \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY"

# 3. Use the atlas_asset_id in a video generation request
#    Pass it in any image field as: asset://<atlas_asset_id>
```

### Using Assets in Video Generation

Once `Active`, pass `asset://<atlas_asset_id>` wherever the model schema accepts an image URL — for example in `reference_images`, `image`, or `last_image`. Example body:

```json
{
  "model": "bytedance/seedance-2.0/image-to-video",
  "prompt": "The character in image 1 dances gracefully to the music",
  "reference_images": [
    "asset://asset-2026xxxx-xxxxx"
  ]
}
```

### Error Codes

| HTTP Status | Meaning |
|-------------|---------|
| 200 | Success |
| 400 | Invalid request (e.g. malformed body, name too long, image fetch failed) |
| 401 | Missing or invalid API key |
| 404 | Asset not found |
| 500 | Server error |

### Content Review

All uploaded portraits go through automatic moderation. Non-compliant content will be blocked — not every image can be uploaded. Make sure you own the rights to the material, that it does not impersonate any real person, and that it does not violate laws, public order, or third-party intellectual property.

## Additional Resources

### Documentation

- [Model Playground](https://www.atlascloud.ai/models/bytedance/seedance-2.0/image-to-video)

A sleek futuristic spaceship slowly orbiting a gigantic planet, the planet’s glowing atmosphere and clouds visible from space, starfield and nebula in the background, smooth orbital movement, cinematic sci-fi scene, epic scale, volumetric lighting, ultra-realistic, 4K, slow camera tracking.

A cinematic product shot with smooth camera movement, detailed lighting, and stable motion.

กำลังโหลด...

1. Introduction

Seedance 2.0 is a state-of-the-art multimodal generative AI model designed for synchronized video and audio content creation. Developed by ByteDance and integrated into the CapCut/Dreamina platform as of March 2026, this model family advances the field of generative multimedia by combining sophisticated diffusion transformer architectures with physics-informed world modeling for realistic motion and spatial consistency.

Seedance 2.0’s significance lies in its Dual-Branch Diffusion Transformer (DB-DiT) architecture that jointly processes video and audio streams, enabling phoneme-level lip synchronization across multiple languages. Compared to previous iterations, it achieves substantially higher output usability rates and faster generation speeds. The two variants target different workloads: Seedance 2.0 delivers high-fidelity, cinematic-quality renders with enhanced lighting and texture detail, while Seedance 2.0 Fast provides a cost-effective, accelerated pipeline optimized for high throughput and rapid prototyping.

2. Key Features & Innovations

Dual-Branch Diffusion Transformer Architecture: Seedance 2.0 integrates separate yet synchronized diffusion branches for video and audio, enabling tight coupling between visual motion and sound generation. This architecture improves motion realism and audio-visual coherence beyond previous generative models.
World Model with Physics Simulation: The model incorporates a physics-based world modeling approach that simulates realistic object motion and spatial consistency over time. This leads to naturalistic dynamics and stable scene composition across generated video sequences.
Rich Multimodal Input Support: Seedance 2.0 accepts diverse input formats including text prompts, up to 9 images, and up to 3 video or audio clips of 15 seconds each. This flexibility allows nuanced content creation workflows combining static, dynamic, and auditory cues.
Phoneme-Level Lip Synchronization: The native audio generation pipeline supports lip-sync at the phoneme granularity in 8+ languages, ensuring high fidelity mouth movements closely match generated speech or singing.
High Usability and Efficiency: The model achieves an estimated 90% usable output rate compared to an industry average of approximately 20%, reducing post-processing overhead. Additionally, it delivers a 30% inference speed advantage over predecessor systems.
API Variants for Different Use Cases: The Seedance 2.0 endpoint is geared toward high fidelity and cinematic visual effects suitable for final production, while the Seedance 2.0 Fast variant offers roughly 3 times faster generation and approximately 91% cost savings at $0.022 per second of output, ideal for rapid iteration and volume workflows.

3. Model Architecture & Technical Details

Seedance 2.0 is built around the Dual-Branch Diffusion Transformer (DB-DiT), which separately processes video and audio streams via transformer-based denoising diffusion models while synchronizing generation steps to enforce audio-visual alignment. The system leverages a World Model that integrates physics simulation modules, enabling consistent spatial and temporal object behaviors within video sequences.

Training was conducted in multiple stages on large-scale, diverse datasets spanning images, videos, text captions, and audio recordings across multiple languages. Initial large-scale pre-training utilized resolutions spanning from 720p to 1080p, followed by supervised fine-tuning (SFT) to improve text and visual prompt conditioning fidelity. Reinforcement Learning with Human Feedback (RLHF) optimized multi-dimensional reward models that simultaneously assess aesthetics, motion coherence, and audio-visual synchronization quality.

The training pipeline supports multiple aspect ratios including 9:16, 16:9, 1:1, and 4:3, and target output lengths from 4 to 60 seconds. Specialized modules enable the @ reference system for fine-grained control of creative elements based on provided input assets.

4. Performance Highlights

Seedance 2.0 was benchmarked on the comprehensive SeedVideoBench-2.0 suite, which evaluates generative video models across over 50 image-based and 24 video-based benchmarks covering diverse content domains and multi-modal tasks.

Rank	Model	Developer	Score/Metric	Release Date
1	Kling 3.0	External	Competitive	2025
2	Sora 2	External	Competitive	2025
3	Seedance 2.0	ByteDance	High audiovisual sync, motion realism	2026
4	Veo 3.1	External	Strong baseline	2025

Seedance 2.0 matches or exceeds these contemporary models in synchronized video-audio generation, demonstrating especially strong performance in phoneme-level lip synchronization and motion naturalism thanks to the World Model component. Its 30% speed improvement and 90% output usability rate reflect notable efficiency advancements.

5. Intended Use & Applications

Social Media Content Creation: Efficiently generate engaging short videos with synchronized audio and visually rich effects, tailored for platforms like TikTok and Instagram.
E-commerce Product Videos: Automatically produce dynamic product showcases combining text, image, and video inputs with realistic motion and sound to enhance online shopping experiences.
Marketing Campaigns: Craft high-quality cinematic promotional content that integrates brand assets via the @ reference system for tailored storytelling and audience engagement.
Music Videos: Generate synchronized visuals with phoneme-accurate lip-syncing for multilingual vocal tracks to support artist and record label promotional needs.
Short Narrative Films: Create compelling narrative-driven video clips with coherent motion and spatial consistency, supporting indie filmmakers and content creators.
Fashion and Luxury Showcases: Produce visually detailed and aesthetic presentations incorporating texture and lighting refinements for high-end brand communications.

Super Resolution

This model supports FlashVSR-backed super-resolution tiers through the resolution parameter:

Resolution	Behavior
`1080p-SR`	Generates a 720p source, then applies FlashVSR to a 1080p target.
`1440p-SR`	Generates a 720p source, then applies FlashVSR to a 1440p QHD target.

Use the SR tiers when you want sharper edges, cleaner texture retention, or a lower-cost HD option compared with native high-resolution generation. Final billing follows the active model pricing configuration for the selected resolution, duration, account, and environment.

Seedance 2.0 Image-to-Video API by ByteDance

อินพุต

เอาต์พุต

พารามิเตอร์

ตัวอย่างโค้ด

ติดตั้ง

การยืนยันตัวตน

HTTP Headers

ส่งคำขอ

ส่งคำขอ

เนื้อหาคำขอ

การตอบกลับ

ตรวจสอบสถานะ

ตัวอย่างการตรวจสอบสถานะเป็นระยะ

ค่าสถานะ

การตอบกลับที่เสร็จสมบูรณ์

อัปโหลดไฟล์

ตัวอย่างการอัปโหลด

การตอบกลับ

Input Schema

ตัวอย่างเนื้อหาคำขอ

Output Schema

ตัวอย่างการตอบกลับ

Atlas Cloud Skills

ไคลเอนต์ที่รองรับ

ติดตั้ง

ตั้งค่า API Key

ความสามารถ

MCP Server

ไคลเอนต์ที่รองรับ

ติดตั้ง

การกำหนดค่า

เครื่องมือที่ใช้ได้

API Schema

เทมเพลต Prompt สำหรับ LLM

1. Introduction

2. Key Features & Innovations

3. Model Architecture & Technical Details

4. Performance Highlights

5. Intended Use & Applications

Super Resolution

สำรวจโมเดลที่คล้ายกัน

Seedance 2.0 Mini Reference-to-Video

Seedance 2.0 Mini Image-to-Video

Seedance 2.0 Mini Text-to-Video

Avatar Omni Human 1.5

Seedance 2.0 Fast Reference-to-Video

Seedance 2.0 Fast Image-to-Video

Seedance 2.0 Fast Text-to-Video

Seedance 2.0 Reference-to-Video

Seedance 2.0 Text-to-Video

Seedance v1.5 Pro Image-to-Video

Seedance v1.5 Pro Text-to-Video

Seedance v1.5 Pro Image-to-Video Fast

Seedance v1.5 Pro Text-to-Video Fast

Seedance v1 Pro Fast Text-to-video

Seedance v1 Pro Fast Image-to-video

Seedance v1 Pro t2v 1080p

API เดียวสำหรับ AI สื่อทุกประเภท

Join our Discord community

อินพุต

เอาต์พุต

พารามิเตอร์

ตัวอย่างโค้ด

ติดตั้ง

การยืนยันตัวตน

HTTP Headers

ส่งคำขอ

ส่งคำขอ

เนื้อหาคำขอ

การตอบกลับ

ตรวจสอบสถานะ

ตัวอย่างการตรวจสอบสถานะเป็นระยะ

ค่าสถานะ

การตอบกลับที่เสร็จสมบูรณ์

อัปโหลดไฟล์

ตัวอย่างการอัปโหลด

การตอบกลับ

Input Schema

ตัวอย่างเนื้อหาคำขอ