Trung tâm AI tất cả trong một

Chào mừng bạn đến với trung tâm công cụ AI của bạn. Được thiết kế để duyệt mượt mà và triển khai nhanh chóng, trang này tập hợp toàn bộ bộ trí tuệ của chúng tôi — từ các mô hình ngôn ngữ lớn (LLM) mạnh mẽ đến các trình tạo ảnh và video tiên tiến. Tại đây, bạn có thể đánh giá và sử dụng đúng công cụ cho quy trình của mình chỉ trong một cái nhìn.

Tìm kiếm phổ biến:
DANH MỤC
Giảm giá (140)
Chức năng Mô hình
Dòng
48 trong số 395 mô hình
Mới
Qwen Image 3.0 Text-to-Image
NEW
Văn bản-Hình ảnh

Qwen Image 3.0 Text-to-Image

Generates images from a text prompt at resolutions up to 2048×2048, with automatic prompt rewriting and prompt-guided resolution selection, building on Qwen strength in complex text rendering and precise prompt adherence

From
$0.04/HÌNH ẢNH
Qwen Image 3.0 Edit
NEW
Hình ảnh-Hình ảnh

Qwen Image 3.0 Edit

Edits images from one to three reference images and a natural-language instruction, preserving key details such as facial features and identity while applying the requested changes

From
$0.04/HÌNH ẢNH
Qwen3.8 Max
NEW
LLM

Qwen3.8 Max

Next-generation flagship model for advanced reasoning, coding, and multimodal AI applications.

FP8
From
$2/6M Đầu vào/Đầu ra
MiniMax H3 Text-to-Video
NEW
Văn bản-Video

MiniMax H3 Text-to-Video

MiniMax H3 text-to-video: generate a cinematic video from a text prompt. Supports 2K, 5-15s., and 16:9/9:16/1:1/adaptive aspect ratios.

From
$0.14/GIÂY
MiniMax H3 Image-to-Video
NEW
Hình ảnh-Video

MiniMax H3 Image-to-Video

MiniMax H3 image-to-video: animate a first-frame image (optionally with a last frame) driven by a text prompt. Supports 2K, 5-15s.

From
$0.14/GIÂY
MiniMax H3 Reference-to-Video
NEW
Hình ảnh-Video

MiniMax H3 Reference-to-Video

MiniMax H3 reference-to-video: generate a video that keeps the subject from a reference image, driven by a text prompt. Supports 2K, 5-15s.

From
$0.14/GIÂY
Reve 2.1 Remix
NEW
Hình ảnh-Hình ảnh

Reve 2.1 Remix

Reve 2.1 Remix composes one to six reference images with a natural-language prompt into a single coherent image at native 4K, blending subject, style, and background while keeping references consistent.

From
$0.24/HÌNH ẢNH
Reve 2.1 Edit
NEW
Hình ảnh-Hình ảnh

Reve 2.1 Edit

Reve 2.1 Edit applies precise, instruction-driven, element-level edits to a single input image at native 4K, changing targeted regions while preserving the rest of the scene.

From
$0.24/HÌNH ẢNH
Reve 2.1 Text-to-Image
NEW
Văn bản-Hình ảnh

Reve 2.1 Text-to-Image

Reve 2.1 is a layout-first text-to-image model that turns natural-language prompts into sharp, production-ready images at native 4K, with best-in-class typography and high prompt adherence.

From
$0.24/HÌNH ẢNH
Kimi K3
NEW
HOT
LLM

Kimi K3

Top-performing open-weight model optimized for frontier reasoning, coding, and enterprise AI applications.

INT4CODE
From
$3/15M Đầu vào/Đầu ra
Grok 4.5
NEW
HOT
LLM

Grok 4.5

Flagship conversational model built for real-time knowledge exploration, sharp reasoning, and highly engaging AI interactions.

From
$2/6M Đầu vào/Đầu ra
Youchuan V8.2 Image-to-Video
NEW
Hình ảnh-Video

Youchuan V8.2 Image-to-Video

Youchuan V8.2 animates an input image into four 5-second videos at 480p or 720p.

From
$0.086/GIÂY
Youchuan V8.2 Remove Background
NEW
Hình ảnh-Hình ảnh

Youchuan V8.2 Remove Background

Youchuan automatically removes the background from an input image, returning one transparent-background result.

From
$0.086/HÌNH ẢNH
Youchuan V8.2 Style Transfer
NEW
Hình ảnh-Hình ảnh

Youchuan V8.2 Style Transfer

Youchuan retexture changes the artistic style of an input image while preserving its composition, returning four restyled results.

From
$0.129/HÌNH ẢNH
Youchuan V8.2 Blend
NEW
Hình ảnh-Hình ảnh

Youchuan V8.2 Blend

Youchuan V8.2 blends two to five input images into four fused results, with an optional guiding prompt and native 2K HD.

From
$0.086/HÌNH ẢNH
Youchuan V8.2 Image-to-Image
NEW
Hình ảnh-Hình ảnh

Youchuan V8.2 Image-to-Image

Youchuan V8.2 re-imagines an input image guided by a text prompt, returning four variations. Supports native 2K HD, style reference, and aspect-ratio / stylize / chaos / weird controls.

From
$0.086/HÌNH ẢNH
Youchuan V8.2 Text-to-Image
NEW
Văn bản-Hình ảnh

Youchuan V8.2 Text-to-Image

Youchuan V8.2 generates four images from a text prompt, with optional native 2K HD, a style reference, and aspect-ratio / stylize / chaos / weird controls.

From
$0.086/HÌNH ẢNH
Wan 2.7 Spicy Reference-to-Video
NEW
Tham chiếu-Video

Wan 2.7 Spicy Reference-to-Video

An image-reference-to-video model that generates one continuous video from one to four reference images with precise subject binding and prompt control. Supports image references only; reference video and audio are not accepted.

WAN-2.7SPICY
From
$0.1/GIÂY
Wan 2.7 Spicy Image-to-Video
NEW
HOT
Hình ảnh-Video

Wan 2.7 Spicy Image-to-Video

AtlasCloud Wan 2.7 Spicy Image-to-Video turns a first-frame image into short cinematic motion with stable temporal detail and expressive character movement.

From
$0.1/GIÂY
Seedream v5.0 Pro Edit
NEW
HOT
Hình ảnh-Hình ảnh
PRO

Seedream v5.0 Pro Edit

ByteDance flagship next-generation image editing model. Supports up to 10 reference images while preserving identity, lighting, and color tones for professional-quality modifications.

From
$0.045/HÌNH ẢNH
Seedream v5.0 Pro Text-to-Image
NEW
HOT
Văn bản-Hình ảnh
PRO

Seedream v5.0 Pro Text-to-Image

ByteDance flagship next-generation image generation model with stronger prompt adherence, refined typography, and photorealistic detail. Single-image output at 1.5K and 2K tiers with JPEG and PNG support.

From
$0.045/HÌNH ẢNH
KAT Coder Pro V2.5
NEW
HOT
LLM
PRO

KAT Coder Pro V2.5

High-end coding agent built for complex software engineering, repository-scale tasks, and autonomous development workflows.

CODE
From
$0.74/2.96M Đầu vào/Đầu ra
KAT Coder Air V2.5
NEW
HOT
LLM

KAT Coder Air V2.5

Fast, lightweight coding model designed for interactive development, rapid code iteration, and efficient coding agents.

CODE
From
$0.15/0.6M Đầu vào/Đầu ra
Nano Banana 2 Lite Edit Developer
NEW
Hình ảnh-Hình ảnh
DEV

Nano Banana 2 Lite Edit Developer

Google's fastest and most cost-efficient Nano Banana image model for editing, applying natural-language edits and multi-image composition to up to 14 reference images with low latency.

From$0.04/HÌNH ẢNH
$0.028/HÌNH ẢNH
-30%
Nano Banana 2 Lite Text-to-Image Developer
NEW
Văn bản-Hình ảnh
DEV

Nano Banana 2 Lite Text-to-Image Developer

Google's fastest and most cost-efficient Nano Banana image model, turning natural-language text prompts into high-quality 1k images in as little as 4 seconds for rapid, high-volume generation.

From$0.04/HÌNH ẢNH
$0.028/HÌNH ẢNH
-30%
Nano Banana 2 Lite Edit
NEW
Hình ảnh-Hình ảnh

Nano Banana 2 Lite Edit

Nano banana lite is the efficiency-focused model in the image generation family. Sub-2 second latency with cost-effective generation and editing, fast multi-turn local edits, and 14 supported aspect ratios.

From
$0.04/HÌNH ẢNH
Nano Banana 2 Lite Text-to-image
NEW
Văn bản-Hình ảnh

Nano Banana 2 Lite Text-to-image

Nano banana lite is the efficiency-focused model in the image generation family. Sub-2 second latency with cost-effective generation and editing, fast multi-turn local edits, and 14 supported aspect ratios.

From
$0.04/HÌNH ẢNH
Seed Audio 1.0
NEW
Văn bản-Âm thanh

Seed Audio 1.0

Doubao‑Audio‑Generate‑1.0 is Doubao Voice’s next‑generation audio‑generation engine. The industry‑first commercial tool creates film‑grade audio with just one prompt. It eliminates cumbersome audio‑engineering work. Creators generate publish‑ready radio dramas, podcasts and branded audio easily, shifting from a simple voice‑generator to an AI audio director. It serves audiobooks, serialized episodes and commercial audio for high‑quality narrative‑driven production.

AUDIO-GENERATION
From
$0.143/phút
Doubao Seed Character
NEW
HOT
LLM

Doubao Seed Character

Flagship model delivering premium reasoning, coding, multimodal understanding, and enterprise-grade performance.

From
$0.2/0.8M Đầu vào/Đầu ra
Doubao Seed 2.1 Turbo
NEW
HOT
LLM
TURBO

Doubao Seed 2.1 Turbo

Turbo model optimized for ultra-low latency, high throughput, and responsive AI experiences.

From
$0.45/2.25M Đầu vào/Đầu ra
Doubao Seed 2.1 Pro
NEW
HOT
LLM
PRO

Doubao Seed 2.1 Pro

Flagship model delivering premium reasoning, coding, multimodal understanding, and enterprise-grade performance.

From
$0.9/4.5M Đầu vào/Đầu ra
Seedance 2.0 Mini Reference-to-Video
NEW
Hình ảnh-Video

Seedance 2.0 Mini Reference-to-Video

Lightweight, economical multimodal video generation from reference images, videos, and audio with native audio.

From
$0.056/GIÂY
Seedance 2.0 Mini Image-to-Video
NEW
Hình ảnh-Video

Seedance 2.0 Mini Image-to-Video

Lightweight, economical video generation from a first-frame image (and optional last-frame) with native audio.

From
$0.056/GIÂY
Seedance 2.0 Mini Text-to-Video
NEW
Văn bản-Video

Seedance 2.0 Mini Text-to-Video

Lightweight, economical video generation from text prompts with native audio.

From
$0.056/GIÂY
HappyHorse-1.1 Text-to-video
NEW
Văn bản-Video

HappyHorse-1.1 Text-to-video

Generates videos from text prompts with HappyHorse 1.1, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

From
$0.14/GIÂY
HappyHorse-1.1 Image-to-video
NEW
Hình ảnh-Video

HappyHorse-1.1 Image-to-video

Animates a first-frame image into video with optional prompt guidance, 720P or 1080P output, and durations from 3 to 15 seconds.

From
$0.14/GIÂY
HappyHorse-1.1 Reference-to-video
NEW
Tham chiếu-Video

HappyHorse-1.1 Reference-to-video

Generates videos from one to nine reference images and a text prompt, supporting 720P or 1080P output, flexible aspect ratios, and durations from 3 to 15 seconds.

From
$0.14/GIÂY
GLM 5.2
NEW
HOT
LLM

GLM 5.2

Agent-oriented model built for complex reasoning, tool use, and autonomous task execution.

From$1.4/4.4M Đầu vào/Đầu ra
$1.26/3.96M Đầu vào/Đầu ra
-10%
Gemini Omni Flash Reference-to-Video
NEW
Tham chiếu-Video

Gemini Omni Flash Reference-to-Video

A natively multimodal Google DeepMind model that generates cinematic, sound-enabled videos from a text prompt plus 1-5 reference images, carrying a consistent subject, scene, or style across generations.

From
$0.135/GIÂY
Gemini Omni Flash Image-to-Video
NEW
Hình ảnh-Video

Gemini Omni Flash Image-to-Video

A natively multimodal Google DeepMind model that animates a still image into a cinematic, sound-enabled video guided by a text prompt while preserving the source subject and composition.

From
$0.13/GIÂY
Gemini Omni Flash Video Edit
NEW
Video-Video

Gemini Omni Flash Video Edit

A natively multimodal Google DeepMind model that edits an existing video from a text prompt with optional reference images, applying scene-consistent changes and native audio while preserving the untouched footage.

VIDEO-EDIT
From
$0.14/GIÂY
Gemini Omni Flash Text-to-Video
NEW
Văn bản-Video

Gemini Omni Flash Text-to-Video

A natively multimodal Google DeepMind model that generates cinematic videos with synchronized native audio from a text prompt alone, grounded in real-world physics for controllable, high-speed video generation.

From
$0.125/GIÂY
Kimi K2.7 Code
NEW
HOT
LLM

Kimi K2.7 Code

Powerful coding model for programming, debugging, and AI developer workflows.

INT4CODE
From
$0.95/4M Đầu vào/Đầu ra
Gemini Omni Flash Reference-to-Video Developer
NEW
Video-Video
DEV

Gemini Omni Flash Reference-to-Video Developer

Gemini Omni Flash is Google's multimodal video generation model. This reference-to-video variant transforms existing video clips using reference images and text prompts, enabling video style transfer, scene editing, and character insertion.

From
$0.12/GIÂY
MiniMax M3
NEW
HOT
LLM

MiniMax M3

MiniMax M3 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world capability while maintaining exceptional latency, scalability, and cost efficiency.

CODE
From$0.6/2.4M Đầu vào/Đầu ra
$0.3/1.2M Đầu vào/Đầu ra
-50%
Avatar Omni Human 1.5
NEW
HOT
Âm thanh-Video

Avatar Omni Human 1.5

OmniHuman 1.5 is ByteDance's digital-human model that turns a single portrait plus an audio track into a lifelike video of that character speaking or singing, with lip-sync, expressions, and gestures generated straight from the audio.

From
$0.12/GIÂY
Kling V3.0 Turbo Image-to-Video
NEW
Hình ảnh-Video
TURBO

Kling V3.0 Turbo Image-to-Video

Kling V3.0 Turbo Image-to-Video transforms static images into dynamic cinematic videos using MVL technology. Supports first/last frame control and audio generation.

From$0.112/GIÂY
$0.095/GIÂY
-15%
Kling V3.0 Turbo Text-to-Video
NEW
Văn bản-Video
TURBO

Kling V3.0 Turbo Text-to-Video

Kling V3.0 Turbo Text-to-Video generates dynamic cinematic videos from text prompts using MVL technology. Supports first/last frame control and audio generation.

From$0.112/GIÂY
$0.095/GIÂY
-15%