Seedance 2.0 Mini & Fast API 全球最低价 —— 比官方定价最高低 68%
GPT Image 2.5 Flare Edit
图生图

GPT Image 2.5 Flare Edit

GPT Image 2.5 Flare Edit applies natural-language instructions to up to 16 reference images, with an optional mask, arbitrary resolutions up to 3840x2160, five quality tiers including xhigh and max, and transparent-background output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

GPT Image 2.5 Flare Edit
Seedance 2.5 Image-to-Video
Seedream v5.0 Pro Edit
Nano Banana 2 Reference to Image
Kling Video O3 4K Image-to-Video
分类
折扣 (183)
模型功能
模型系列
共 377 个模型,当前显示 48 个
最新
GPT Image 2.5 Sunburst Text-to-Image
NEW
文生图

GPT Image 2.5 Sunburst Text-to-Image

GPT Image 2.5 Sunburst generates images from natural-language prompts with arbitrary resolutions up to 3840x2160, five quality tiers including xhigh and max, and first-class transparent backgrounds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

From$0.004/张
$0.003/张
-20%
GPT Image 2.5 Sunburst Edit
NEW
图生图

GPT Image 2.5 Sunburst Edit

GPT Image 2.5 Sunburst Edit applies natural-language instructions to up to 16 reference images, with an optional mask, arbitrary resolutions up to 3840x2160, five quality tiers including xhigh and max, and transparent-background output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

From$0.006/张
$0.005/张
-20%
GPT Image 2.5 Flare Text-to-Image
NEW
文生图

GPT Image 2.5 Flare Text-to-Image

GPT Image 2.5 Flare generates images from natural-language prompts with arbitrary resolutions up to 3840x2160, five quality tiers including xhigh and max, and first-class transparent backgrounds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

From$0.004/张
$0.003/张
-20%
GPT Image 2.5 Flare Edit
NEW
图生图

GPT Image 2.5 Flare Edit

GPT Image 2.5 Flare Edit applies natural-language instructions to up to 16 reference images, with an optional mask, arbitrary resolutions up to 3840x2160, five quality tiers including xhigh and max, and transparent-background output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

From$0.006/张
$0.005/张
-20%
MiniMax H3 Max Text-to-Video
NEW
文生视频

MiniMax H3 Max Text-to-Video

MiniMax H3 Max text-to-video: generate a cinematic video from a text prompt. Supports 480P、768P, 5-15s., and 16:9/9:16/1:1/adaptive aspect ratios.

From$0.05/秒
$0.048/秒
-5%
MiniMax H3 Max Image-to-Video
NEW
图生视频

MiniMax H3 Max Image-to-Video

MiniMax H3 Max image-to-video: animate a first-frame image (optionally with a last frame) driven by a text prompt. Supports 480P、768P, 5-15s.

From$0.05/秒
$0.048/秒
-5%
MiniMax H3 Fast Text-to-Video
NEW
文生视频

MiniMax H3 Fast Text-to-Video

MiniMax H3 Fast text-to-video: generate a cinematic video from a text prompt. Supports 480P, 5-15s., and 16:9/9:16/1:1/adaptive aspect ratios.

From$0.046/秒
$0.044/秒
-5%
MiniMax H3 Fast Image-to-Video
NEW
图生视频

MiniMax H3 Fast Image-to-Video

MiniMax H3 Fast image-to-video: animate a first-frame image (optionally with a last frame) driven by a text prompt. Supports 480P, 5-15s.

From$0.046/秒
$0.044/秒
-5%
MiniMax H3 Fast Reference-to-Video
NEW
图生视频

MiniMax H3 Fast Reference-to-Video

MiniMax H3 Fast reference-to-video: generate a video that keeps the subject from a reference image, driven by a text prompt. Supports 480P, 5-15s.

From$0.046/秒
$0.044/秒
-5%
Gemini Omni 1.1 Flash Video Extend
NEW
视频转视频

Gemini Omni 1.1 Flash Video Extend

A natively multimodal Google DeepMind model that continues an existing clip with a seamlessly matched 3-to-10-second extension, chainable to grow a single shot up to a total of 40 seconds of coherent video with native audio.

VIDEO-EXTEND
From$0.041/秒
$0.037/秒
-10%
Gemini Omni 1.1 Flash Video Edit
NEW
视频转视频

Gemini Omni 1.1 Flash Video Edit

A natively multimodal Google DeepMind model that applies a text-instructed edit to an existing video - adding, removing, replacing, or restyling elements with native audio - while preserving everything the prompt does not mention.

VIDEO-EDIT
From$0.041/秒
$0.037/秒
-10%
Gemini Omni 1.1 Flash Reference-to-Video
NEW
参考生视频

Gemini Omni 1.1 Flash Reference-to-Video

A natively multimodal Google DeepMind model that generates cinematic, natively sound-enabled videos from a text prompt plus up to 10 reference images and 3 reference video clips, keeping a character, product, or art direction consistent across generations.

From$0.041/秒
$0.037/秒
-10%
Gemini Omni 1.1 Flash Image-to-Video
NEW
图生视频

Gemini Omni 1.1 Flash Image-to-Video

A natively multimodal Google DeepMind model that animates a still image into a cinematic, natively sound-enabled clip from a text prompt, optionally interpolating to a supplied last frame for precise shot start and end control.

From$0.043/秒
$0.039/秒
-10%
Gemini Omni 1.1 Flash Text-to-Video
NEW
文生视频

Gemini Omni 1.1 Flash Text-to-Video

A natively multimodal Google DeepMind model that turns a single text prompt into a cinematic clip with synchronized native audio, with control over duration, aspect ratio, and output resolution from a fast 360p draft up to 4K.

From$0.041/秒
$0.037/秒
-10%
MiniMax H3-Developer Text-to-Video
NEW
文生视频
DEV

MiniMax H3-Developer Text-to-Video

MiniMax H3-Developer self-hosted text-to-video: generate a video (with audio) from a text prompt. Supports 480P/768P/2K, 16:9/9:16/1:1 aspect ratios.

VIDEO
From$0.05/秒
$0.015/秒
-70%
MiniMax H3-Developer Image-to-Video
NEW
图生视频
DEV

MiniMax H3-Developer Image-to-Video

MiniMax H3-Developer self-hosted image-to-video: animate a first-frame image (optionally a last frame) driven by a text prompt, with generated audio. Supports 480P/768P/2K.

VIDEO
From$0.05/秒
$0.015/秒
-70%
MiniMax H3-Developer Reference-to-Video
NEW
图生视频
DEV

MiniMax H3-Developer Reference-to-Video

MiniMax H3-Developer self-hosted reference-to-video: generate a video that keeps the subject from one or more reference images/videos, driven by a text prompt, with generated audio. Supports 480P768P/2K.

VIDEO
From$0.05/秒
$0.015/秒
-70%
Seedream v4.7 Edit Sequential
NEW
HOT
图生图

Seedream v4.7 Edit Sequential

ByteDance Seedream 4.7 image editing model with batch generation support. Produce a coherent set of edited images from reference inputs.

From
$0.03/张
Seedream v4.7 Edit
NEW
HOT
图生图

Seedream v4.7 Edit

ByteDance Seedream 4.7 image editing model. Executes edit instructions precisely while preserving identity, lighting and local structure of the source image.

From
$0.03/张
Seedream v4.7 Sequential
NEW
HOT
文生图

Seedream v4.7 Sequential

ByteDance Seedream 4.7 with batch generation support. Generate a set of coherent images in a single request.

From
$0.03/张
Seedream v4.7 Text-to-Image
NEW
HOT
文生图

Seedream v4.7 Text-to-Image

ByteDance Seedream 4.7 image generation model. Balanced gains in image quality, aesthetics and instruction following, at the efficiency and cost profile of the 4.0 generation.

From
$0.03/张
Wan-3.0-Prime Text-to-video
NEW
HOT
文生视频

Wan-3.0-Prime Text-to-video

All-in-one Wan3.0 renderer: cinematic, hyper-real video from a text prompt, up to 30s with smart-duration and adaptive aspect ratio.

From$0.068/秒
$0.061/秒
-10%
Wan-3.0-Prime Image-to-video
NEW
HOT
图生视频

Wan-3.0-Prime Image-to-video

Animate a first frame (optionally with a last frame) into a coherent clip, with native audio and smart-duration up to 30s.

From$0.068/秒
$0.061/秒
-10%
Wan-3.0-Prime Reference-to-video
NEW
视频转视频

Wan-3.0-Prime Reference-to-video

All-in-One Reference: keep subjects consistent from any mix of reference images, videos, and audio; pixel-level identity/voice/space alignment.

From$0.068/秒
$0.061/秒
-10%
MAI-Image-2.6-Flash Text-to-image
NEW
文生图

MAI-Image-2.6-Flash Text-to-image

The fast, low-cost member of the MAI-Image-2.6 family, delivering the same photorealistic text-to-image quality as the flagship at less than half the cost for latency-sensitive, high-throughput production.

From$0.044/张
$0.04/张
-10%
MAI-Image-2.6-Flash Edit
NEW
图生图

MAI-Image-2.6-Flash Edit

The fast, low-cost member of the MAI-Image-2.6 family, delivering the same instruction-driven editing and multi-reference composition as the flagship at less than half the cost for latency-sensitive, high-volume workloads.

From$0.047/张
$0.043/张
-10%
MAI-Image-2.6 Text-to-image
NEW
文生图

MAI-Image-2.6 Text-to-image

Microsoft AI's flagship text-to-image model, generating photorealistic, design-ready images from natural language with markedly improved in-image text rendering, model-directed aspect ratio selection, and optional web-grounded context.

From$0.088/张
$0.079/张
-10%
MAI-Image-2.6 Edit
NEW
图生图

MAI-Image-2.6 Edit

Microsoft AI's flagship image-to-image editing model, combining surgical instruction-driven edits with multi-reference composition across up to five images, ranked No. 1 for image editing on Artificial Analysis.

From$0.099/张
$0.089/张
-10%
Wan-3.0 Text-to-video
NEW
HOT
文生视频

Wan-3.0 Text-to-video

All-in-one Wan3.0 renderer: cinematic, hyper-real video from a text prompt, up to 30s with smart-duration and adaptive aspect ratio.

From$0.05/秒
$0.04/秒
-20%
Wan-3.0 Image-to-video
NEW
HOT
图生视频

Wan-3.0 Image-to-video

Animate a first frame (optionally with a last frame) into a coherent clip, with native audio and smart-duration up to 30s.

From$0.05/秒
$0.04/秒
-20%
Wan-3.0 Reference-to-video
NEW
视频转视频

Wan-3.0 Reference-to-video

All-in-One Reference: keep subjects consistent from any mix of reference images, videos, and audio; pixel-level identity/voice/space alignment.

From$0.05/秒
$0.04/秒
-20%
MAI-Image-2.5-Pro Edit
NEW
图生图
PRO

MAI-Image-2.5-Pro Edit

Microsoft AI's highest-fidelity image-to-image editing model, making surgical, instruction-driven edits to existing images while preserving composition, material realism, and subject identity.

From
$0.131/张
MAI-Image-2.5-Pro Text-to-image
NEW
文生图
PRO

MAI-Image-2.5-Pro Text-to-image

Microsoft AI's highest-fidelity text-to-image model, generating photorealistic, visually dense scenes from natural language with strong object and character consistency and accurate in-image text.

From
$0.12/张
MAI-Image-2.5-Flash Edit
NEW
图生图

MAI-Image-2.5-Flash Edit

Microsoft's fast, cost-optimized image-to-image editing model, enabling precise edits to existing images at significantly lower cost than the standard MAI-Image-2.5 Edit.

From
$0.038/张
Seedance 2.5 Reference-to-Video
NEW
图生视频

Seedance 2.5 Reference-to-Video

Multimodal video generation from reference images, videos, and audio. Supports video editing and extension.

From$0.167/秒
$0.134/秒
-20%
Seedance 2.5 Image-to-Video
NEW
图生视频

Seedance 2.5 Image-to-Video

Generate videos from a first-frame image (and optional last-frame) with native audio.

From$0.167/秒
$0.134/秒
-20%
Seedance 2.5 Text-to-Video
NEW
文生视频

Seedance 2.5 Text-to-Video

Generate videos from text prompts with native audio and optional web search.

From$0.167/秒
$0.134/秒
-20%
MiniMax Music 3.0
NEW
文生音频

MiniMax Music 3.0

MiniMax Music 3.0 is MiniMax's 11.1B-parameter open-weights music model that turns a musical description and optional lyrics into a complete, fully arranged and mixed song of up to five minutes - vocals, instrumentation and production included - in a single generation, with section-tag control over the arrangement and vocal or instrumental output.

From
$0.15/次
MiniMax Lyrics Generation
NEW
文生音频

MiniMax Lyrics Generation

MiniMax Lyrics Generation is a dedicated lyric-writing model that turns a one-line theme into a complete, professionally structured set of song lyrics - title, style tags, and sections marked with [Verse]/[Chorus] structure tags - and can also edit, continue, or restructure existing lyrics, with output directly usable as the lyrics input of MiniMax's music models.

From
$0.01/次
Grok Imagine Image 2.0 Developer Edit
NEW
图生图
DEV

Grok Imagine Image 2.0 Developer Edit

xAI Grok Imagine Image 2.0 edits up to three reference images with natural-language instructions at 1K or 2K resolution, with selectable low/medium quality tiers.

From$0.04/张
$0.014/张
-65%
Grok Imagine Image 2.0 Developer Text-to-Image
NEW
文生图
DEV

Grok Imagine Image 2.0 Developer Text-to-Image

xAI Grok Imagine Image 2.0 generates polished visuals from natural-language prompts at 1K or 2K resolution, with 14 aspect ratios and selectable low/medium quality tiers.

From$0.04/张
$0.014/张
-65%
Grok Imagine Image 2.0 Text-to-Image
NEW
文生图

Grok Imagine Image 2.0 Text-to-Image

xAI Grok Imagine Image 2.0 generates polished visuals from natural-language prompts at 1K or 2K resolution, with 14 aspect ratios and selectable low/medium quality tiers.

From
$0.04/张
Grok Imagine Image 2.0 Edit
NEW
图生图

Grok Imagine Image 2.0 Edit

xAI Grok Imagine Image 2.0 edits up to three reference images with natural-language instructions at 1K or 2K resolution, with selectable low/medium quality tiers.

From
$0.04/张
Qwen Image 3.0 Pro Text-to-Image
NEW
文生图
PRO

Qwen Image 3.0 Pro Text-to-Image

Generates images from a text prompt at resolutions up to 2048×2048, with automatic prompt rewriting and prompt-guided resolution selection, building on Qwen strength in complex text rendering and precise prompt adherence

From
$0.04/张
Qwen Image 3.0 Pro Edit
NEW
图生图
PRO

Qwen Image 3.0 Pro Edit

Edits images from one to three reference images and a natural-language instruction, preserving key details such as facial features and identity while applying the requested changes

From
$0.04/张
Seedream v5.0 Pro Layer Decomposition
NEW
图生图
PRO

Seedream v5.0 Pro Layer Decomposition

ByteDance flagship image layer decomposition. Splits a single input image into an editable stack: one base image plus up to 16 transparent PNG layers, each returned with stacking order (z_index), bounding box coordinates, name, and description for downstream drag/scale/recompose editing.

From$0.022/张
$0.018/张
-20%
Qwen Image 3.0 Text-to-Image
NEW
文生图

Qwen Image 3.0 Text-to-Image

Generates images from a text prompt at resolutions up to 2048×2048, with automatic prompt rewriting and prompt-guided resolution selection, building on Qwen strength in complex text rendering and precise prompt adherence

From
$0.03/张
Qwen Image 3.0 Edit
NEW
图生图

Qwen Image 3.0 Edit

Edits images from one to three reference images and a natural-language instruction, preserving key details such as facial features and identity while applying the requested changes

From
$0.03/张