Seedance 2.0 Mini & Fast API dengan harga terendah di dunia — diskon hingga 68% dari harga resmi
GPT Image 2.5 Flare Edit
gambar-ke-gambar

GPT Image 2.5 Flare Edit

GPT Image 2.5 Flare Edit applies natural-language instructions to up to 16 reference images, with an optional mask, arbitrary resolutions up to 3840x2160, five quality tiers including xhigh and max, and transparent-background output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

GPT Image 2.5 Flare Edit
Seedance 2.5 Image-to-Video
Seedream v5.0 Pro Edit
Nano Banana 2 Reference to Image
Kling Video O3 4K Image-to-Video
KATEGORI
Diskon (183)
Fungsi model
Seri
14 dari 14 model
Baru
MiniMax Music 3.0
NEW
teks-ke-audio

MiniMax Music 3.0

MiniMax Music 3.0 is MiniMax's 11.1B-parameter open-weights music model that turns a musical description and optional lyrics into a complete, fully arranged and mixed song of up to five minutes - vocals, instrumentation and production included - in a single generation, with section-tag control over the arrangement and vocal or instrumental output.

From
$0.15/kali
MiniMax Lyrics Generation
NEW
teks-ke-audio

MiniMax Lyrics Generation

MiniMax Lyrics Generation is a dedicated lyric-writing model that turns a one-line theme into a complete, professionally structured set of song lyrics - title, style tags, and sections marked with [Verse]/[Chorus] structure tags - and can also edit, continue, or restructure existing lyrics, with output directly usable as the lyrics input of MiniMax's music models.

From
$0.01/kali
Seed Audio 1.0
NEW
teks-ke-audio

Seed Audio 1.0

Doubao‑Audio‑Generate‑1.0 is Doubao Voice’s next‑generation audio‑generation engine. The industry‑first commercial tool creates film‑grade audio with just one prompt. It eliminates cumbersome audio‑engineering work. Creators generate publish‑ready radio dramas, podcasts and branded audio easily, shifting from a simple voice‑generator to an AI audio director. It serves audiobooks, serialized episodes and commercial audio for high‑quality narrative‑driven production.

AUDIO-GENERATION
From
$0.143/menit
MiniMax Speech 2.6 Turbo
NEW
teks-ke-audio
TURBO

MiniMax Speech 2.6 Turbo

MiniMax text-to-speech (Turbo): fast, low-latency speech from text with selectable preset voices, speed, volume and pitch.

From$0.06/K chars
$0.048/K chars
-20%
MiniMax Speech 2.6 HD
NEW
teks-ke-audio

MiniMax Speech 2.6 HD

MiniMax text-to-speech (HD): natural, high-fidelity speech from text with selectable preset voices, speed, volume and pitch.

From$0.1/K chars
$0.08/K chars
-20%
xAI TTS v1
NEW
teks-ke-audio

xAI TTS v1

xAI TTS v1 is a high-fidelity text-to-speech model that converts text into natural, expressive speech with sub-second latency, supporting 20 languages and 80+ voices with fine-grained delivery control.

From
$0.015/K chars
MiniMax Music 2.6
teks-ke-audio

MiniMax Music 2.6

MiniMax text-to-music (latest): generate a full vocal or instrumental song from a style prompt plus lyrics with [Verse]/[Chorus] structure tags. Synchronous single-call generation.

From
$0.15/kali
ElevenLabs v3 Text-to-Speech
NEW
teks-ke-audio

ElevenLabs v3 Text-to-Speech

ElevenLabs v3 Text-to-Speech model. High-quality speech generation from text prompts.

From
$0.1/K chars
Suno chirp-v6
NEW
teks-ke-audio

Suno chirp-v6

Suno V6 text-to-music via APIMart: inspiration mode (custom=false) turns a description into a song; custom mode (custom=true) uses your own lyrics, with variety and target duration controls. Async; returns 2 tracks per generation.

From
$0.132/kali
Suno chirp-v6-wild
NEW
teks-ke-audio

Suno chirp-v6-wild

Suno V6 Wild via APIMart: the experimental V6 variant with bolder, less conventional arrangements. Same modes and pricing as chirp-v6. Async; returns 2 tracks per generation.

From
$0.132/kali
Suno chirp-v6-mini
NEW
teks-ke-audio

Suno chirp-v6-mini

Suno V6 Mini via APIMart: the lighter, faster V6 variant for drafts and short clips. Same modes and pricing as chirp-v6. Async; returns 2 tracks per generation.

From
$0.132/kali
Gemini 3.1 Flash TTS
NEW
teks-ke-audio

Gemini 3.1 Flash TTS

Gemini text-to-speech (3.1 Flash): latest-generation expressive speech from text with 30 prebuilt voices and natural-language style control. Powered by gemini-3.1-flash-tts-preview.

From
$0.15/K chars
Gemini 2.5 Flash TTS
NEW
teks-ke-audio

Gemini 2.5 Flash TTS

Gemini text-to-speech: expressive, controllable speech from text with 30+ prebuilt voices (Kore, Puck, Charon, Aoede...). Powered by gemini-2.5-flash-preview-tts.

From
$0.04/K chars
Gemini 2.5 Pro TTS
NEW
teks-ke-audio
PRO

Gemini 2.5 Pro TTS

Gemini text-to-speech (Pro): studio-quality, expressive speech from text with 30 prebuilt voices and natural-language style control. Powered by gemini-2.5-pro-preview-tts.

From
$0.08/K chars