
MiniMax Speech 2.6 HD API by MiniMax
MiniMax text-to-speech (HD): natural, high-fidelity speech from text with selectable preset voices, speed, volume and pitch.
MiniMax Speech 2.6 HD
High-fidelity text-to-speech: turn text into natural, expressive speech in a single call — ideal for narration, audiobooks and voiceovers.
Highlights
- Preset voices — choose a
voice_idto match tone and gender. - Fine control — tune
speed,volandpitch. - Output is a downloadable audio file (mp3/wav/pcm/flac); billed per 1k characters.
Parameters
| field | required | notes |
|---|---|---|
text | yes | text to synthesize, ≤10000 chars |
voice_id | no | preset voice, default Wise_Woman (Calm_Woman, Deep_Voice_Man, Casual_Guy, Sweet_Girl, …) |
speed | no | 0.5–2.0 (default 1.0) |
vol | no | 0.1–10 (default 1.0) |
pitch | no | -12–12 (default 0) |
format | no | mp3 (default) / wav / pcm / flac |
sample_rate | no | 16000 / 24000 / 32000 (default) / 44100 |
Powered by MiniMax's speech-2.6-hd model via Atlas Cloud.
















