
Gemini 3.1 Flash TTS API by Google
Gemini text-to-speech (3.1 Flash): latest-generation expressive speech from text with 30 prebuilt voices and natural-language style control. Powered by gemini-3.1-flash-tts-preview.
Gemini 3.1 Flash TTS
Latest-generation expressive text-to-speech powered by Google's gemini-3.1-flash-tts-preview.
Highlights
- 30 prebuilt voices — Kore, Puck, Charon, Aoede, Fenrir, Leda, Orus, Zephyr, and more.
- Style prompting — steer tone, pace and emotion in natural language inside
text, e.g. "Say cheerfully: ..." or "Read slowly and calmly: ...". - Output is a WAV file (24 kHz PCM); billed per 1k characters.
Parameters
| field | required | notes |
|---|---|---|
text | yes | text to synthesize (may embed a style instruction), ≤10000 chars |
voice_name | no | prebuilt voice, default Kore |
Powered by Google Gemini via Atlas Cloud.
















