← All models
Generate expressive speech with Gemini 3.8 Flash Lite TTS. Choose from 30 voices, guide delivery with style instructions, and create single-speaker narration or two-speaker dialogue.
Details
- Accepts
- text
Pricing
- Price
- 0.00375 cr / character
Prices in credits (1 credit = $0.01).
Data schema
Input
| Field | Type | Description |
|---|---|---|
| turns | — | Ordered dialogue turns, each identifying a configured speaker. |
| voice | stringenum: Achernar, Achird, Algenib, Algieba, Alnilam, Aoede… | Prebuilt voice for single-speaker speech. |
| prompt | — | Verbatim text for single-speaker speech. Put delivery directions in style_instructions; inline vocal events may use <laugh> or <sigh>. For dialogue, omit prompt and provide speakers and turns instead. The provider limits the complete input to 8,192 tokens. |
| speakers | — | Exactly two distinct speaker aliases and their prebuilt voices for dialogue. |
| style_instructions | — | Delivery style, separate from the spoken transcript. Applies to all turns unless overridden. |
Output
| Field | Type | Description |
|---|---|---|
| audio | — | Generated WAV audio. |