← All models
Generate speech with Eleven v4 Turbo from ElevenLabs. Choose a voice and control delivery with audio tags, stability, similarity settings, and IPA pronunciation.
Details
- Accepts
- text
Pricing
- Price
- 0.00500 cr / character
Prices in credits (1 credit = $0.01).
Data schema
Input
| Field | Type | Description |
|---|---|---|
| seed | — | Seed for best-effort reproducibility. Identical output is not guaranteed. |
| textrequired | string | The text to convert to speech. Supports audio tags such as [whispering] and IPA pronunciation enclosed in forward slashes. |
| voice | string | The voice to use for speech generation |
| stability | number | Voice stability. Lower values allow more expressive delivery; higher values make delivery more consistent. |
| timestamps | boolean | Whether to return character-level timing information with the generated audio. |
| language_code | — | Language code (ISO 639-1) for speech generation and text normalization. |
| output_format | stringenum: mp3_22050_32, mp3_44100_32, mp3_44100_64, mp3_44100_96, mp3_44100_128, mp3_44100_192… | Output format of the generated audio. Formatted as codec_sample_rate_bitrate. |
| similarity_boost | number | How closely the output follows the reference voice. Higher values increase similarity but may reduce naturalness. |
| apply_text_normalization | stringenum: auto, on, off | Whether to normalize text such as numbers and dates before generation. |
Output
| Field | Type | Description |
|---|---|---|
| audio | — | The generated audio file |
| timestamps | — | Timestamps for each word in the generated speech. Only returned if `timestamps` is set to True in the request. |