← All models
Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.
Details
- Accepts
- text
Pricing
- Price
- 0.00313 cr / character
Prices in credits (1 credit = $0.01).
Data schema
Input
| Field | Type | Description |
|---|---|---|
| seed | — | Random seed for reproducible results. Set to 0 for random generation, or provide a specific number for consistent outputs. |
| textrequired | string | The text to be converted to speech (maximum 300 characters). Supports 23 languages including English, French, German, Spanish, Italian, Portuguese, Hindi, Arabic, Chinese, Japanese, Korean, and more. |
| voice | string | Language code for synthesis. In case using custom please provide audio url and select custom_audio_language. |
| cfg_scale | number | Configuration/pace weight controlling generation guidance (0.0-1.0). Use 0.0 for language transfer to mitigate accent inheritance. |
| temperature | number | Controls randomness and variation in generation (0.05-2.0). Higher values create more varied speech patterns. |
| exaggeration | number | Controls speech expressiveness and emotional intensity (0.25-2.0). 0.5 is neutral, higher values increase expressiveness. Extreme values may be unstable. |
| custom_audio_language | — | If using a custom audio URL, specify the language of the audio here. Ignored if voice is not a custom url. |
Output
| Field | Type | Description |
|---|---|---|
| audio | — | The generated multilingual speech audio file |