infery
← All models

Chatterbox

chatterbox-text-to-speech

Text to Speechby Resemble AI

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

Details

Accepts
text + audio

Pricing

Price
1.88 cr / minute

Prices in credits (1 credit = $0.01).

Data schema

Input

FieldTypeDescription
cfgnumber
seedUseful to control the reproducibility of the generated audio. Assuming all other properties didn't change, a fixed seed should always generate the exact same audio file. Set to 0 for random seed..
textstringThe text to be converted to speech (maximum 5000 characters). You can additionally add the following emotive tags: <laugh>, <chuckle>, <sigh>, <cough>, <sniffle>, <groan>, <yawn>, <gasp>
audio_urlOptional URL to an audio file to use as a reference for the generated speech. If provided, the model will try to match the style and tone of the reference audio.
temperaturenumberTemperature for generation (higher = more creative).
exaggerationnumberExaggeration factor for the generated speech (0.0 = no exaggeration, 1.0 = maximum exaggeration).
source_audio_urlstring
target_voice_audio_urlRequired URL to an audio file to use as the target reference voice for speech-to-speech voice conversion.

Output

FieldTypeDescription
audioThe generated speech audio