← All models
Speech-to-text model powered by GPT-4o Mini
Details
- Accepts
- audio
Pricing
- Input
- 156 cr / 1M tokens
- Output
- 625 cr / 1M tokens
- Audio input
- 375 cr / 1M tokens
Prices in credits (1 credit = $0.01).
Data schema
approximateInput
| Field | Type | Description |
|---|---|---|
| audiorequired | string | Input audio to transcribe. |
| response_format | stringenum: json, text, srt, verbose_json, vtt |
Output
| Field | Type | Description |
|---|---|---|
| text | string | Transcribed text. |