← All models
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. 1,048,576 token context window, maximum output of 65,536 tokens. Higher uptime with 2 providers.
chatstreamingvisionpdftoolsjson mode
Details
- Accepts
- text
- Context window
- 1M tokens
Pricing
- Input
- 32.5 cr / 1M tokens
- Output
- 195 cr / 1M tokens
- Cached input
- 3.25 cr / 1M tokens
Prices in credits (1 credit = $0.01).
Data schema
approximateInput
| Field | Type | Description |
|---|---|---|
| promptrequired | string | The text prompt. |
Output
| Field | Type | Description |
|---|---|---|
| text | string | Generated text (choices[0].message.content). |