infery
← All models

Gemini 2.5 Flash-Lite

gemini-2.5-flash-lite

Chat & Textby Google

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. 1,048,576 token context window, maximum output of 65,535 tokens. Higher uptime with 2 providers. Includes independent benchmarks from Artificial Analysis.

chatstreamingvisionpdftoolsjson modeimages

Details

Accepts
text
Context window
1M tokens

Pricing

Input
13 cr / 1M tokens
Output
52 cr / 1M tokens
Cached input
1.3 cr / 1M tokens
Audio input
39 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).