infery
← All models

Gemma 3 4B

gemma-3-4b-it

Chat & Textby Google

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. 131,072 token context window, maximum output of 16,384 tokens. Includes independent benchmarks from Artificial Analysis.

chatstreamingvisionjson mode

Details

Accepts
text
Context window
131K tokens

Pricing

Input
6.25 cr / 1M tokens
Output
12.5 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).