infery
← All models

Gemma 3 12B

gemma-3-12b-it

Chat & Textby Google

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. 131,072 token context window, maximum output of 16,384 tokens. Includes independent benchmarks from Artificial Analysis.

chatstreamingvisiontoolsjson mode

Details

Accepts
text
Context window
131K tokens

Pricing

Input
6.25 cr / 1M tokens
Output
18.75 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).