infery
← All models

Llama 3.2 3B Instruct

llama-3.2-3b-instruct

Chat & Textby Meta

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. 131,072 token context window, maximum output of 131,072 tokens. Higher uptime with 2 providers. Includes independent benchmarks from Artificial Analysis.

chatstreaming

Details

Accepts
text
Context window
80K tokens

Pricing

Input
7.5 cr / 1M tokens
Output
7.5 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).