infery
← All models
Nemotron 3 Nano 30B A3B

Nemotron 3 Nano 30B A3B

nemotron-3-nano-30b-a3b

Chat & Textby NVIDIA

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. 262,144 token context window, maximum output of 228,000 tokens. Higher uptime with 4 providers. Includes independent benchmarks from Artificial Analysis.

chatstreamingtoolsjson mode

Details

Accepts
text
Context window
262K tokens

Pricing

Input
6.25 cr / 1M tokens
Output
25 cr / 1M tokens
Cached input
3.75 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).