Nemotron 3 Nano 30B A3B
View websiteReleased December 2025 under the NVIDIA Nemotron Open Model License, a 30B hybrid model activating 3.5B parameters per token across 23 Mamba-2 and MoE layers plus 6 attention layers, with 128 routed experts and 1 shared expert. The smallest of the Nemotron 3 line, sized for local and edge deployment.
At a glance
- Context window
- 262K tokens
- Knowledge cutoff
- Jun 2025
- Modalities
- Text Text → Text Text
Capabilities
Function calling
Connect to external tools, APIs, and systems.
Structured output
Return responses in structured formats like JSON.
Pricing by provider
| Provider | Input / 1M tokens | Output / 1M tokens | |
|---|---|---|---|
|
|
$0.05 | $0.20 | View |
|
|
$0.05 | $0.20 | View |
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.
Compare with other models
Estimated prices shown. Actual costs may vary based on context length, batch size, caching, and provider-specific pricing tiers.