Nvidia: Nemotron 3 Nano 30B A3B
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
Specifications
| Attribute | Value |
|---|---|
| Lab | Nvidia |
| Tags | Fast Agentic |
| Release Date | 2025-12 |
| Context Window | 262,144 tokens |
| Input Price / 1M | $0.05 |
| Output Price / 1M | $0.20 |
| Input Modalities | Text |
| Output Modalities | Text |
Strengths
- Efficient MoE with 30B total / 3B active params
- Very low latency and cost
- Strong for agentic and tool-use tasks
- NVIDIA ecosystem integration
Weaknesses
- Limited deep reasoning capability
- Below larger Nemotron models
- Narrower general knowledge
Best For
- Latency-sensitive agent tasks
- Cost-efficient deployments
- NVIDIA-optimized inference
In Depth: Nemotron 3 Nano 30B A3B
Summary
Nemotron 3 Nano 30B A3B is an AI model from Nvidia.
Released 2025-12. It currently appears in the Overall category on LMRank. It supports Text input and produces Text output, with a context window of 262,144 tokens. Input pricing is $0.05 per 1M tokens and output is $0.20 per 1M tokens on OpenRouter.