Nvidia: Nemotron 3 Nano 30B A3B

by Nvidia Fast Agentic

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

Choose a model to compare against Nemotron 3 Nano 30B A3B

Specifications

Specifications for Nemotron 3 Nano 30B A3B
AttributeValue
Lab Nvidia
Tags Fast Agentic
Release Date 2025-12
Context Window 262,144 tokens
Input Price / 1M $0.05
Output Price / 1M $0.20
Input Modalities Text
Output Modalities Text

Strengths

  • Efficient MoE with 30B total / 3B active params
  • Very low latency and cost
  • Strong for agentic and tool-use tasks
  • NVIDIA ecosystem integration

Weaknesses

  • Limited deep reasoning capability
  • Below larger Nemotron models
  • Narrower general knowledge

Best For

  • Latency-sensitive agent tasks
  • Cost-efficient deployments
  • NVIDIA-optimized inference

In Depth: Nemotron 3 Nano 30B A3B

Summary

Nemotron 3 Nano 30B A3B is an AI model from Nvidia.

Released 2025-12. It currently appears in the Overall category on LMRank. It supports Text input and produces Text output, with a context window of 262,144 tokens. Input pricing is $0.05 per 1M tokens and output is $0.20 per 1M tokens on OpenRouter.

Sources & Further Reading

Related Models