Nemotron 3 Nano 30B A3B (free)
Archived
NVIDIA🇺🇸
nvidia/nemotron-3-nano-30b-a3b:freeNemotron 3 Nano is NVIDIA's compute-efficient open model in the Nemotron 3 family, a 30B mixture-of-experts model with 3B active parameters optimized for high-throughput, low-cost agentic and multi-agent workloads. This is the free-tier variant of the same model.
Cost rate
Free
Context
256K
Released
Dec 14, 2025
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Performance
Median latency and throughput measured across recent requests.
Throughput225 tok/s
Latency