Nemotron 3 Nano 30B A3B
NVIDIA · Open weights · Released Dec 14, 2025 · nvidia/nemotron-3-nano-30b-a3b
262K contextToolsReasoningStructured outputPrompt cachingOpen weights
Available from 3 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.05 | $0.20 | $0.025 | 262K | fp4 | 0.83s | 441 tok/s | 100.0% | |
| $0.05 | $0.20 | — | 262K | fp4 | 0.88s | 282 tok/s | 100.0% | |
| $0.05 | $0.20 | $0.03 | 262K | fp8 | 0.43s | 221 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 44 hours ago
Pricing variants · USD / MTok in / out
STANDARD
$0.05 / $0.20
per MTok in / out
CACHE READ $0.025
About
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
Specifications
Context window262,144
Max output262,144
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedDec 14, 2025
TokenizerOther
WeightsHugging Face ↗
Open weights
safetensorspytorch
Related
Embed badge
[](https://modelindex.ai/models/nvidia/nemotron-3-nano-30b-a3b)Sources
OpenRouter API4 minutes ago
AGGREGATORHugging Face4 minutes ago
COMMUNITYModelIndex benchmarks44 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.