Models/NVIDIA/Nemotron 3 Nano 30B A3B

Nemotron 3 Nano 30B A3B

NVIDIA · Open weights · Released Dec 14, 2025 · nvidia/nemotron-3-nano-30b-a3b
Best $0.05 in / $0.20 out via DeepInfra · fp4
CompareEstimate cost
262K contextToolsReasoningStructured outputPrompt cachingOpen weights
Available from 3 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 8 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 8 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$0.05 / $0.20
per MTok in / out
CACHE READ $0.025
Recent changesRSS ↗
Oct 8, 2026endpoint removed · crusoe4 → 3
Aug 24, 2026endpoint added · nebius3 → 4
About

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

Specifications
Context window262,144
Max output235,929
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedDec 14, 2025
TokenizerOther
Open weights
Downloads907.1K
Likes825
Licenseother
safetensorspytorch
Embed badge
Nemotron 3 Nano 30B A3B price and speed badge
[![Nemotron 3 Nano 30B A3B](https://modelindex.ai/badge/nvidia/nemotron-3-nano-30b-a3b)](https://modelindex.ai/models/nvidia/nemotron-3-nano-30b-a3b)
Sources
OpenRouter API23 minutes ago
AGGREGATOR
Hugging Face23 minutes ago
COMMUNITY
ModelIndex benchmarks8 days ago
MEASURED
Verified 23 minutes ago · list prices, not negotiated rates.