Nemotron 3 Super
NVIDIA · Open weights · Released Mar 11, 2026 · nvidia/nemotron-3-super-120b-a12b
Best $0.08 in / $0.40 out via DekaLLM · fp8
262K contextToolsReasoningStructured outputOpen weights
Available from 2 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 24 HOURS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| DEDekaLLM | $0.08 | $0.45 | — | 262K | fp8 | 2.1s | 26 tok/s | 99.7% |
| $0.085 | $0.40 | — | 262K | bf16 | ✕ | ✕ | 100.0% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 24 hours ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$0.08 / $0.45
per MTok in / out
FREE
Free / Free
per MTok in / out
Recent changesRSS ↗
Sep 12, 2026endpoint added · dekallm1 → 2
Sep 9, 2026endpoint removed · digitalocean2 → 1
Sep 9, 2026context changed1000000 → 262144
About
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Specifications
Context window262,144
Max output16,384
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedMar 11, 2026
TokenizerOther
WeightsHugging Face ↗
Open weights
safetensorspytorch
Related
Embed badge
[](https://modelindex.ai/models/nvidia/nemotron-3-super-120b-a12b)Sources
OpenRouter API23 minutes ago
AGGREGATORHugging Face23 minutes ago
COMMUNITYModelIndex benchmarks12 days ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.