Other notable labs · General language
NVIDIA Nemotron 3 Super 120B-A12B
The practical sweet spot of the Nemotron line: 12B active params keeps throughput high for high-volume ticket automation and RAG while retaining a 256K default window.
Open weights · Community or custom license · Data checked September 6, 2026 · Not a hands-on evaluation
Model ID nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16
- Parameters
- 120B total · 12B active120B total / 12B active (LatentMoE, Mamba-2 + MoE)
- Architecture
- Mixture of experts
- Context window
- 1,000,000 tokensMaximum input
- Max output
- Not recordedPer response
- Input price
- Not published
- Output price
- Not published
- Cached input
- Not published
- Blended price
- Not available3:1 input to output
- Knowledge cutoff
- 2025-06 (pre-training), 2026-02 (post-training)
- Released
- 2026-03-11
- Status
- ga
- Size band
- 40B to 120B
License and openness
NVIDIA Nemotron 3 Super 120B-A12B is released under the NVIDIA Nemotron Open Model License. This is a community or custom license with its own conditions, which can include acceptable-use rules, user thresholds or naming requirements. Read it before commercial use. License terms can change between versions; confirm the text published with the weights you download.
Capabilities
| Input modalities | text |
|---|---|
| Output modalities | text |
Cost at list price
| Input | Not published |
|---|---|
| Output | Not published |
| Cached input | Not published |
| Blended, 3:1 | Not available |
The provider publishes no per-token price for this model. Any price you see elsewhere belongs to a third-party host, and running the weights yourself has hardware costs instead.
Estimate a workload across all models ↗Sources
- Source: huggingface.co ↗
- Other notable labs pricing ↗
- API docs ↗
- Full reference record for NVIDIA Nemotron 3 Super 120B-A12B ↗
- Other notable labs in the reference ↗
Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.
Other open-weight models from Other notable labs
- Jamba Large 1.7 · 398B total · 94B active
- Jamba Mini 1.7 · 52B total · 12B active
- Jamba2 Mini · 52B total · 12B active
- Jamba2 3B · 3B
- Jamba Reasoning 3B · 3B
- Reka Edge (reka-edge-2603) · 7B
- Olmo 3.1 32B Think · 32B
- Olmo 3.1 32B Instruct · 32B
- Olmo 3 32B Base · 32B
- Olmo 3 7B Instruct · 7B
- Olmo 3 7B Think · 7B
- Molmo 2 8B · 9B