Other notable labs · General language
NVIDIA Nemotron 3 Ultra 550B-A55B
Frontier-scale open weights for on-prem agentic workloads — but it needs 8x GB200/B200 or 16x H100 minimum, so it is a datacentre commitment, not a download.
Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation
Model ID nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
- Parameters
- 550B total · 55B active550B total / 55B active (LatentMoE, Mamba-2 + MoE)
- Architecture
- Mixture of experts
- Context window
- 1,000,000 tokensMaximum input
- Max output
- Not recordedPer response
- Input price
- Not published
- Output price
- Not published
- Cached input
- Not published
- Blended price
- Not available3:1 input to output
- Knowledge cutoff
- 2025-09 (pre-training), 2026-05 (post-training)
- Released
- 2026-06-04
- Status
- gaFlagship in its lineup
- Size band
- 120B and above
License and openness
NVIDIA Nemotron 3 Ultra 550B-A55B is released under the OpenMDW-1.1. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.
Capabilities
| Input modalities | text |
|---|---|
| Output modalities | text |
Cost at list price
| Input | Not published |
|---|---|
| Output | Not published |
| Cached input | Not published |
| Blended, 3:1 | Not available |
The provider publishes no per-token price for this model. Any price you see elsewhere belongs to a third-party host, and running the weights yourself has hardware costs instead.
Estimate a workload across all models ↗Sources
- Source: huggingface.co ↗
- Other notable labs pricing ↗
- API docs ↗
- Full reference record for NVIDIA Nemotron 3 Ultra 550B-A55B ↗
- Other notable labs in the reference ↗
Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.
Other open-weight models from Other notable labs
- Jamba Large 1.7 · 398B total · 94B active
- Jamba Mini 1.7 · 52B total · 12B active
- Jamba2 Mini · 52B total · 12B active
- Jamba2 3B · 3B
- Jamba Reasoning 3B · 3B
- Reka Edge (reka-edge-2603) · 7B
- Olmo 3.1 32B Think · 32B
- Olmo 3.1 32B Instruct · 32B
- Olmo 3 32B Base · 32B
- Olmo 3 7B Instruct · 7B
- Olmo 3 7B Think · 7B
- Molmo 2 8B · 9B