All models

NVIDIA Nemotron 3 Ultra 550B

byNVIDIANVIDIA· 04 Jun 2026
General purpose

The big version of the Nemotron series, which uses LatentMoE to be even faster than comparable models. Just like the other Nemotron models, the vast majority of the data is open source. And, to top it all off: NVIDIA commits to using the OpenMDW license, which is tailored specifically for model weights (and data) and drops its custom license. While MIT and Apache are in the same spirit as OpenMDW, only the latter really covers model weights, while the former are software licenses that do not really apply to model weights.

Specs
Params550B
LicenseOpenMDW-1.1
Capability · Artificial Analysis
AA Index
38.3
Months Behind Frontier
6.5 mo
gemini-3-pro
18 Nov 2025
Adoption · Hugging Face
RAM @ 30d
Relative Adoption Metric: 0.30×. Below benchmark. Measured at 30 days.
Hugging Face Downloads
428.7K
last 30d
839.1K
all time
HF Likes
329

Relative Adoption Metric contextualizes downloads against the model's size bucket.

Inference · OpenRouter
Tokens/Day
631B
7d avg · listed 7/7
Peak Tokens/Day
737.1B
22 Aug 2026
Peak Rank
#4
12 Jul 2026

OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean NVIDIA Nemotron 3 Ultra 550B fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.

Related Models