All models

Nemotron-3-Super-120B-NVFP4

byNVIDIANVIDIA· 11 Mar 2026
General purpose

The long-awaited mid-sized model from NVIDIA is finally here: 120B total params with 12B active, a 1M context window, and support for multiple popular languages. Furthermore, the model is based on LatentMoE and uses NVFP4 during pre-training, which is a first for open models. Like other things from NVIDIA, it comes with an in-depth tech report plus pre-training and post-training datasets, with the vast majority of the data being openly released.

Specs
Params120B
LicenseNVIDIA Nemotron Open Model License
Capability · Artificial Analysis
AA Index
25.7
Months Behind Frontier
12.5 mo
claude-3-7-sonnet-thinking
24 Feb 2025
Adoption · Hugging Face
RAM @ 30d
Relative Adoption Metric: 14.5×. Breakout. Measured at 30 days.
Hugging Face Downloads
1.5M
last 30d
9M
all time
HF Likes
429

Relative Adoption Metric contextualizes downloads against the model's size bucket.

Inference · OpenRouter
Tokens/Day
58.5B
7d avg · listed 7/7
Peak Tokens/Day
103.6B
09 Apr 2026
Peak Rank
#7
09 Apr 2026

OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Nemotron-3-Super-120B-NVFP4 fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.

Related Models