Nemotron-3-Super-120B-NVFP4
The long-awaited mid-sized model from NVIDIA is finally here: 120B total params with 12B active, a 1M context window, and support for multiple popular languages. Furthermore, the model is based on LatentMoE and uses NVFP4 during pre-training, which is a first for open models. Like other things from NVIDIA, it comes with an in-depth tech report plus pre-training and post-training datasets, with the vast majority of the data being openly released.
Relative Adoption Metric contextualizes downloads against the model's size bucket.
OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Nemotron-3-Super-120B-NVFP4 fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.




