A model Nvidia recently fine-tuned that scored extremely high on LLM-as-a-judge evals like MT Bench and ArenaHard. The final jury will be LMSYS, but it's exciting to see fine tunes continuing to climb upwards. This was trained on the HelpSteer2 data. At the same time, we need actually good models and not just high evaluations.
Explore other models with behavioral similarity to Llama-3.1-Nemotron-70B-Instruct-HF.
Relative Adoption Metric contextualizes downloads against the model's size bucket.
OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Llama-3.1-Nemotron-70B-Instruct-HF fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub sums them into one daily total per model.
