All models

Kimi K2 Thinking

byMoonshotMoonshot· 06 Nov 2025
General purpose

The best open model, competitive with some of the best closed models. However, independent evaluation is a problem as third-party API providers struggle to implement the model correctly, something which we have seen, for example, with the release of GPT-OSS. As an example, running the agentic "Vending-Bench" from Andon Labs with a third party provider as opposed to the official API makes a huge difference:

This is a huge problem plaguing open models. Moonshot also documents the tool calling accuracy in a repo, where a lot of providers perform sub-par, including vLLM with a schema accuracy <90%.

Specs
Params1T (32B active)
LicenseModified MIT
Capability · Artificial Analysis
AA Index
33.5
Months Behind Frontier
3.9 mo
grok-4
10 Jul 2025
Adoption · Hugging Face
RAM @ 30d
Relative Adoption Metric: 0.75×. Below benchmark. Measured at 30 days.
Hugging Face Downloads
48.9K
last 30d
1.8M
all time
HF Likes
1.7K

Relative Adoption Metric contextualizes downloads against the model's size bucket.

Inference · OpenRouter
Tokens/Day
below top 50 this week
Peak Tokens/Day
14B
29 Jan 2026
Peak Rank
#23
29 Jan 2026

OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Kimi K2 Thinking fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.

Related Models