Another strong MoE base model from the DeepSeek team. Some people are questioning the very high MMLU scores, which is a similar story to the next model (Yi 1.5). Regardless, open models are making a lot of progress on MoE models. Scaling MoE models from this 20B active range to 100+ is supposedly an almighty engineering challenge.
Similarity · VAIL
VAIL Fingerprint
00ac:00e5:0121:01be:02f6:0546:0b3a:675e
Explore other models with behavioral similarity to DeepSeek-V2.
Capability · Artificial Analysis
AA Index
3.6
Adoption · Hugging Face
RAM score
—
Hugging Face Downloads
10.2K
last 30d
855K
all time
HF Likes
334
Relative Adoption Metric contextualizes downloads against the model's size bucket.
Related Models
