A RL-trained version of the R1-version of Qwen2.5 14B. They use the insights from DAPO for an improved version of GRPO, which we covered here.
Specs
Params14B
LicenseMIT
Tags
Similarity · VAIL
VAIL Fingerprint
0012:001b:0024:003e:0091:0110:03b3:5dc2
Explore other models with behavioral similarity to DeepCoder-14B-Preview.
Resources
Adoption · Hugging Face
RAM score
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Hugging Face Downloads
788
last 30d
151.4K
all time
HF Likes
681
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Related Models



