All models

DeepCoder-14B-Preview

byAgenticaAgentica· 07 Apr 2025
Special purpose

A RL-trained version of the R1-version of Qwen2.5 14B. They use the insights from DAPO for an improved version of GRPO, which we covered here.

Specs
Params14B
LicenseMIT
Similarity · VAIL
VAIL
VAIL Fingerprint
0012:001b:0024:003e:0091:0110:03b3:5dc2

Explore other models with behavioral similarity to DeepCoder-14B-Preview.

Adoption · Hugging Face
RAM score
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Hugging Face Downloads
788
last 30d
151.4K
all time
HF Likes
681

Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.

Related Models