A RL-tuned version of R1-Qwen2.5 1.5B, which improves the math performance compared to the original, distilled model significantly. One of their main findings is that the incorrect responses were three times longer than the correct ones.
Specs
Params1.5B
LicenseMIT
Tags
Similarity · VAIL
VAIL Fingerprint
0076:00b0:00f2:0120:0225:036c:0654:4758
Explore other models with behavioral similarity to DeepScaleR-1.5B-Preview.
Resources
Adoption · Hugging Face
RAM score
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Hugging Face Downloads
9.6K
last 30d
743K
all time
HF Likes
584
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Related Models



