All models

Llama-3.1-Tulu-3.1-8B

byAi2Ai2· 07 Feb 2025
General purpose

Ai2 released an updated version of Tülu. Trained on the same data as the previous version, but with GRPO (instead of PPO), the same algorithm used by R1. This results in better performance across the board, most notably in math benchmarks. It also is evidence against the argument that GRPO is "poor man's PPO". Full reasoning models from Ai2 are still "coming soon."

Specs
Params8B
LicenseLlama 3.1 Community License
Similarity · VAIL
VAIL
VAIL Fingerprint
0322:0435:0594:06d4:0c1d:102b:18e4:5fc6

Explore other models with behavioral similarity to Llama-3.1-Tulu-3.1-8B.

Adoption · Hugging Face
RAM score
Hugging Face Downloads
849
last 30d
58.9K
all time
HF Likes
40

Relative Adoption Metric contextualizes downloads against the model's size bucket.

Related Models