All models

SearchR1-3B-PPo

byPeterJinGoPeterJinGo· 12 Mar 2025
General purpose

A RL-trained model that learns to use search engines with multi-turn support. Those experiments are important first steps to training an (open) replication of OpenAI Deep Research.

Specs
Params3B
LicenseUnknown
Similarity · VAIL
VAIL
VAIL Fingerprint
04f9:0660:094a:0bd1:1802:2284:279d:5e64

Explore other models with behavioral similarity to SearchR1-3B-PPo.

Adoption · Hugging Face
RAM score
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Hugging Face Downloads
32
last 30d
1.7K
all time
HF Likes
0

Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.

Related Models