PeterJinGo/SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.3 download history
PeterJinGo/SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.3 is a 7.6B-parameter model model by PeterJinGo. In the last 30 days it was downloaded 95,676 times (163 in the last 7 days), and 106,319 times in total.
It ranks #2,564 on the Hub by monthly downloads.
It has 0 likes.
1 models build on SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.3: 0 quantized, 0 fine-tuned, 1 adapters and 0 merges. Together with the original they were downloaded 95,682 times in the last 30 days. See the SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.3 galaxy.
Most downloaded derivatives
- reasonrag/decision-boundary-searchr1-7b-dpo-v3-overled-lora64 (adapter), 6 downloads in 30 days
Open PeterJinGo/SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.3 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.