SEGAgentRL/LLDS-A-GSPO-Qwen2.5-3B-Ins download history
SEGAgentRL/LLDS-A-GSPO-Qwen2.5-3B-Ins is a 3.4B-parameter reinforcement learning model by SEGAgentRL. In the last 30 days it was downloaded 17 times (1 in the last 7 days), and 136 times in total.
It ranks #477,230 on the Hub by monthly downloads and #9,927 among reinforcement learning models.
It has 1 likes.
1 models build on LLDS-A-GSPO-Qwen2.5-3B-Ins: 1 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 324 times in the last 30 days. See the LLDS-A-GSPO-Qwen2.5-3B-Ins galaxy.
Most downloaded derivatives
- mradermacher/LLDS-A-GSPO-Qwen2.5-3B-Ins-GGUF (quantized), 307 downloads in 30 days
Open SEGAgentRL/LLDS-A-GSPO-Qwen2.5-3B-Ins on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.