SEGAgentRL/LLDS-R-GSPO-Qwen2.5-3B-Ins download history

SEGAgentRL/LLDS-R-GSPO-Qwen2.5-3B-Ins is a 3.4B-parameter reinforcement learning model by SEGAgentRL. In the last 30 days it was downloaded 12 times (1 in the last 7 days), and 78 times in total.

It ranks #670,866 on the Hub by monthly downloads and #12,994 among reinforcement learning models.

It has 1 likes.

2 models build on LLDS-R-GSPO-Qwen2.5-3B-Ins: 2 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 877 times in the last 30 days. See the LLDS-R-GSPO-Qwen2.5-3B-Ins galaxy.

Most downloaded derivatives

Open SEGAgentRL/LLDS-R-GSPO-Qwen2.5-3B-Ins on Hugging Face