SEGAgentRL/LLDS-A-GRPO-Qwen2.5-7B-Ins download history

SEGAgentRL/LLDS-A-GRPO-Qwen2.5-7B-Ins is a 7.6B-parameter reinforcement learning model by SEGAgentRL. In the last 30 days it was downloaded 27 times (5 in the last 7 days), and 164 times in total.

It ranks #334,611 on the Hub by monthly downloads and #5,288 among reinforcement learning models.

It has 2 likes.

It is a fine-tune of Qwen/Qwen2.5-7B-Instruct.

2 models build on LLDS-A-GRPO-Qwen2.5-7B-Ins: 2 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 1,169 times in the last 30 days. See the LLDS-A-GRPO-Qwen2.5-7B-Ins galaxy.

Most downloaded derivatives

Open SEGAgentRL/LLDS-A-GRPO-Qwen2.5-7B-Ins on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.