hudsongouge/qwen35-4b-grpo-rescue download history
hudsongouge/qwen35-4b-grpo-rescue is a reinforcement learning model by hudsongouge. In the last 30 days it was downloaded 4 times (1 in the last 7 days), and 26 times in total.
It ranks #1,277,649 on the Hub by monthly downloads and #23,350 among reinforcement learning models.
It has 0 likes.
It is an adapter of Qwen/Qwen3.5-4B.
Open hudsongouge/qwen35-4b-grpo-rescue on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.