rizkysulaeman/Qwen-3-GRPO-Healthcare-Deep-Research download history

rizkysulaeman/Qwen-3-GRPO-Healthcare-Deep-Research is a 1.5B-parameter reinforcement learning model by rizkysulaeman. In the last 30 days it was downloaded 95 times (9 in the last 7 days), and 386 times in total.

It ranks #191,614 on the Hub by monthly downloads and #685 among reinforcement learning models.

It has 0 likes.

It is a quantization of unsloth/Qwen2.5-1.5B-Instruct.

Open rizkysulaeman/Qwen-3-GRPO-Healthcare-Deep-Research on Hugging Face