harshavardhan88858/deepseek-qwen-grpo-reasoning-v1 download history
harshavardhan88858/deepseek-qwen-grpo-reasoning-v1 is a 7.6B-parameter text generation model by harshavardhan88858. In the last 30 days it was downloaded 18 times (8 in the last 7 days), and 689 times in total.
It ranks #453,072 on the Hub by monthly downloads and #143,285 among text generation models.
It has 0 likes.
It is a fine-tune of unsloth/DeepSeek-R1-Distill-Qwen-7B-bnb-4bit.
Open harshavardhan88858/deepseek-qwen-grpo-reasoning-v1 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.