notbdq/Qwen2.5-14B-Instruct-1M-GRPO-Reasoning download history

notbdq/Qwen2.5-14B-Instruct-1M-GRPO-Reasoning is a 14.8B-parameter text generation model by notbdq. In the last 30 days it was downloaded 19 times (4 in the last 7 days), and 801 times in total.

It ranks #432,033 on the Hub by monthly downloads and #133,723 among text generation models.

It has 4 likes.

11 models build on Qwen2.5-14B-Instruct-1M-GRPO-Reasoning: 7 quantized, 0 fine-tuned, 0 adapters and 4 merges. Together with the original they were downloaded 2,752 times in the last 30 days. See the Qwen2.5-14B-Instruct-1M-GRPO-Reasoning galaxy.

Most downloaded derivatives

Open notbdq/Qwen2.5-14B-Instruct-1M-GRPO-Reasoning on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.