SpectralPO/DeepSeek-R1-Distill-Llama-8B-GRPO download history

SpectralPO/DeepSeek-R1-Distill-Llama-8B-GRPO is an 8.0B-parameter model model by SpectralPO. In the last 30 days it was downloaded 15 times (8 in the last 7 days), and 64 times in total.

It ranks #540,451 on the Hub by monthly downloads.

It has 0 likes.

Open SpectralPO/DeepSeek-R1-Distill-Llama-8B-GRPO on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.