sleeepeer/Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42 download history
sleeepeer/Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42 is an 8.0B-parameter model model by sleeepeer. In the last 30 days it was downloaded 8 times (1 in the last 7 days), and 24 times in total.
It ranks #971,478 on the Hub by monthly downloads.
It has 0 likes.
2 models build on Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42: 0 quantized, 2 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 31 times in the last 30 days. See the Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42 galaxy.
Most downloaded derivatives
- sleeepeer/Llama-3.1-8B-Instruct-GRPO-AT-short-10-NEW-TRAIN-directly-output-easy-AT-1 (finetune), 12 downloads in 30 days
- sleeepeer/Llama-3.1-8B-Instruct-GRPO-AT-short-10-NEW-TRAIN-directly-output-easy-2-AT-1 (finetune), 11 downloads in 30 days
Open sleeepeer/Llama-3.1-8B-Instruct-GRPO-alpaca-combine-100-no-KL-42 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.