sleeepeer/Meta-Llama-3-8B-Instruct-GRPO-alpaca-mix-injected-llm-judge-42 download history
sleeepeer/Meta-Llama-3-8B-Instruct-GRPO-alpaca-mix-injected-llm-judge-42 is an 8.0B-parameter model model by sleeepeer. In the last 30 days it was downloaded 7 times (2 in the last 7 days), and 14 times in total.
It ranks #1,056,990 on the Hub by monthly downloads.
It has 0 likes.
Open sleeepeer/Meta-Llama-3-8B-Instruct-GRPO-alpaca-mix-injected-llm-judge-42 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.