dslighfdsl/Llama-3.1-8B-Instruct-Env-SFT-GRPO download history
dslighfdsl/Llama-3.1-8B-Instruct-Env-SFT-GRPO is a 1.2B-parameter model model by dslighfdsl. In the last 30 days it was downloaded 11 times (5 in the last 7 days), and 26 times in total.
It ranks #739,005 on the Hub by monthly downloads.
It has 0 likes.
Open dslighfdsl/Llama-3.1-8B-Instruct-Env-SFT-GRPO on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.