gupta-tanish/llama3-8b-instruct-on-policy-mpo-iteration2 download history
gupta-tanish/llama3-8b-instruct-on-policy-mpo-iteration2 is an 8.0B-parameter text generation model by gupta-tanish. In the last 30 days it was downloaded 8 times (4 in the last 7 days), and 41 times in total.
It ranks #971,478 on the Hub by monthly downloads and #368,705 among text generation models.
It has 0 likes.
Open gupta-tanish/llama3-8b-instruct-on-policy-mpo-iteration2 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.