gupta-tanish/llama3-8b-instruct-on-policy-mpo-iteration2-v3 download history
gupta-tanish/llama3-8b-instruct-on-policy-mpo-iteration2-v3 is an 8.0B-parameter text generation model by gupta-tanish. In the last 30 days it was downloaded 9 times (5 in the last 7 days), and 104 times in total.
It ranks #891,661 on the Hub by monthly downloads and #344,818 among text generation models.
It has 0 likes.
Open gupta-tanish/llama3-8b-instruct-on-policy-mpo-iteration2-v3 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.