abhayesian/llama-3.3-70b-reward-model-biases-dpo-merged download history

abhayesian/llama-3.3-70b-reward-model-biases-dpo-merged is a 70.6B-parameter text generation model by abhayesian. In the last 30 days it was downloaded 23 times (7 in the last 7 days), and 3,350 times in total.

It ranks #371,996 on the Hub by monthly downloads and #108,195 among text generation models.

It has 0 likes.

2 models build on llama-3.3-70b-reward-model-biases-dpo-merged: 2 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 856 times in the last 30 days. See the llama-3.3-70b-reward-model-biases-dpo-merged galaxy.

Most downloaded derivatives

Open abhayesian/llama-3.3-70b-reward-model-biases-dpo-merged on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.