TMLR-Group-HF/Co-rewarding-I-Llama-3.2-3B-Instruct-DAPO14k download history

TMLR-Group-HF/Co-rewarding-I-Llama-3.2-3B-Instruct-DAPO14k is a 3.6B-parameter model model by TMLR-Group-HF. In the last 30 days it was downloaded 21 times (13 in the last 7 days), and 38 times in total.

It ranks #397,481 on the Hub by monthly downloads.

It has 0 likes.

1 models build on Co-rewarding-I-Llama-3.2-3B-Instruct-DAPO14k: 1 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 568 times in the last 30 days. See the Co-rewarding-I-Llama-3.2-3B-Instruct-DAPO14k galaxy.

Most downloaded derivatives

Open TMLR-Group-HF/Co-rewarding-I-Llama-3.2-3B-Instruct-DAPO14k on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.