mzhaoshuai/Llama-3.3-70B-Inst-awq_ultrafeedback_1in3 download history

mzhaoshuai/Llama-3.3-70B-Inst-awq_ultrafeedback_1in3 is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 28 times (6 in the last 7 days), and 544 times in total. It ranks #289,747 among datasets by monthly downloads.

Generated Reference Answers for Language Model Alignment This dataset contains responses generated for the research presented in the paper Learning from Reference Answers: Versatile Language Model Alignment without Binary Human Preference Data. The paper introduces RefAlign, a versatile R

Models trained on Llama-3.3-70B-Inst-awq_ultrafeedback_1in3

3 models list it as training data.

Open mzhaoshuai/Llama-3.3-70B-Inst-awq_ultrafeedback_1in3 on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.