mzhaoshuai/Llama-3.3-70B-Inst-awq_SafeRLHF download history
mzhaoshuai/Llama-3.3-70B-Inst-awq_SafeRLHF is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 57 times (9 in the last 7 days), and 481 times in total. It ranks #169,558 among datasets by monthly downloads.
Llama-3.3-70B-Inst-awq Responses for RefAlign Safety Alignment This dataset contains responses generated for the paper Learning from Reference Answers: Versatile Language Model Alignment without Binary Human Preference Data, which introduces the RefAlign alignment algorithm. Code Reposito
Models trained on Llama-3.3-70B-Inst-awq_SafeRLHF
3 models list it as training data.
- mradermacher/alpaca-7b-ref-meteor-GGUF 485 downloads in 30 days
- mzhaoshuai/alpaca-7b-ref-meteor 40 downloads in 30 days
- mzhaoshuai/alpaca-7b-ref-bertscore 24 downloads in 30 days
Open mzhaoshuai/Llama-3.3-70B-Inst-awq_SafeRLHF on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.