Cornell-AGI/Ultrafeedback-Llama-3-Armo-iter_1 download history

Cornell-AGI/Ultrafeedback-Llama-3-Armo-iter_1 is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 59 times (4 in the last 7 days), and 1,123 times in total. It ranks #166,158 among datasets by monthly downloads.

Dataset Card for Ultrafeedback-Llama-3-Armo-iter_1 This dataset was used to train REBEL-Llama-3-Armo-iter_1. We generate 5 responses using Meta-Llama-3-8B-Instruct and collect the rewards with ArmoRM-Llama3-8B-v0.1. The best response in terms of reward is selected as chosen while the wors

Open Cornell-AGI/Ultrafeedback-Llama-3-Armo-iter_1 on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.