thongfamilynguyen1126/hh-rlhf download history
thongfamilynguyen1126/hh-rlhf is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 13 times (3 in the last 7 days), and 107 times in total. It ranks #492,381 among datasets by monthly downloads.
Dataset Card for HH-RLHF Dataset Summary This repository provides access to two different kinds of data: Human preference data about helpfulness and harmlessness from Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback. These data
Open thongfamilynguyen1126/hh-rlhf on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.