Anthropic/hh-rlhf download history

Anthropic/hh-rlhf is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 29,535 times (6,231 in the last 7 days), and 2,049,830 times in total. It ranks #962 among datasets by monthly downloads.

Dataset Card for HH-RLHF Dataset Summary This repository provides access to two different kinds of data: Human preference data about helpfulness and harmlessness from Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback. These data

Models trained on hh-rlhf

493 models list it as training data.

Spaces using hh-rlhf

Open Anthropic/hh-rlhf on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.