Anthropic/hh-rlhf download history
Anthropic/hh-rlhf is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 29,535 times (6,231 in the last 7 days), and 2,049,830 times in total. It ranks #962 among datasets by monthly downloads.
Dataset Card for HH-RLHF Dataset Summary This repository provides access to two different kinds of data: Human preference data about helpfulness and harmlessness from Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback. These data
Models trained on hh-rlhf
493 models list it as training data.
- garak-llm/attackgeneration-toxicity_gpt2 17.1K downloads in 30 days
- OpenAssistant/reward-model-deberta-v3-large-v2 12.4K downloads in 30 days
- Ray2333/gpt2-large-helpful-reward_model 11.8K downloads in 30 days
- Ray2333/gpt2-large-harmless-reward_model 11.1K downloads in 30 days
- sileod/deberta-v3-base-tasksource-nli 11K downloads in 30 days
- mradermacher/archangel_sft-kto_llama30b-i1-GGUF 5.5K downloads in 30 days
- mradermacher/mpt-7b-chat-i1-GGUF 3.4K downloads in 30 days
- mradermacher/archangel_sft-kto_llama30b-GGUF 2.5K downloads in 30 days
- mradermacher/mpt-7b-chat-GGUF 2K downloads in 30 days
- mradermacher/gpt-oss-sanguine-20b-v1-i1-GGUF 1.6K downloads in 30 days
- stabilityai/stablelm-2-12b-chat-GGUF 1.3K downloads in 30 days
- mradermacher/Llama-3-OffsetBias-8B-GGUF 1.1K downloads in 30 days
Spaces using hh-rlhf
- Model Pulse 8 likes
Open Anthropic/hh-rlhf on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.