ai-safety-institute/harmful-advice-dataset download history
ai-safety-institute/harmful-advice-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 116 times (16 in the last 7 days), and 827 times in total. It ranks #102,500 among datasets by monthly downloads.
Harmful Advice Dataset Developed by: Lennart Luettgau1, Henry Davidson1, Elizabeth Nguyen2, Daria Butuc2, Christopher Summerfield1 1 UK AI Security Institute, 2 Pareto AI This dataset contains advice requests and responses with harm level annotations from multiple graders (human domain ex
Models trained on harmful-advice-dataset
1 models list it as training data.
- ai-safety-institute/Llama-3.1-8B-harmful-advice-classifier 0 downloads in 30 days
Open ai-safety-institute/harmful-advice-dataset on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.