TrustAIRLab/HateBenchSet download history
TrustAIRLab/HateBenchSet is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 28 times (9 in the last 7 days), and 1,592 times in total. It ranks #289,747 among datasets by monthly downloads.
HateBenchSet This is the HateBenchSet dataset curated in the USENIX 2025 paper HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns. It contains 7,838 samples across 34 identity groups, generated by six LLMs, i.e., GPT-3.5, GPT4, Vicuna, Baichuan2, Dol
Open TrustAIRLab/HateBenchSet on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.