hsbharadwaj/multilingual_toxicity_dataset download history

hsbharadwaj/multilingual_toxicity_dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 294 times (91 in the last 7 days), and 357 times in total. It ranks #49,911 among datasets by monthly downloads.

Multilingual Toxicity Detection Dataset [2025] We extend our binary toxicity classification dataset to more languages! Now also covered: Italian, French, Hebrew, Hindglish, Japanese, Tatar. The data is prepared for TextDetox 2025 shared task. [2024] For the shared task TextDetox 2024,

Open hsbharadwaj/multilingual_toxicity_dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.