textdetox/multilingual_toxicity_dataset download history

textdetox/multilingual_toxicity_dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,564 times (352 in the last 7 days), and 21,794 times in total. It ranks #13,725 among datasets by monthly downloads.

Multilingual Toxicity Detection Dataset [2025] We extend our binary toxicity classification dataset to more languages! Now also covered: Italian, French, Hebrew, Hindglish, Japanese, Tatar. The data is prepared for TextDetox 2025 shared task. [2024] For the shared task TextDetox 2024, we

Models trained on multilingual_toxicity_dataset

11 models list it as training data.

Open textdetox/multilingual_toxicity_dataset on Hugging Face