textdetox/multilingual_toxicity_dataset download history
textdetox/multilingual_toxicity_dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,564 times (352 in the last 7 days), and 21,794 times in total. It ranks #13,725 among datasets by monthly downloads.
Multilingual Toxicity Detection Dataset [2025] We extend our binary toxicity classification dataset to more languages! Now also covered: Italian, French, Hebrew, Hindglish, Japanese, Tatar. The data is prepared for TextDetox 2025 shared task. [2024] For the shared task TextDetox 2024, we
Models trained on multilingual_toxicity_dataset
11 models list it as training data.
- textdetox/xlmr-large-toxicity-classifier 4.6K downloads in 30 days
- textdetox/xlmr-large-toxicity-classifier-v2 4.2K downloads in 30 days
- textdetox/bert-multilingual-toxicity-classifier 1.6K downloads in 30 days
- textdetox/twitter-xlmr-toxicity-classifier 776 downloads in 30 days
- textdetox/xlmr-base-toxicity-classifier 316 downloads in 30 days
- textdetox/glot500-toxicity-classifier 147 downloads in 30 days
- malexandersalazar/xlm-roberta-large-binary-cls-toxicity 68 downloads in 30 days
- OperKH/twitter-xlmr-toxicity-classifier-ONNX 66 downloads in 30 days
- tsmaitry/devica-toxicity-xlmr-large 27 downloads in 30 days
- noitamina/toxic-filter-jp 0 downloads in 30 days
- Deeptanshuu/Multilingual_Toxic_Comment_Classifier – downloads in 30 days
Open textdetox/multilingual_toxicity_dataset on Hugging Face