Bonkh/multilingual-benchmark-dataset download history
Bonkh/multilingual-benchmark-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 497 times (61 in the last 7 days), and 4,780 times in total. It ranks #33,285 among datasets by monthly downloads.
Multilingual benchmark dataset This dataset contains cleaned text data organized by language, and can be used as a benchmark for evaluating language identification (LID) models. Overview This dataset consists of samples from multiple existing benchmark and multilingual text dat