minhleduc/multilang-classify-dataset-02 download history
minhleduc/multilang-classify-dataset-02 is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 30 times (3 in the last 7 days), and 467 times in total. It ranks #273,959 among datasets by monthly downloads.
Dataset Card for Multilingual Language Detection Dataset Details This dataset is a comprehensive resource for multilingual text classification, specifically designed for language identification. It contains over 100,000 text samples from 36 different languages, sourced from var
Models trained on multilang-classify-dataset-02
1 models list it as training data.
- minhleduc/xlm-roberta-multilang-finetuned-00 7 downloads in 30 days
Open minhleduc/multilang-classify-dataset-02 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.