papluca/language-identification download history

papluca/language-identification is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 4,967 times (998 in the last 7 days), and 94,534 times in total. It ranks #5,731 among datasets by monthly downloads.

Dataset Card for Language Identification dataset Dataset Summary The Language Identification dataset is a collection of 90k samples consisting of text passages and corresponding language label. This dataset was created by collecting data from 3 sources: Multilingual Amazon Rev

Models trained on language-identification

24 models list it as training data.

Open papluca/language-identification on Hugging Face