laurievb/open-lid-dataset download history

laurievb/open-lid-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 2,221 times (305 in the last 7 days), and 24,124 times in total. It ranks #10,416 among datasets by monthly downloads.

Dataset Card for "open-lid-dataset" Dataset Summary The OpenLID dataset covers 201 languages and is designed for training language identification models. The majority of the source datasets were derived from news sites, Wikipedia, or religious text, though some come from other

Models trained on open-lid-dataset

1 models list it as training data.

Open laurievb/open-lid-dataset on Hugging Face