laurievb/open-lid-dataset download history
laurievb/open-lid-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 2,221 times (305 in the last 7 days), and 24,124 times in total. It ranks #10,416 among datasets by monthly downloads.
Dataset Card for "open-lid-dataset" Dataset Summary The OpenLID dataset covers 201 languages and is designed for training language identification models. The majority of the source datasets were derived from news sites, Wikipedia, or religious text, though some come from other
Models trained on open-lid-dataset
1 models list it as training data.
- davidschulte/ESM_laurievb__open-lid-dataset_default 6 downloads in 30 days