joelniklaus/eurlex_resources download history

joelniklaus/eurlex_resources is a fill mask dataset on the Hugging Face Hub. In the last 30 days it was downloaded 3,286 times (2,400 in the last 7 days), and 36,305 times in total. It ranks #7,801 among datasets by monthly downloads.

Dataset Card for EurlexResources: A Corpus Covering the Largest EURLEX Resources Dataset Summary This dataset contains large text resources (~179GB in total) from EURLEX that can be used for pretraining language models. Use the dataset like this: from datasets import load_datas

Models trained on eurlex_resources

32 models list it as training data.

Open joelniklaus/eurlex_resources on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.