joelniklaus/eurlex_resources download history
joelniklaus/eurlex_resources is a fill mask dataset on the Hugging Face Hub. In the last 30 days it was downloaded 3,286 times (2,400 in the last 7 days), and 36,305 times in total. It ranks #7,801 among datasets by monthly downloads.
Dataset Card for EurlexResources: A Corpus Covering the Largest EURLEX Resources Dataset Summary This dataset contains large text resources (~179GB in total) from EURLEX that can be used for pretraining language models. Use the dataset like this: from datasets import load_datas
Models trained on eurlex_resources
32 models list it as training data.
- BSC-LT/salamandra-7b-instruct 26.4K downloads in 30 days
- BSC-LT/salamandra-2b-instruct 3.2K downloads in 30 days
- BSC-LT/ALIA-40b 1.8K downloads in 30 days
- BSC-LT/salamandra-7b 1.7K downloads in 30 days
- BSC-LT/salamandra-2b 1.6K downloads in 30 days
- mradermacher/ALIA-40b-i1-GGUF 1.2K downloads in 30 days
- mradermacher/salamandra-2b-instruct-i1-GGUF 879 downloads in 30 days
- mradermacher/salamandra-2b-instruct-GGUF 662 downloads in 30 days
- mradermacher/ALIA-40b-GGUF 492 downloads in 30 days
- tensorblock/ALIA-40b-GGUF 269 downloads in 30 days
- tensorblock/salamandra-2b-instruct-GGUF 222 downloads in 30 days
- tensorblock/salamandra-7b-instruct-GGUF 189 downloads in 30 days
Open joelniklaus/eurlex_resources on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.