escontra/mdm_data download history
escontra/mdm_data is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 16 times (6 in the last 7 days), and 357 times in total. It ranks #434,073 among datasets by monthly downloads.
The WikiText language modeling dataset is a collection of over 100 million tokens extracted from the set of verified Good and Featured articles on Wikipedia. The dataset is available under the Creative Commons Attribution-ShareAlike License.
Open escontra/mdm_data on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.