nkandpa2/cccc_all_domains download history

nkandpa2/cccc_all_domains is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 11 times (7 in the last 7 days), and 2,341 times in total. It ranks #540,266 among datasets by monthly downloads.

🔓 Dolma 🍇 Creative Commons Common Crawl 🕸️ Subset of the Common Crawl corpus containing English documents with Creative Commons licenses. Snapshot Unicode Words Documents CC-MAIN-2013-20 3,851,018,197 5,529,294 CC-MAIN-2013-48 4,544,197,252 6,997,831 CC-MAIN-2014-10 4,429,2

Open nkandpa2/cccc_all_domains on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.