alexliap/greek-cc download history
alexliap/greek-cc is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 12,481 times (4,249 in the last 7 days), and 12,514 times in total. It ranks #2,493 among datasets by monthly downloads.
Greek Common Crawl A FineWeb-style Greek-language text dataset extracted from Common Crawl, following the FineWeb-2 recipe adapted for Greek (ell_Grek). Pipeline source: github.com/alexliap/greek-cc. Crawl coverage starts at CC-MAIN-2024-22 rather than Common Crawl's earliest snapsho
Spaces using greek-cc
- Model Pulse 8 likes