Joemn/project-madurai-corpus-tamil download history

Joemn/project-madurai-corpus-tamil is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 158 times (101 in the last 7 days), and 186 times in total. It ranks #81,504 among datasets by monthly downloads.

Project Madurai Structured Tamil Literary Corpus Structured Tamil and English literary text extracted from the Unicode HTML catalogue published by Project Madurai. The release preserves source URLs, SHA-256 hashes, catalogue metadata, Project Madurai provenance headers, section hierar

Open Joemn/project-madurai-corpus-tamil on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.