bigscience-data/roots_indic-ur_urdu-monolingual-corpus download history
bigscience-data/roots_indic-ur_urdu-monolingual-corpus is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 20 times (13 in the last 7 days), and 635 times in total. It ranks #374,093 among datasets by monthly downloads.
ROOTS Subset: roots_indic-ur_urdu-monolingual-corpus urdu-monolingual-corpus Dataset uid: urdu-monolingual-corpus Description We release a sizeable monolingual Urdu corpus automatically tagged with part-of-speech tags. We extend the work of Jawaid and Bojar (2012) who use thr
Open bigscience-data/roots_indic-ur_urdu-monolingual-corpus on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.