cfilt/IITB-IndicMonoDoc download history

cfilt/IITB-IndicMonoDoc is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 10,005 times (1,075 in the last 7 days), and 423,912 times in total. It ranks #3,019 among datasets by monthly downloads.

IITB Document level Monolingual Corpora for Indian languages. 22 scheduled languages of India + English (1) Assamese, (2) Bengali, (3) Gujarati, (4) Hindi, (5) Kannada, (6) Kashmiri, (7) Konkani, (8) Malayalam, (9) Manipuri, (10) Marathi, (11) Nepali, (12) Oriya, (13) Punjabi, (14) Sanskrit, (15) S

Spaces using IITB-IndicMonoDoc

Open cfilt/IITB-IndicMonoDoc on Hugging Face