nlpctx/tamil-english-corpus download history

nlpctx/tamil-english-corpus is a text retrieval dataset on the Hugging Face Hub. In the last 30 days it was downloaded 16 times (5 in the last 7 days), and 110 times in total. It ranks #433,426 among datasets by monthly downloads.

Tamil-English Retrieval Corpus A high-quality multilingual retrieval corpus constructed from the Mozhi Tamil Corpus and machine-translated into English using IndicTrans2. Dataset Summary This dataset contains Tamil documents paired with English translations. The corpus

Open nlpctx/tamil-english-corpus on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.