SPRINGLab/BPCC_cleaned download history
SPRINGLab/BPCC_cleaned is a translation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 527 times (26 in the last 7 days), and 1,538 times in total. It ranks #31,827 among datasets by monthly downloads.
A curated subset of Bharat Parallel Corpus Collection (BPCC) for 8 Indian languages. Translation pairs are filtered with LABSE score(>0.9) and further preprocessed. Useful for training high-quality translation models.
Models trained on BPCC_cleaned
1 models list it as training data.
- SPRINGLab/shiksha-MT-nllb-3.3B 18 downloads in 30 days