SPRINGLab/BPCC_cleaned download history

SPRINGLab/BPCC_cleaned is a translation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 527 times (26 in the last 7 days), and 1,538 times in total. It ranks #31,827 among datasets by monthly downloads.

A curated subset of Bharat Parallel Corpus Collection (BPCC) for 8 Indian languages. Translation pairs are filtered with LABSE score(>0.9) and further preprocessed. Useful for training high-quality translation models.

Models trained on BPCC_cleaned

1 models list it as training data.

Open SPRINGLab/BPCC_cleaned on Hugging Face