hotchpotch/nllb-sampled-500k download history

hotchpotch/nllb-sampled-500k is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 114 times (9 in the last 7 days), and 4,195 times in total. It ranks #103,765 among datasets by monthly downloads.

nllb-sampled-500k This dataset contains sampled bilingual sentence pairs derived from allenai/nllb. It is published in a convenient Hugging Face layout with one subset per language pair, for example ace_Latn-ban_Latn. The dataset is intended for multilingual representation learning, t

Open hotchpotch/nllb-sampled-500k on Hugging Face