slone/nllb-200-10M-sample download history

slone/nllb-200-10M-sample is a translation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 362 times (74 in the last 7 days), and 4,782 times in total. It ranks #42,486 among datasets by monthly downloads.

Dataset Card for "nllb-200-10M-sample" This is a sample of nearly 10M sentence pairs from the NLLB-200 mined dataset allenai/nllb, scored with the model facebook/blaser-2.0-qe described in the SeamlessM4T paper. The sample is not random; instead, we just took the top n sentence pairs f

Models trained on nllb-200-10M-sample

1 models list it as training data.

Open slone/nllb-200-10M-sample on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.