SEACrowd/ted_en_id download history

SEACrowd/ted_en_id is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 106 times (17 in the last 7 days), and 1,768 times in total. It ranks #109,173 among datasets by monthly downloads.

TED En-Id is a machine translation dataset containing Indonesian-English parallel sentences collected from the TED talk transcripts. We split the dataset and use 75% as the training set, 10% as the validation set, and 15% as the test set. Each of the datasets is evaluated in both directions, i.e., E

Open SEACrowd/ted_en_id on Hugging Face