yammdd/vietnamese-error-correction-corpus download history

yammdd/vietnamese-error-correction-corpus is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 66 times (16 in the last 7 days), and 506 times in total. It ranks #152,630 among datasets by monthly downloads.

Data Summary The model is trained on a Vietnamese text error correction dataset constructed from real-world noisy inputs. The dataset contains approximately 70,000 sentence pairs and is split into training, validation, and test sets. • Data Source: Crawled Vietnamese social media comm

Models trained on vietnamese-error-correction-corpus

1 models list it as training data.

Open yammdd/vietnamese-error-correction-corpus on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.