seyoungsong/Open-Korean-Historical-Corpus download history
seyoungsong/Open-Korean-Historical-Corpus is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 884 times (104 in the last 7 days), and 7,921 times in total. It ranks #21,201 among datasets by monthly downloads.
Open Korean Historical Corpus Dataset Description The Open Korean Historical Corpus is a large-scale, openly licensed dataset created to address the lack of accessible data for Korean NLP and historical linguistics. It contains 17.7 million documents (5.1 billion tokens) compil
Models trained on Open-Korean-Historical-Corpus
2 models list it as training data.
- LinkinShan/hanja-wsd-base 141 downloads in 30 days
- LinkinShan/hanja-wsd-large 58 downloads in 30 days
Open seyoungsong/Open-Korean-Historical-Corpus on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.