nlp-chula/ThaiTrees download history

nlp-chula/ThaiTrees is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 357 times (281 in the last 7 days), and 357 times in total. It ranks #42,938 among datasets by monthly downloads.

ThaiTrees A 342M-token corpus of Thai drawn from news, Wikipedia, spoken transcripts and social media, automatically parsed under the Universal Dependencies framework. It is released as three artefacts: a raw text corpus, a frequency lexicon, and a dependency-parsed corpus in CoNLL-U.

Open nlp-chula/ThaiTrees on Hugging Face