nlp-chula/ThaiTrees download history
nlp-chula/ThaiTrees is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 357 times (281 in the last 7 days), and 357 times in total. It ranks #42,938 among datasets by monthly downloads.
ThaiTrees A 342M-token corpus of Thai drawn from news, Wikipedia, spoken transcripts and social media, automatically parsed under the Universal Dependencies framework. It is released as three artefacts: a raw text corpus, a frequency lexicon, and a dependency-parsed corpus in CoNLL-U.