billingsmoore/tibetan-word-segmentation-ds download history
billingsmoore/tibetan-word-segmentation-ds is a token classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 21 times (6 in the last 7 days), and 207 times in total. It ranks #361,848 among datasets by monthly downloads.
Tibetan Word Segmentation Annotations Human expert word boundary annotations for 100 utterances of modern Lhasa Tibetan news speech, drawn from the NICT-Tib1 ASR corpus. Dataset Description Tibetan orthography does not mark spaces between words; syllables are delimited
Open billingsmoore/tibetan-word-segmentation-ds on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.