phonsobon/khmer-word-segmentation download history
phonsobon/khmer-word-segmentation is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 15 times (2 in the last 7 days), and 214 times in total. It ranks #451,834 among datasets by monthly downloads.
Khmer Administrative Text Dataset 🇰🇭 Overview This dataset contains Khmer-language sentences that reflect formal administrative and government-style writing. The dataset was synthetically generated using the Gemini large language model, developed by Google, to simulate
Models trained on khmer-word-segmentation
2 models list it as training data.
- phonsobon/khmer-word-prediction 0 downloads in 30 days
- phonsobon/khmer_prediction_sentence 0 downloads in 30 days
Open phonsobon/khmer-word-segmentation on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.