Digital-Divide-Data/khmer-speech-dataset download history
Digital-Divide-Data/khmer-speech-dataset is an automatic speech recognition dataset on the Hugging Face Hub. In the last 30 days it was downloaded 4,511 times (913 in the last 7 days), and 33,301 times in total. It ranks #6,191 among datasets by monthly downloads.
Khmer ASR Cultural Dataset 727.94 hours of manually curated speech-text pairs by native speakers in the Khmer language about Cambodian cultural topics. On average, each recording is 8 seconds. Speaker metadata (gender, age group, and origin city) is provided. Language: Khmer (khm). S
Open Digital-Divide-Data/khmer-speech-dataset on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.