keshan/large-sinhala-asr-dataset download history

keshan/large-sinhala-asr-dataset is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 45 times (8 in the last 7 days), and 1,396 times in total. It ranks #201,034 among datasets by monthly downloads.

This data set contains ~185K transcribed audio data for Sinhala. The data set consists of wave files, and a TSV file. The file utt_spk_text.tsv contains a FileID, anonymized UserID and the transcription of audio in the file. The data set has been manually quality checked, but there might still be er

Open keshan/large-sinhala-asr-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.