parambharat/bengali_asr_corpus download history
parambharat/bengali_asr_corpus is an automatic speech recognition dataset on the Hugging Face Hub. In the last 30 days it was downloaded 54 times (15 in the last 7 days), and 966 times in total. It ranks #175,582 among datasets by monthly downloads.
The corpus contains roughly 500 hours of audio and transcripts in Bangla language. The transcripts have beed de-duplicated using exact match deduplication and audio has be converted to 16000 samples
Open parambharat/bengali_asr_corpus on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.