rmndrnts/wikipedia_asr_splitted download history

rmndrnts/wikipedia_asr_splitted is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 346 times (35 in the last 7 days), and 3,090 times in total. It ranks #43,945 among datasets by monthly downloads.

Synthetic dataset based on highly specialised texts from Wikipedia. This version is splitted by category. Voiced using Yandex SpeechKit with random voices, roles and speech rate. Can be used to evaluate ASR models not trained on given domains and to identify areas that the model does not handle wel

Open rmndrnts/wikipedia_asr_splitted on Hugging Face