TalTechNLP/voxlingua107_wds download history

TalTechNLP/voxlingua107_wds is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 6,952 times (4,063 in the last 7 days), and 82,392 times in total. It ranks #4,286 among datasets by monthly downloads.

VoxLingua107 VoxLingua107 is a speech dataset for training spoken language identification models. The dataset consists of short speech segments automatically extracted from YouTube videos and labeled according the language of the video title and description, with some post-processing ste

Spaces using voxlingua107_wds

Open TalTechNLP/voxlingua107_wds on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.