TalTechNLP/voxlingua107_wds download history
TalTechNLP/voxlingua107_wds is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 6,952 times (4,063 in the last 7 days), and 82,392 times in total. It ranks #4,286 among datasets by monthly downloads.
VoxLingua107 VoxLingua107 is a speech dataset for training spoken language identification models. The dataset consists of short speech segments automatically extracted from YouTube videos and labeled according the language of the video title and description, with some post-processing ste
Spaces using voxlingua107_wds
- Model Pulse 8 likes
Open TalTechNLP/voxlingua107_wds on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.