vnahata/SpokenWikipedia-retrieval download history

vnahata/SpokenWikipedia-retrieval is an automatic speech recognition dataset on the Hugging Face Hub. In the last 30 days it was downloaded 83 times (7 in the last 7 days), and 126 times in total. It ranks #129,953 among datasets by monthly downloads.

Spoken Wikipedia speech-text retrieval (MTEB) Volunteer readings of Wikipedia articles paired with the article lead, in Dutch, English, German, Spanish and French. Recordings come from Wikimedia Commons, which is free by site policy, and the lead text from each Wikipedia, which is CC-

Open vnahata/SpokenWikipedia-retrieval on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.