openbank-uz/youtube_transcriptions download history
openbank-uz/youtube_transcriptions is an automatic speech recognition dataset on the Hugging Face Hub. In the last 30 days it was downloaded 754 times (235 in the last 7 days), and 7,535 times in total. It ranks #24,003 among datasets by monthly downloads.
Dataset Description A speech dataset of Uzbek language audio clips sourced from YouTube videos. Audio segments were extracted, separated by speaker using vocal isolation, and transcribed using Google's Gemini 2.0 Flash model. Speaker identities were clustered using ECAPA-TDNN embeddings.
Open openbank-uz/youtube_transcriptions on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.