toloka/CrowdSpeech download history
toloka/CrowdSpeech is a summarization dataset on the Hugging Face Hub. In the last 30 days it was downloaded 49 times (12 in the last 7 days), and 3,457 times in total. It ranks #188,270 among datasets by monthly downloads.
CrowdSpeech is a publicly available large-scale dataset of crowdsourced audio transcriptions. It contains annotations for more than 50 hours of English speech transcriptions from more than 1,000 crowd workers.
Models trained on CrowdSpeech
1 models list it as training data.
- toloka/t5-large-for-text-aggregation 32 downloads in 30 days
Open toloka/CrowdSpeech on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.