mesolitica/AudioSet-Audio-Instructions download history
mesolitica/AudioSet-Audio-Instructions is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 354 times (53 in the last 7 days), and 3,692 times in total. It ranks #43,214 among datasets by monthly downloads.
AudioSet-Audio-Instructions Convert AudioSet up to 527 audio labels to Speech Instruction dataset. For Speech, we transcribe first using Whisper Large V3 after that use the transcription with the label to generate the synthetic instructions.
Open mesolitica/AudioSet-Audio-Instructions on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.