mispeech/MECAT-Caption download history

mispeech/MECAT-Caption is an audio classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 311 times (88 in the last 7 days), and 8,069 times in total. It ranks #47,737 among datasets by monthly downloads.

MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks ๐Ÿ“– Paper | ๐Ÿ› ๏ธ GitHub | ๐ŸŽง Demo | ๐Ÿ”Š MECAT-QA (HF) Dataset Description MECAT (Multi-Expert Chain for Audio Tasks) is a comprehensive benchmark constructed on large-scale data to evaluate machine unders

Models trained on MECAT-Caption

1 models list it as training data.

Open mispeech/MECAT-Caption on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.