mispeech/MECAT-Caption download history
mispeech/MECAT-Caption is an audio classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 311 times (88 in the last 7 days), and 8,069 times in total. It ranks #47,737 among datasets by monthly downloads.
MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks ๐ Paper | ๐ ๏ธ GitHub | ๐ง Demo | ๐ MECAT-QA (HF) Dataset Description MECAT (Multi-Expert Chain for Audio Tasks) is a comprehensive benchmark constructed on large-scale data to evaluate machine unders
Models trained on MECAT-Caption
1 models list it as training data.
- ModelsLab/midashenglm-gen-wer-lora 10 downloads in 30 days
Open mispeech/MECAT-Caption on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.