prometheus-eval/MM-Eval download history
prometheus-eval/MM-Eval is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 79 times (32 in the last 7 days), and 2,771 times in total. It ranks #134,617 among datasets by monthly downloads.
Multilingual Meta-EVALuation benchmark (MM-Eval) 👨💻Code | 📄Paper | 🤗 MMQA MM-Eval is a multilingual meta-evaluation benchmark consisting of five core subsets—Chat, Reasoning, Safety, Language Hallucination, and Linguistics—spanning 18 languages and a Language Resource subset spanning