prometheus-eval/MM-Eval download history

prometheus-eval/MM-Eval is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 79 times (32 in the last 7 days), and 2,771 times in total. It ranks #134,617 among datasets by monthly downloads.

Multilingual Meta-EVALuation benchmark (MM-Eval) 👨‍💻Code | 📄Paper | 🤗 MMQA MM-Eval is a multilingual meta-evaluation benchmark consisting of five core subsets—Chat, Reasoning, Safety, Language Hallucination, and Linguistics—spanning 18 languages and a Language Resource subset spanning

Open prometheus-eval/MM-Eval on Hugging Face