PromptEval/PromptEval_MMLU_correctness download history

PromptEval/PromptEval_MMLU_correctness is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 6,540 times (83 in the last 7 days), and 149,792 times in total. It ranks #4,522 among datasets by monthly downloads.

MMLU Multi-Prompt Evaluation Data (correctness scores) Overview This dataset contains the results of a comprehensive evaluation of various Large Language Models (LLMs) using multiple prompt templates on the Massive Multitask Language Understanding (MMLU) benchmark. The data is

Spaces using PromptEval_MMLU_correctness

Open PromptEval/PromptEval_MMLU_correctness on Hugging Face