vigneshwar234/llm-eval-benchmark download history
vigneshwar234/llm-eval-benchmark is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 62 times (12 in the last 7 days), and 363 times in total. It ranks #159,286 among datasets by monthly downloads.
LLM Evaluation Benchmark A 1,200-sample curated benchmark dataset for evaluating LLMs on factual accuracy and truthfulness. Sourced from MMLU and TruthfulQA, cleaned and formatted for the LLM Evaluation Framework. Dataset Summary Split Samples Us
Open vigneshwar234/llm-eval-benchmark on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.