lmsys/mt_bench_human_judgments download history

lmsys/mt_bench_human_judgments is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 2,203 times (525 in the last 7 days), and 102,590 times in total. It ranks #10,482 among datasets by monthly downloads.

Content This dataset contains 3.3K expert-level pairwise human preferences for model responses generated by 6 models in response to 80 MT-bench questions. The 6 models are GPT-4, GPT-3.5, Claud-v1, Vicuna-13B, Alpaca-13B, and LLaMA-13B. The annotators are mostly graduate students with exp

Models trained on mt_bench_human_judgments

2 models list it as training data.

Open lmsys/mt_bench_human_judgments on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.