lmsys/mt_bench_human_judgments download history
lmsys/mt_bench_human_judgments is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 2,203 times (525 in the last 7 days), and 102,590 times in total. It ranks #10,482 among datasets by monthly downloads.
Content This dataset contains 3.3K expert-level pairwise human preferences for model responses generated by 6 models in response to 80 MT-bench questions. The 6 models are GPT-4, GPT-3.5, Claud-v1, Vicuna-13B, Alpaca-13B, and LLaMA-13B. The annotators are mostly graduate students with exp
Models trained on mt_bench_human_judgments
2 models list it as training data.
- samratduttaofficial/WaterSheep 32 downloads in 30 days
- NikolayKozloff/Mixtral_AI_CyberTron_Swahili_7b-GGUF 22 downloads in 30 days
Open lmsys/mt_bench_human_judgments on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.