zou-lab/BioMed-R1-Eval download history
zou-lab/BioMed-R1-Eval is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 673 times (172 in the last 7 days), and 11,942 times in total. It ranks #26,179 among datasets by monthly downloads.
Disentangling Reasoning and Knowledge in Medical Large Language Models This is the evaluation dataset accompanying our paper, comprising 11 publicly available biomedical benchmarks. We disentangle each benchmark question into either medical reasoning or medical knowledge categories. Addit
Open zou-lab/BioMed-R1-Eval on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.