HANI-LAB/Med-REFL-DPO download history
HANI-LAB/Med-REFL-DPO is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 23 times (3 in the last 7 days), and 849 times in total. It ranks #337,693 among datasets by monthly downloads.
News [2025/06/10] We are releasing the Med-REFL dataset, which is split into two subsets: Reasoning Enhancement Data and Reflection Enhancement Data. Introduction This is the Direct Preference Optimization (DPO) dataset created by the Med-REFL framework, designed to improve the