abhi26/dpo-scientific-reasoning download history
abhi26/dpo-scientific-reasoning is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 34 times (6 in the last 7 days), and 898 times in total. It ranks #248,325 among datasets by monthly downloads.
Dpo Scientific Reasoning This dataset contains 100 high-quality examples for Direct Preference Optimization (DPO) training, focused on scientific reasoning and analysis. Dataset Description This dataset was generated using an enhanced DSPy-based pipeline that creates structured
Open abhi26/dpo-scientific-reasoning on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.