abhi26/openpipe-dpo-scientific-reasoning download history

abhi26/openpipe-dpo-scientific-reasoning is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 31 times (7 in the last 7 days), and 472 times in total. It ranks #266,938 among datasets by monthly downloads.

Openpipe Dpo Scientific Reasoning This dataset contains 100 high-quality examples for Direct Preference Optimization (DPO) training, formatted for OpenPipe fine-tuning, focused on scientific reasoning and analysis. Dataset Description This dataset was generated using an enhance

Open abhi26/openpipe-dpo-scientific-reasoning on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.