SLoonker/RL-OpenCodeReasoning-DPO download history

SLoonker/RL-OpenCodeReasoning-DPO is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 25 times (8 in the last 7 days), and 274 times in total. It ranks #316,899 among datasets by monthly downloads.

RL-OpenCodeReasoning-DPO Alpaca-format dataset. Columns: instruction, input, output, rejected from datasets import load_dataset ds = load_dataset("SLoonker/RL-OpenCodeReasoning-DPO", split="train") Made from (nvidia/OpenCodeReasoning)[https://huggingface.co/datasets/nvidia/OpenCodeReason

Open SLoonker/RL-OpenCodeReasoning-DPO on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.