hassanjbara/LONG-DPO download history

hassanjbara/LONG-DPO is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 20 times (2 in the last 7 days), and 1,390 times in total. It ranks #374,093 among datasets by monthly downloads.

This dataset is a DPO version of the hassanjbara/LONG dataset. The preference model used for choosing and rejecting responses is Hello-SimpleAI/chatgpt-detector-roberta, with the idea being creating a dataset for training a model to "beat" that detector. Full script for generating the dataset inclu

Models trained on LONG-DPO

2 models list it as training data.

Open hassanjbara/LONG-DPO on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.