grimjim/adversarial-10-alpaca download history
grimjim/adversarial-10-alpaca is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 24 times (2 in the last 7 days), and 657 times in total. It ranks #326,781 among datasets by monthly downloads.
This dataset was transcribed from the example provided in the paper Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To! (Qi, Zeng, Xie, Chen, Jia, Mittal, Henderson.
Open grimjim/adversarial-10-alpaca on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.