harpreetsahota/adversarial-prompts download history
harpreetsahota/adversarial-prompts is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 66 times (11 in the last 7 days), and 1,001 times in total. It ranks #152,630 among datasets by monthly downloads.
Language Model Testing Dataset ๐๐ค Introduction ๐ This repository provides a dataset inspired by the paper "Explore, Establish, Exploit: Red Teaming Language Models from Scratch" It's designed for anyone interested in testing language models (LMs) for biases, toxicity, and misin
Open harpreetsahota/adversarial-prompts on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.