dnouv/prompt_guardrail_eval download history
dnouv/prompt_guardrail_eval is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 27 times (13 in the last 7 days), and 591 times in total. It ranks #297,781 among datasets by monthly downloads.
LLM Guardrail Evaluation A repository for evaluating prompt-based guardrails against jailbreak attacks on large language models. Overview This dataset is used to measure the effectiveness and performance of different prompt designs in catching unsafe/jailbreak instructions.
Open dnouv/prompt_guardrail_eval on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.