This dataset accompanies the paper "[PromptShield: Deployable Detection for Prompt Injection Attacks]" (ArXiv Link) and is built from a curated selection of open-source datasets and published prompt injection attack strategies.
Task: Binary classification of prompt injection attempts.
Fields:
prompt: The full text of the prompt, including instructions, inputs… See the full description on the dataset page:
https://huggingface.co/datasets/gaurav-nimbalkar/PromptShield.