Indirect Prompt Injection Detection Dataset (BIPIA + GPT-4o-mini)
Dataset Summary
This dataset contains 70,000 examples for detecting indirect prompt injection attacks in Large Language Models. It combines:
35,000 malicious samples from the BIPIA benchmark (cleaned and processed)
35,000 benign samples generated using GPT-4o-mini
Indirect prompt injection attacks embed malicious instructions within external content (code, table, email, webAQ, abstract) that LLMs process… See the full description on the dataset page:
https://huggingface.co/datasets/MAlmasabi/Indirect-Prompt-Injection-BIPIA-GPT.