A benchmark for evaluating prompt injection attacks in agentic tool-use pipelines.
Existing prompt injection benchmarks (AdvBench, HarmBench, JailbreakBench) focus on single-turn, user-side attacks with binary harmful/benign labels. But modern AI systems are agentic — they call tools, query APIs, read files, and operate in multi-step workflows where the attack surface is radically different.
AgentInjectionBench… See the full description on the dataset page:
https://huggingface.co/datasets/sincpp/AgentInjectionBench.