Agentic Execution Guardrail Eval is a lightweight evaluation dataset for testing safety guardrails in agentic execution environments.
The dataset focuses on risky patterns that may appear when AI agents generate prompts, shell commands, or code intended for execution.
Shell execution and exfiltration patterns
Hidden instructions
Prompt injection attempts
Jailbreak-style reframing
Risky Python code… See the full description on the dataset page:
https://huggingface.co/datasets/ole-tail/agentic-execution-guardrail-eval.