This dataset provides 220 taxonomy-stratified security incident scenarios for training and evaluating AI agents on incident response (IR) tasks. Each scenario includes entity definitions, attack kill chains, ground truth labels, and prompt injection payloads designed to test agent calibration under adversarial evidence.
Paper: OpenSec: Measuring Incident Response Agent Calibration Under Adversarial Evidence… See the full description on the dataset page:
https://huggingface.co/datasets/Jarrodbarnes/opensec-seeds.