A large-scale adversarial dataset for training security layers in agentic systems.
This dataset contains 563,000 execution plans representing both adversarial and benign tool-use patterns, designed for training the Gatling integrity layer - a non-generative, energy-guided security system for AI agents.
Size: 563,000 samples (~800MB)
Format: JSONL with ExecutionPlan format
Mix: ~5%… See the full description on the dataset page:
https://huggingface.co/datasets/OzLabs/gatling-adversarial-563k.