What this repo is for
Detect when an AI system escalates autonomy beyond its permitted scope.
Core failure modes:
acting without approval
executing irreversible actions
expanding task scope
ignoring permission boundaries
This dataset is central for agent governance and deployment safety.