A curated corpus of multi-agent execution trajectories paired with step-level decisive-error annotations for online auditing of LLM-based multi-agent systems.
Companion code: github.com/ZBox1005/AgentForesight
Project page: zbox1005.github.io/agent-foresight
AFTraj-2K contains 1,162 verified-safe and 1,114 unsafe multi-agent trajectories (2,276 total) spanning three deployment-faithful domains. Each unsafe trajectory is annotated with a… See the full description on the dataset page:
https://huggingface.co/datasets/ZBox008003/AFTraj.