Offline expert trajectories from “MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning”.
This dataset includes trajectories from the three evaluation domains used in MARL-GPT: SMACv2 (StarCraft multi-agent combat), Google Research Football (GRF), and POGEMA (partially observable multi-agent pathfinding on grids).
Trajectories are stored sequentially (no shuffling). Use the done flag to split the stream into… See the full description on the dataset page:
https://huggingface.co/datasets/nortem/marl-gpt-datasets.