This is a large dataset of 10M video frames and actions collected from the Assault atari environment (Bellemare et al., 2012) in order to train world models.The dataset enables reproducible, large-scale experiments in action-conditioned video prediction. It is meant to be used with Jasmine, our JAX-based world modeling codebase.
Environment: Atari Learning Environment
Frames: 10 million
Resolution: 84 × 84
Format:… See the full description on the dataset page:
https://huggingface.co/datasets/p-doom/atari-assault-dataset.