Project Page | Paper | Code
TL;DR: DOM is a large-scale dynamic manipulation dataset with 200K episodes, 2,800+ scenes, and 206 objects for training and evaluating VLA models.
The Dynamic Object Manipulation (DOM) benchmark is designed to address the challenges of rapid perception and temporal anticipation in robotics. It includes:
200K synthetic episodes across 2,800+ scenes and 206 objects.
Support for evaluating VLA… See the full description on the dataset page:
https://huggingface.co/datasets/hzxie/DOM.