This is the RoboTwin 2.0 fine-tuning dataset used to train the X-WAM unified 4D World Action Model. It packages dual-arm bimanual manipulation demonstrations into a unified multi-view RGB-D video + low-dimensional state/action format, where each episode provides synchronized RGB videos, depth videos, dual-arm end-effector proprioception, actions, and a… See the full description on the dataset page:
https://huggingface.co/datasets/sharinka0715/X-WAM-RoboTwin.