This dataset contains 100 episodes of humans performing the pick-and-place task.
The data captured consists of multiple modalities. These include rgb-color footage, depth footage and gyro-accelerometer data captured using a head mounted camera-IMU unit,
as well as full-body motion capture data recorded by an inertial-based MoCap suit.