Name: SAM (Spatial Audio-driven human Motion)
Purpose: This dataset provides paired human motion sequences and corresponding spatial audio for modeling and analysis.
Source: Captured using a Vicon motion capture system.
Size: Over 9 hours of recordings at 120 frames per second (FPS).
Added audio.txt reference file. This file maps audio IDs to their corresponding class labels. The data is segmented by the human subject who performed the… See the full description on the dataset page:
https://huggingface.co/datasets/JimSYXu/SAM.