Did the bystander move toward (approach) or away from (retreat) the camera-wearer?
One layer of the Social-Affective Filter (SAF) — dehydrated social-signal metadata extracted from egocentric (first-person) video so robots can learn to read human reactions. No raw pixels and no audio. Each row is one source video, keyed by video_id; rehydrate against your own legally-obtained Ego4D copies (below).
Rows: 862 — videos in the evaluation… See the full description on the dataset page:
https://huggingface.co/datasets/louisye/social-robotics-proxemic-kinematics.