S-EMBER is a benchmark for streaming episodic memory over long egocentric
(wearable-camera) video. Given a video and a natural-language question about
something that happened earlier in the recording, a model must recall the
relevant moment and answer.
This repository is an anonymized mirror provided for peer review. It
contains the complete benchmark. Author, institution, and provenance
information has been intentionally omitted for double-blind review.… See the full description on the dataset page:
https://huggingface.co/datasets/paper-review-only/S-EMBER.