Action-free video pre-training data for AWM. Egocentric go-to navigation clips,
open-vocabulary route+landmark / object-level instructions (Gemini v3 harvest),
across indoor (object-level) and outdoor (route+landmark) domains. All video is
normalized to 30 fps (native 60fps down-sampled, native 30fps kept; 24/25fps
clips dropped — no clean resample). fps_map.json records each clip's ORIGINAL
fps for reference.… See the full description on the dataset page:
https://huggingface.co/datasets/Yangyihui/awm-nav-pretrain.