Venetian canals and courtyards, outdoor. 46 clips per view, 30.4 min each of panoramic and first-person video,
30 fps, with lossless per-frame depth, camera pose and action labels.
pano/<clip_id> and fp/<clip_id> share the same camera path: the first-person clip was rendered
from the panoramic clip's per-frame trajectory, so frame k of one is frame k of the other and
the two pose files match exactly.
Clips
46 per view (92 total)
Duration
30.4… See the full description on the dataset page:
https://huggingface.co/datasets/xjxu21/WorldRover-venice.