Per-frame depth annotations for the ScanNet dataset, produced by
Depth-Anything-3 (DA3) and then aligned to each scene's sparse depth
from the original ScanNet reconstruction. We use this refined dataset for 3D world generation and reconstruction in our Puffin-World.
Each clip is a 1×3 comparison — RGB | Original Depth | Our Aligned Depth —
with depth rendered by vision banana representation. It shows
how the DA3-aligned… See the full description on the dataset page:
https://huggingface.co/datasets/KangLiao/ScanNet-Depth-DA3-Aligned.