Per-image camera parameter annotations for the ADE20K dataset
(a scene-parsing benchmark; 27,574 images = 25,574 train + 2,000 val), captioned by the Puffin-World model. More captioned datasets are provided in our Puffin-16M website.
The collage above visualizes the camera maps on sample images — each
pair shows the up field (green arrows: the projected gravity-up direction)
and the latitude field (colored contours: angle above/below the horizon).… See the full description on the dataset page:
https://huggingface.co/datasets/KangLiao/ADE20K-Camera.