Per-image camera parameter annotations for the full ImageNet-1K dataset
(1,000 training classes + the 50,000-image validation split, ~1.35M images), captioned by the Puffin-World model. More captioned datasets are provided in our Puffin-16M website.
The collage above visualizes the camera maps on sample images — each
pair shows the up field (green arrows: the projected gravity-up direction)
and the latitude field (colored contours: angle above/below the… See the full description on the dataset page:
https://huggingface.co/datasets/KangLiao/ImageNet-1K-Camera.