This dataset is a collection of 3D point clouds generated from images in the ImageNet-1k VL Enriched dataset (visual-layer/imagenet-1k-vl-enriched).
Each 2D image is converted into a point cloud where the (X, Y) coordinates correspond to pixel locations, and the Z coordinate (depth/elevation) is derived from the pixel's grayscale intensity. The original image colors are retained as the colors of the points. The point clouds are stored in the GLB format.… See the full description on the dataset page:
https://huggingface.co/datasets/RAY-AUTRA-TECHNOLOGY/img_pointV1.