RAY-AUTRA-TECHNOLOGY/img_pointV1
img_pointV1 This dataset is a collection of 3D point clouds generated from images in the ImageNet-1k VL Enriched dataset (visual-layer/imagenet-1k-vl-enriched). Each 2D image is converted into a point cloud where the (X, Y) coordinates correspond to pixel locations, and the Z coordinate (depth/elevation) is derived from the pixel's grayscale intensity. The original image colors are retained as the colors of the points. The point clouds are stored in the GLB format.… See the full description on the dataset page: https://huggingface.co/datasets/RAY-AUTRA-TECHNOLOGY/img_pointV1.

img_pointV1
This dataset is a collection of 3D point clouds generated from images in the ImageNet-1k VL Enriched dataset (visual-layer/imagenet-1k-vl-enriched).
Each 2D image is converted into a point cloud where the (X, Y) coordinates correspond to pixel locations, and the Z coordinate (depth/elevation) is derived from the pixel's grayscale intensity. The original image colors are retained as the colors of the points. The point clouds are stored in the GLB format.
Dataset Structure
The dataset consists of two main parts:
- `dataset_cloudV1.arrow`: An Apache Arrow file containing the metadata for each point cloud.
- `glb_files/`: A directory containing the individual GLB files, each representing a 3D point cloud.
This process creates a 3D representation of the image, similar to a heightmap, where luminance dictates depth.
RAY AUTRA TECHNOLOGY 2025
