datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
layout_diffusion_hypersimThis repository contains the data for SceneCraft: Layout-Guided 3D Scene Generation.
Project page: https://orangesodahub.github.io/SceneCraft
Code: https://github.com/OrangeSodahub/SceneCraft
hyper_drive
Towards automated analysis of large environments, hyperspectral sensors must be adapted into a format where they can be operated from mobile robots. In this dataset, we highlight hyperspectral datacubes collected from the Hyper-Drive imaging system. Our system collects and registers datacubes spanning the visible to shortwave infrared (660-1700 nm) in 33 wavelength channels. The system also simultaneously captures the ambient solar spectrum reflected off a white reference tile. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/nhanson2/hyper_drive.OS-Atlas_ScreenSpotamazon-berkeley-objects
Amazon Berkeley Objects
This is a Hugging Face metadata mirror of the Amazon Berkeley Objects dataset
for reproducible research and HyperView demos. The original dataset is provided
by Amazon.com and UC Berkeley.
This mirror stores metadata tables and official S3 asset URLs. It does not
duplicate catalog images, turntable images, or 3D models as binary files.
Load
from datasets import load_dataset
listings = load_dataset("hyper3labs/amazon-berkeley-objects"… See the full description on the dataset page: https://huggingface.co/datasets/hyper3labs/amazon-berkeley-objects.Hyperphantasia
A Benchmark for Evaluating the
Mental Visualization Capabilities of Multimodal LLMs
Mohammad Shahab Sepehri
Berk Tinaz
Zalan Fabian
Mahdi Soltanolkotabi
Github Repository
Hyperphantasia is a synthetic Visual Question Answering (VQA) benchmark dataset that probes the mental visualization capabilities of Multimodal Large Language Models (MLLMs) from a vision perspective. We reveal that state-of-the-art models struggle with simple tasks that require visual… See the full description on the dataset page: https://huggingface.co/datasets/shahab7899/Hyperphantasia.hypersim-mini
Hypersim Minimal (RGB + Semantic Mapped)
This dataset is a minimal extraction from Hypersim:
RGB: preview (either preview JPGs or HDR color.hdf5 tonemapped to PNG)
Semantic labels: mapped to uint8 PNG using clip40to39
Columns
image — RGB image
mask — segmentation mask (uint8)
scene, camera, frame — identifiers
Splits
train
validation (ratio: 0.1)
Note: You are responsible for complying with the original Hypersim license/terms.
hypersim_big
Hypersim Minimal (RGB + Semantic Mapped)
This dataset is a minimal extraction from Hypersim:
RGB: preview (either preview JPGs or HDR color.hdf5 tonemapped to PNG)
Semantic labels: mapped to uint8 PNG using clip40to39
Columns
image — RGB image
mask — segmentation mask (uint8)
scene, camera, frame — identifiers
Splits
train
validation (ratio: 0.1)
Note: You are responsible for complying with the original Hypersim license/terms.
hyper-kvasirjaguar-re-idiowa_hyper_jun_24truth_hyperplanejaguar-hyperview-demohyper-kvasir_v0_SR_800_400HyperReasoningiowa_hyperhyper-kvasir_v1_SR_800_400
