CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Ever2after /3d-spatial-reasoning-2image10K<n<100K0 likes13k downloads5mo agoHugging Face02EPFL-CVLAB-SPACECRAFT /PocketQubeimage0 likes11k downloads7mo agoHugging Face03ZeroOneCreative /amara-spatial-10k AmaraSpatial-10K A Semantically Anchored, Metric-Scale 3D Dataset for Embodied AI and Spatial Computing 10,071 AI-generated 3D meshes across 10 top-level categories and 476 subcategories — from basilisks to bassoons, cottages to cosmic stations — curated by Zero One Creative to close the spatial alignment gap that makes most generative 3D repositories unusable for zero-shot deployment in game engines, robotics simulators, and AR/VR pipelines. Every asset is… See the full description on the dataset page: https://huggingface.co/datasets/ZeroOneCreative/amara-spatial-10k.imagetext-to-3d10K<n<100K11 likes11k downloads5mo agoHugging Face04EasonXiao-888 /SpatialEdit-500K SpatialEdit-500K SpatialEdit-500K is a synthetic training dataset for fine-grained image spatial editing. It is built for learning geometry-aware edits such as object moving, object rotation, and camera viewpoint change. The dataset was introduced in the paper SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing. It is generated with a controllable rendering pipeline to provide structured spatial transformations at scale. Project Resources GitHub Repository:… See the full description on the dataset page: https://huggingface.co/datasets/EasonXiao-888/SpatialEdit-500K.imageimage-to-image100K<n<1M14 likes11k downloads6mo agoHugging Face05deinal /spacecast-data Vlasiator Dataset for Machine Learning Studies The data is stored in Zarr. It can be downloaded to a local data directory with: from huggingface_hub import snapshot_download snapshot_download( repo_id="deinal/spacecast-data", repo_type="dataset", local_dir="data" ) This will yield a local data folder that can be used with spacecast: data/ ├── graph/ - Directory containing graphs for training ├── run_1.zarr/ - Vlasiator run 1 with ρ = 0.5 cm⁻³… See the full description on the dataset page: https://huggingface.co/datasets/deinal/spacecast-data.imagen<1K0 likes11k downloads10mo agoHugging Face06Spawning /pd12m-fullThis dataset is the downloaded variant of Spawning/PD12M. More specifically, this dataset is compatible with webdataset. It was made public after obtaining permission from the original authors of the dataset. You can use the following to explore the dataset with webdataset: import webdataset as wds dataset_path = "pipe:curl -s -f -L https://huggingface.co/datasets/sayakpaul/pd12m-full/resolve/main/{00155..02480}.tar" dataset = ( wds.WebDataset(dataset_path… See the full description on the dataset page: https://huggingface.co/datasets/Spawning/pd12m-full.image10M<n<100M21 likes10k downloads2y agoHugging Face07spatialverse /SAGE-3D_Collision_Mesh SAGE-3D Collision Mesh: Physics-Enabled Collision Bodies for 3D Gaussian Scenes Paper | Project Page | Code High-precision collision geometry dataset extracted from 1,000 indoor Mesh scenes, enabling physically accurate navigation and interaction in virtual environments. Collision Mesh of InteriorGS data captured on Issac Sim 5.0. 📢 News 2025-12-15: Released SAGE-3D Collision Mesh dataset with collision bodies for 1000 InteriorGS scenes.… See the full description on the dataset page: https://huggingface.co/datasets/spatialverse/SAGE-3D_Collision_Mesh.imageroboticsn<1K104 likes8.2k downloads2mo agoHugging Face08spatial-reason /qwen_trajectories_finalimage1K<n<10K0 likes7.8k downloads6mo agoHugging Face09mila-ai4h /mid-space MID-Space: Aligning Diverse Communities’ Needs to Inclusive Public Spaces A new version of the dataset will be released soon, incorporating user identity markers and expanded annotations. LIVS PAPER Click below to see more: Overview The MID-Space dataset is designed to align AI-generated visualizations of urban public spaces with the preferences of diverse and marginalized communities in Montreal. It includes textual prompts, Stable Diffusion… See the full description on the dataset page: https://huggingface.co/datasets/mila-ai4h/mid-space.imagetext-to-image1K<n<10K1 likes7.3k downloads2y agoHugging Face10EPFL-CVLAB-SPACECRAFT /SwissCubeimage10K<n<100K3 likes6.2k downloads2y agoHugging Face11RiceD2KLab /SWiM-SpacecraftWithMasks SWiM: Spacecraft With Masks A large-scale instance segmentation dataset of nearly 64k annotated spacecraft images created using real spacecraft models, superimposed on a mixture of real and synthetic backgrounds generated using NASA's TTALOS pipeline. To mimic camera distortions and noise in real-world image acquisition, we added different types of noise and distortion. Dataset Summary The dataset contains over 63,917 annotated images with instance masks for varied… See the full description on the dataset page: https://huggingface.co/datasets/RiceD2KLab/SWiM-SpacecraftWithMasks.imageimage-segmentation2 likes6k downloads6mo agoHugging Face12Voxel51 /spatial_lm_dataset Dataset Card for Spatial LM This is a FiftyOne 3D dataset with 19,992 samples representing indoor room point clouds with structured 3D layout and object annotations from the SpatialLM benchmark. Each sample is an .fo3d scene containing a coloured point cloud with overlaid 3D bounding box annotations for walls, doors, windows, and furniture/objects — all browsable and queryable in the FiftyOne App. Installation If you haven't already, install FiftyOne: pip install -U… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/spatial_lm_dataset.3dobject-detection1K<n<10K2 likes4.9k downloads6mo agoHugging Face13amayuelas /aya-mm-exams-spanish-medicalMedical Spanish Exams for the Multimodal Aya Exams Projects. Questions available in file: data.json Images stored in: /images Original data and file available here: link imagemultiple-choicen<1K0 likes4.5k downloads2y agoHugging Face14stdKonjac /Sparkle Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance Ziyun Zeng, Yiqi Lin, Guoqiang Liang, and Mike Zheng Shou 📦 Dataset Sparkle is a large-scale video background replacement dataset comprising ~140K high-quality source–edited video pairs. It is fully open-sourced at 🤗stdKonjac/Sparkle. For full methodology and dataset details, please refer to our paper. The dataset is organized into five themes along different… See the full description on the dataset page: https://huggingface.co/datasets/stdKonjac/Sparkle.imagetext-to-video100K<n<1M1 likes4.4k downloads5mo agoHugging Face15LLDDSS /Awesome_Spatial_VQA_Benchmarksimage10K<n<100K1 likes4.1k downloads1y agoHugging Face16Spawning /PD12M PD12M Summary At 12.4 million image-caption pairs, PD12M is the largest public domain image-text dataset to date, with sufficient size to train foundation models while minimizing copyright concerns. Through the Source.Plus platform, we also introduce novel, community-driven dataset governance mechanisms that reduce harm and support reproducibility over time. Jordan Meyer Nicholas Padgett Cullen Miller Laura Exline Paper Datasheet Project About… See the full description on the dataset page: https://huggingface.co/datasets/Spawning/PD12M.image10M<n<100M186 likes3.4k downloads2y agoHugging Face17a8cheng /SpatialRGPT-Benchimage1K<n<10K13 likes2.3k downloads1y agoHugging Face18links-ads /spada-dataset SPADA Dataset This dataset contains images and sparse labels used in the paper Land Cover Segmentation with Sparse Annotations from Sentinel-2 Imagery , published at IGARSS 2023. Repository: https://github.com/links-ads/igarss-spada Paper: https://paperswithcode.com/paper/land-cover-segmentation-with-sparse Dataset Preparation The dataset has been compressed into segmented tarballs for ease of use within Git LFS (that is, tar > gzip > split). To revert the process… See the full description on the dataset page: https://huggingface.co/datasets/links-ads/spada-dataset.imageimage-segmentation10M<n<100M1 likes1.9k downloads2y agoHugging Face19spatialverse /SAGE-3D_InteriorGS_usdz SAGE-3D InteriorGS USDZ: USDZ-Format 3D Gaussian Scenes for Isaac Sim Paper | Project Page | Code InteriorGS dataset converted to USDZ format for seamless integration with NVIDIA Omniverse and Isaac Sim platforms. USDZ format InteriorGS data captured on Issac Sim 5.0. 📢 News 2025-12-15: Released SAGE-3D InteriorGS USDZ dataset with 1,000 converted scenes. 📋 Overview While the original InteriorGS dataset provides high-quality 3D… See the full description on the dataset page: https://huggingface.co/datasets/spatialverse/SAGE-3D_InteriorGS_usdz.3droboticsn<1K89 likes1.5k downloads2mo agoHugging Face20Avi2006 /spatial-moe-resultsimage1K<n<10K0 likes1.4k downloads4h agoHugging Face21jasonzhango /SPAR-Bench 🎯 Spatial Perception And Reasoning Benchmark (SPAR-Bench) A benchmark to evaluate spatial perception and reasoning in vision-language models (VLMs), with high-quality QA across 20 diverse tasks. SPAR-Bench is a high-quality benchmark for evaluating spatial perception and reasoning in vision-language models (VLMs). It covers 20 diverse spatial tasks across single-view, multi-view, and video settings, with a total of 7,207 manually verified… See the full description on the dataset page: https://huggingface.co/datasets/jasonzhango/SPAR-Bench.image1K<n<10K6 likes1.3k downloads1y agoHugging Face22juliensimon /spacex-launches SpaceX Launch History Credit: NASA Part of a dataset collection on Hugging Face. Dataset description Complete record of every SpaceX launch from spacex.com, including mission descriptions, pre/post-launch timelines, and photo galleries. Covers Falcon 1, Falcon 9, Falcon Heavy, and Starship missions. The data is sourced from the official SpaceX content API and organized into three tables that can be joined on the slug field: launches (one row per mission… See the full description on the dataset page: https://huggingface.co/datasets/juliensimon/spacex-launches.imagetabular-classificationn<1K0 likes1.3k downloads1d agoHugging Face23remyxai /vqasynth_spacellava VQASynth_spacellava Uses the VQASynth pipeline to synthesize spatialVQA samples, mixed with general VQA samples used to fine-tune LLaVA-v1.5-13b. imagevisual-question-answering10K<n<100K14 likes1.2k downloads2y agoHugging Face24spatial-reason /llava_trajectoriesimage1K<n<10K0 likes1.2k downloads7mo agoHugging Face25lerobot /libero_spatial_imageThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 432, "total_frames": 52970, "total_tasks": 10, "chunks_size": 1000, "fps": 10, "splits": { "train": "0:432" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path": "videos/{video_key}/chunk-{chunk_index:03d}/file-{file_index:03d}.mp4"… See the full description on the dataset page: https://huggingface.co/datasets/lerobot/libero_spatial_image.imagerobotics10K<n<100K7 likes1.2k downloads7mo agoHugging Face26dronefreak /SPA-Data SPA-Data: Real-World Single-Image Deraining Dataset (Unofficial, Subsampled Mirror) Unofficial, subsampled redistribution of SPA-Data, the real-world rain/rain-free image dataset from SPANet (Wang et al., CVPR 2019), packaged in a directory layout directly consumable by ClearView's SPADataDataset parser. Disclaimer This repository is not an official release of SPA-Data. SPA-Data was created by Tianyu Wang, Xin Yang, Ke Xu, Shaozhe Chen, Qiang Zhang… See the full description on the dataset page: https://huggingface.co/datasets/dronefreak/SPA-Data.imageimage-to-image10K<n<100K1 likes1.2k downloads6d agoHugging Face27spatial-reason /gpt5_trajectories_finalimage1K<n<10K0 likes1.1k downloads6mo agoHugging Face28austinpatel /libero_gen_spatial_combination_train_openpiimage1M<n<10M0 likes1.1k downloads5mo agoHugging Face29sunyoung00 /spao-aiimage1K<n<10K0 likes1k downloads5mo agoHugging Face30lmms-lab-eval /Spatial457image10K<n<100K0 likes883 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.