datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
3DA-VTG
3DA-VTG
3DA-VTG is a visuo-tactile grasp-stability dataset prepared for the public
SGA-GSN release. It contains paired visual and tactile observations, object-level
metadata, and binary grasp-stability labels.
License: Creative Commons Attribution-NonCommercial-ShareAlike 4.0
International (CC BY-NC-SA 4.0). The dataset is derived from GraspNet-1Billion
assets and follows the GraspNet non-commercial redistribution terms. Commercial
use requires permission from the GraspNet team.… See the full description on the dataset page: https://huggingface.co/datasets/robotic-vt-grasp-project/3DA-VTG.3DSRBench
3DSRBench Circular Evaluation Package
This upload contains the processed TSV required by PhysBrainEvalKit for 3DSRBench circular evaluation. The TSV embeds the evaluation images as base64 data, so no separate image archive is required.
Files in this repository
3dsrbench_v1_vlmevalkit_circular.tsv: processed evaluation data used by PhysBrainEvalKit.
compute_3drbench_results_circular.py: optional standalone result computation script.
.gitattributes: large-file… See the full description on the dataset page: https://huggingface.co/datasets/VLyb/3DSRBench.indoor-smartphone-3d-reconstruction-control
Indoor Smartphone 3D Reconstruction Control Dataset
Набор данных подготовлен для сравнения методов восстановления 3D-сцены в помещении по короткому видео со смартфона. Он содержит три небольшие indoor-сцены, очищенные публикационные видеоролики, подвыборки кадров, ручную CVAT-разметку и физически измеренные контрольные расстояния.
Датасет предназначен для оценки геометрической согласованности результатов 3D-реконструкции. Разметка не является обучающей dense-разметкой глубины:… See the full description on the dataset page: https://huggingface.co/datasets/Maksonchek/indoor-smartphone-3d-reconstruction-control.welsh-speech-3d-meshes
Welsh Speech Dataset - 3D Facial Meshes
3D facial reconstructions from the Welsh Speech Dataset.
Contents
3D meshes (.obj files) - One per frame
Texture maps (.png files) - Fused left-right stereo images from 3DMD
Captured using 3DMD 6-camera system
~330 zip files (one per speaker-phrase sequence)
File Structure
Files are organized as zip archives in the meshes/ directory, one zip per speaker-phrase sequence:
meshes/
├── speaker_01_phrase_01.zip
├──… See the full description on the dataset page: https://huggingface.co/datasets/arvinsingh/welsh-speech-3d-meshes.colon-3d-crc1
