CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01labelmaker /arkit_labelmaker ARKit Labelmaker: A New Scale for Indoor 3D Scene Understanding [arxiv] [website] [checkpoints] [code] We complement ARKitScenes dataset with dense semantic annotations that are automatically generated at scale. This produces the first large-scale, real-world 3D dataset with dense semantic annotations. Training on this auto-generated data, we push forward the state-of-the-art performance on ScanNet and ScanNet200 with prevalent 3D semantic segmentation models. image-segmentation1K<n<10K3 likes69k downloads2y agoHugging Face02rerun /arkitscenes-rrd ARKitScenes → Rerun (.rrd) 5,015 ARKitScenes indoor iPhone/iPad captures, converted into layered Rerun recordings — including the data that exists only inside the dataset's .mov containers and appears in no published asset: 60 Hz camera poses (ARKit visionTransform, 6× denser than the published 10 Hz trajectory) — validated against the published trajectory per sequence and used when a rigid fit agrees within 3° / 10 cm (pose_source = mebx_stream_4_vision_transform), otherwise… See the full description on the dataset page: https://huggingface.co/datasets/rerun/arkitscenes-rrd.3d0 likes4.4k downloads2mo agoHugging Face03GaussianWorld /arkitscenes_mcmc_3dgsgated Data Statistics Scenes Mean PSNR ↑ Mean SSIM ↑ Mean LPIPS ↓ Mean Depth L1 ↓ Mean #3DGS Total #3DGS 1,290 31.63 dB 0.907 0.217 0.0051 m 1.149M 1.483B 3dother1K<n<10K2 likes2.5k downloads2mo agoHugging Face04pablovela5620 /arkitscenes-rrd ARKitScenes → Rerun (.rrd) 5,015 ARKitScenes indoor iPhone/iPad captures, converted into layered Rerun recordings — including the data that exists only inside the dataset's .mov containers and appears in no published asset: 60 Hz camera poses (ARKit visionTransform, 6× denser than the published 10 Hz trajectory) — validated against the published trajectory per sequence and used when a rigid fit agrees within 3° / 10 cm (pose_source = mebx_stream_4_vision_transform), otherwise… See the full description on the dataset page: https://huggingface.co/datasets/pablovela5620/arkitscenes-rrd.3d0 likes2.2k downloads2mo agoHugging Face05cywan /arkit_lowres_processed0 likes462 downloads2mo agoHugging Face06ZhengGuangze /ARKitScenes_preprocessed ARKitScenes Dataset for LBM Preprocessed ARKitScenes Dataset following CUT3R. Each scene contains the following structure when extracted: 40753679/ ├── lowres_depth | ├── 40753679_6790.148.png | ├── ... ├── vga_wide | ├── 40753679_6790.148.jpg | ├── ... ├── new_scene_metadata.npz └── scene_metadata.npz Data Format Details: vga_wide: RGB images in .jpg format lowres_depth: Depth data in .png format (numpy arrays) new_scene_metadata.npz: Camera data processed by… See the full description on the dataset page: https://huggingface.co/datasets/ZhengGuangze/ARKitScenes_preprocessed.2 likes180 downloads1y agoHugging Face07myned-ai /audio2face-mediapipe-arkit-teacher audio2face-mediapipe-arkit-teacher Left: source video frame (face-cropped). Middle: MediaPipe FaceLandmarker's 478 landmark points. Right: an illustrative subset of mp_bs — the 52-channel ARKit blendshape vector shipped in this dataset — as horizontal bars updating per frame. 14,703 emotional-speech clips, each annotated with a 52-channel ARKit blendshape sequence extracted by MediaPipe FaceLandmarker from the source video (or from audio-driven synthesis where no… See the full description on the dataset page: https://huggingface.co/datasets/myned-ai/audio2face-mediapipe-arkit-teacher.tabularaudio-classification10K<n<100K1 likes171 downloads4mo agoHugging Face08ysmao /arkitscenes-spatiallm ARKitScenes-SpatialLM Dataset ARkitScenes dataset preprocessed in SpatialLM format for oriented object bouding boxes detection with LLMs. Overview This dataset is derived from ARKitScenes 5,047 real-world indoor scenes captured using Apple's ARKit framework, preprocessed and formatted specifically for SpatialLM training. Data Extraction Point clouds and layouts are compressed in zip files. To extract the files, run the following script: cd arkitscenes-spatiallm… See the full description on the dataset page: https://huggingface.co/datasets/ysmao/arkitscenes-spatiallm.3d1K<n<10K1 likes148 downloads1y agoHugging Face09yslan /3dscene-arkitscene0 likes134 downloads11mo agoHugging Face10YangCaoCS /ARKitScenes_processedThe processed ARKitScenes datasets by VGGT-Det. Run 'cat train.tar.part_0{00,01,02} > train.tar' to merge the split archives into a single tar file. Run 'md5sum -c MD5SUMS.txt' to verify that the files were downloaded successfully. If the dataset is helpful, please cite: @inproceedings{ dehghan2021arkitscenes, title={{ARK}itScenes - A Diverse Real-World Dataset for 3D Indoor Scene Understanding Using Mobile {RGB}-D Data}, author={Gilad Baruch and Zhuoyuan Chen and Afshin… See the full description on the dataset page: https://huggingface.co/datasets/YangCaoCS/ARKitScenes_processed.0 likes100 downloads6mo agoHugging Face11myned-ai /audio2face-emotion-arkit-teacher audio2face-emotion-arkit-teacher Nyx avatar (Gaussian-splat head, ARKit-52 blendshape rig) driven by a surprise clip's blendshape labels derived from this dataset. 14,082 emotional-speech clips, each annotated with two parallel 52-channel ARKit blendshape sequences (NVIDIA Audio2Face-3D-v2.3.1-James and LAM_Audio2Expression) plus a 26-dimensional NVIDIA Audio2Emotion conditioning vector. Reference-only dataset — the original audio is not shipped. Each row contains a… See the full description on the dataset page: https://huggingface.co/datasets/myned-ai/audio2face-emotion-arkit-teacher.tabularaudio-classification10K<n<100K3 likes98 downloads4mo agoHugging Face12Pointcept /arkitscenes-compressed2 likes84 downloads2y agoHugging Face13THULab /audio2face-emotion-arkit-teacher Audio2Face Emotion ARKit Teacher Labels (TsFile) Apache TsFile version of myned-ai/audio2face-emotion-arkit-teacher. Overview 14,082 emotional-speech clips, each annotated with two parallel 52-channel ARKit blendshape sequences (NVIDIA Audio2Face-3D-v2.3.1-James and LAM_Audio2Expression) plus a 26-dimensional NVIDIA Audio2Emotion conditioning vector. Released by myned-ai to support teacher-student distillation research: condensing heavy GPU-bound audio2face… See the full description on the dataset page: https://huggingface.co/datasets/THULab/audio2face-emotion-arkit-teacher.timeseriesaudio-classification10K<n<100K0 likes66 downloads1mo agoHugging Face14Carlhahaha /arkitscenes-crossview-inpaint Carlhahaha/arkitscenes-crossview-inpaint Cross-view pairs generated from inpainting results (Topomap; meta root: /mnt/NAS/data/jz4725/topomap). Splits Uploaded splits: test, train, validation Schema (columns) Image_a: Original image of sample A (datasets.Image) Image_b: Original image of sample B (datasets.Image) Inpaint_b: Inpainted image of sample B (datasets.Image) mask_b: Mask used for inpainting on B (datasets.Image) point_b: Normalized centroid of B's… See the full description on the dataset page: https://huggingface.co/datasets/Carlhahaha/arkitscenes-crossview-inpaint.image10K<n<100K0 likes61 downloads1y agoHugging Face15robinwitch /xx_beat_arkit_moshi_2025_07_20_30fps_attntext1K<n<10K0 likes47 downloads1y agoHugging Face16yslan /3dscene-arkitscene-highres0 likes37 downloads11mo agoHugging Face17Icey444 /tis-scenes-arkit-mp3d3dn<1K0 likes34 downloads16d agoHugging Face18Pointcept /concerto_arkitscenes_compressed0 likes32 downloads1y agoHugging Face19xlu11 /arkitscenes-preprocessedtabular1K<n<10K1 likes27 downloads6mo agoHugging Face20escontra /gauss_gym_arkit3d1K<n<10K1 likes24 downloads11mo agoHugging Face21arkitex /synthetic_accented_englishaudio10K<n<100K0 likes20 downloads1y agoHugging Face22Gen3DF /Arkitscenes-Spatiallmgated ARKitScenes-SpatialLM Dataset ARkitScenes dataset preprocessed in SpatialLM format for oriented object bouding boxes detection with LLMs. Overview This dataset is derived from ARKitScenes 5,047 real-world indoor scenes captured using Apple's ARKit framework, preprocessed and formatted specifically for SpatialLM training. Data Extraction Point clouds and layouts are compressed in zip files. To extract the files, run the following script: cd arkitscenes-spatiallm… See the full description on the dataset page: https://huggingface.co/datasets/Gen3DF/Arkitscenes-Spatiallm.3d1K<n<10K5 likes14 downloads1y agoHugging Face23Arkitechtron /items_prompts_fulltext100K<n<1M0 likes12 downloads8mo agoHugging Face24Arkitechtron /items_raw_fulltabular100K<n<1M0 likes8 downloads9mo agoHugging Face25robinwitch /xx_beat_arkit_encodec2_2025_07_31_30fps_attntext1K<n<10K0 likes7 downloads1y agoHugging Face26Arkitechtron /items_raw_litetabular10K<n<100K0 likes7 downloads9mo agoHugging Face27robinwitch /xx_cbh_arkit_face_v1_moshi_2025_08_10_30fps_attntextn<1K0 likes5 downloads1y agoHugging Face282inf /arkitscenes_processedgatedimage0 likes4 downloads1y agoHugging Face29qqqqi /arkit_scan_proc0 likes3 downloads2mo agoHugging Face30Arkiteck /GrupoRTXtextn<1K0 likes2 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.