CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alyaalmsouti /Pleural-Line-Segmentation-Masks Pleural-Line Masks with Stanford LUS Frames Dataset Summary This dataset contains pleural-line masks and their corresponding lung-ultrasound frames for anatomy-guided video classification. The mask set includes: masks created by four human annotators via sam2 model point promting and video aggregation; masks predicted by a U-Net and subsequently reviewed and validated; and the metadata required to reproduce the training pipeline. The ultrasound frames originate… See the full description on the dataset page: https://huggingface.co/datasets/alyaalmsouti/Pleural-Line-Segmentation-Masks.image-segmentation1 likes12k downloads2mo agoHugging Face02SakethVemula /fixed-tokenizer-segments0 likes10k downloads5mo agoHugging Face03nielsr /image-segmentation-toy-dataimagen<1K0 likes9.4k downloads4y agoHugging Face04nielsRocholl /universal-lesion-segmentation Universal Lesion Segmentation Datasets A collection of public medical imaging datasets for lesion segmentation in CT scans. These are the datasets exactly as downloaded from their original sources. Datasets This repository contains the following datasets: CECT - Liver (primary). Luo J, Wang X, Zhang Y, et al. Comprehensive multi-phase three-dimensional contrast-enhanced CT imaging dataset for primary liver cancer. Scientific Data. 2025;12(1):768.… See the full description on the dataset page: https://huggingface.co/datasets/nielsRocholl/universal-lesion-segmentation.3dimage-segmentation2 likes5.1k downloads7mo agoHugging Face05SakethVemula /sslm-corpus-segments0 likes3.3k downloads7mo agoHugging Face06syvai /p1-segmentsgated DR P1 speech segments Dataset Danish speech clips from DR P1, in mono 16 kHz OGG/Opus, with verbatim text, timing, and speaker metadata. Transcript text and speaker attribution may contain automated errors. Source The recordings cover roughly 2006–2022 and come from DR P1 recordings in kb.dk’s DR archive. Audio is sourced through the pinned syvai/p1 revision 449b9c2294026df6d0d37538f279fdec03f565ff. Transcripts were generated with ElevenLabs… See the full description on the dataset page: https://huggingface.co/datasets/syvai/p1-segments.audioautomatic-speech-recognition1M<n<10M4 likes2.8k downloads7d agoHugging Face07tsrobcvai /Synthetic_Dataset_for_Stirrup_Rebar_SegmentationA Synthetic Dataset for Stirrup Rebar Segmentation The dataset contains: A synthetic training set of 12,000 images, and a synthetic validation set of 4,000. A synthetic test set of 4,800 images (only top rebars are annotated). A real-world test set of 233 images (only top rebars are annotated). Diverse rebar specifications, stacking, lighting, distractors, and background conditions. Before usage mkdir -p train_syn/train2017 mv train_syn/train2017_sub{1,2,3}/*… See the full description on the dataset page: https://huggingface.co/datasets/tsrobcvai/Synthetic_Dataset_for_Stirrup_Rebar_Segmentation.3 likes2.7k downloads1y agoHugging Face08tkdgur658 /Imbalanced_Segmentation_Datasetsimage1K<n<10K0 likes2.7k downloads5mo agoHugging Face09hf-internal-testing /mask-for-image-segmentation-testsimagen<1K1 likes2.5k downloads4y agoHugging Face10Ehsan-rmz /lgg-mri-segmentation-research LGG Brain MRI Segmentation with Genomic Clusters This repository provides a Patient-Centric version of the Lower-Grade Glioma (LGG) Segmentation dataset. While other versions of this data exist, they often treat slices as independent images. This version preserves the 3D patient volume and integrates all genomic/clinical labels directly into a multimodal-ready format. 🌟 Why This Version? Developed for Multimodal AI Research, this dataset addresses several limitations… See the full description on the dataset page: https://huggingface.co/datasets/Ehsan-rmz/lgg-mri-segmentation-research.imageimage-segmentationn<1K1 likes2.5k downloads9mo agoHugging Face11SakethVemula /hnet-segmentstext10M<n<100M0 likes2.5k downloads7mo agoHugging Face12bassatbassat /animalclef2026-segmented AnimalCLEF2026 Segmented This Hugging Face dataset repo provides segmented and preprocessed derivatives of the official AnimalCLEF26 dataset released through the Kaggle competition: AnimalCLEF26 @ CVPR & CLEF Kaggle Competition The dataset was processed by applying animal segmentation to the original competition images in order to reduce background noise and improve downstream animal re-identification and classification experiments. This repository is packaged in imagefolder… See the full description on the dataset page: https://huggingface.co/datasets/bassatbassat/animalclef2026-segmented.imageimage-classification10K<n<100K1 likes2.4k downloads5mo agoHugging Face13Jackmin108 /bert-base-uncased-refined-web-segment0 Dataset Card for "bert-base-uncased-refined-web-segment0" More Information needed 100M<n<1B0 likes2.3k downloads3y agoHugging Face14Haitam03 /warsh-segments-v3 Haitam03/warsh-v3 Warsh (Rewayat Warsh A'n Nafi') Quran recitation, segmented at waqf with obadx/recitation-segmenter-v2. Built with warsh-data. Layout path what data/<reciter>/<surah>.parquet one file per source recording, audio embedded as 16 kHz mono FLAC raw/<reciter>/<surah>.mp3 the source recording it came from segment_params.json the settings this corpus was produced with One parquet per source recording, named after it, so re-running a… See the full description on the dataset page: https://huggingface.co/datasets/Haitam03/warsh-segments-v3.audioautomatic-speech-recognition100K<n<1M0 likes2.1k downloads28d agoHugging Face15changelinglab /cv-v1.0-segment CommonVoice v1 Phone-Segment Alignments Phone-level time alignments for 10 languages of Mozilla Common Voice, packaged in a canonical segmentation schema with embedded 16 kHz audio. The phone boundaries come from the charsiu/cv_ali release of MFA alignments; the audio and transcripts come from Common Voice Corpus 13.0 (2023-03-09). Dataset summary lang train rows train hrs val rows val hrs test rows test hrs en 1,008,669 1,354.0 3,537 4.9 1,285 1.7 rw… See the full description on the dataset page: https://huggingface.co/datasets/changelinglab/cv-v1.0-segment.audioautomatic-speech-recognition1M<n<10M3 likes2k downloads6mo agoHugging Face16cg1177 /hacs_segment_internvideo2_6b_w16_s80 likes1.8k downloads3y agoHugging Face17Charlie911 /tmmluplus_CKIP_segmentedtext10K<n<100K0 likes1.8k downloads2y agoHugging Face18CATMuS /medieval-segmentation Dataset Card for CATMuS Medieval (Segmentation Version) Join our Discord to ask questions about the dataset: Dataset Details CATMuS Medieval Segmentation (Consistent Approaches to Transcribing Manuscripts) is a specialized dataset designed for layout analysis of medieval manuscripts using the SegmOnto vocabulary for region and line classification. This dataset addresses the challenges associated with establishing consistent ground truth in layout analysis tasks… See the full description on the dataset page: https://huggingface.co/datasets/CATMuS/medieval-segmentation.imageimage-segmentation1K<n<10K7 likes1.7k downloads2y agoHugging Face19Voxel51 /Football-Player-Segmentation Dataset Card for football-player-segmentation This dataset is specifically designed for computer vision tasks related to player detection and segmentation in foot goalkeeperders, and forwards, captured from various angles and distances. This is a FiftyOne dataset with 512 samples. Installation If you haven't already, install FiftyOne: pip install -U fiftyone Usage import fiftyone as fo import fiftyone.utils.huggingface as fouh # Load the dataset #… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/Football-Player-Segmentation.imageobject-detectionn<1K5 likes1.6k downloads2y agoHugging Face20Voxel51 /segment_anything_video_subset51 Dataset Card for Segment Anything Video (Subset 51) This is a FiftyOne dataset containing 917 video samples from the SA-V (Segment Anything Video) dataset. The videos are at 6 fps (matching the annotation cadence) and include both manual and automatic masklet (object mask tracklets) annotations for video object segmentation tasks. These are the videos from Subset 51 of the full dataset. Installation If you haven't already, install FiftyOne: pip install -U fiftyone… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/segment_anything_video_subset51.videon<1K1 likes1.6k downloads7mo agoHugging Face21amaye15 /object-segmentationimagen<1K5 likes1.5k downloads11mo agoHugging Face22Voxel51 /AVM_Segmentation_train Dataset Card for AVM (Around View Monitoring) Semantic Segmentation Dataset This repository provides a FiftyOne-compatible version of the AVM semantic segmentation dataset for autonomous parking systems, with enhanced metadata and visualization capabilities. This is a FiftyOne dataset with 6763 samples. Installation If you haven't already, install FiftyOne: pip install -U fiftyone Usage import fiftyone as fo from fiftyone.utils.huggingface import… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/AVM_Segmentation_train.imageimage-classification1K<n<10K1 likes1.4k downloads11mo agoHugging Face23ExtendedRealityLab /RGB-D-SegmentEgocentricBodiesannotations_creators: - other language: - en language_creators: - other license: - odc-by multilinguality: - monolingual pretty_name: 'RGB-D-SegmentEgocentricBodies ' size_categories: - 1K<n<10K source_datasets: - original tags: - egocentric segmentation - extended reality - xr - human-body - mixed-reality - avatar task_categories: - image-segmentation - depth-estimation task_ids: - semantic-segmentation - features: - name: image dtype: image - name: depth dtype: image… See the full description on the dataset page: https://huggingface.co/datasets/ExtendedRealityLab/RGB-D-SegmentEgocentricBodies.image1 likes1.2k downloads9mo agoHugging Face24julioojalvo /synthetic_kidney_stone_segmentation_dataimage10K<n<100K1 likes1.1k downloads3mo agoHugging Face25espnet /ace-opencpop-segments Citation Information @misc{shi2024singingvoicedatascalingup, title={Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing}, author={Jiatong Shi and Yueqian Lin and Xinyi Bai and Keyi Zhang and Yuning Wu and Yuxun Tang and Yifeng Yu and Qin Jin and Shinji Watanabe}, year={2024}, eprint={2401.17619}, archivePrefix={arXiv}, primaryClass={cs.SD}, url={https://arxiv.org/abs/2401.17619}, } audiotext-to-audio100K<n<1M8 likes1k downloads2y agoHugging Face26lauesa1 /minne-apple-segmentationA version of MinneApple with coco annotations image1K<n<10K0 likes978 downloads1y agoHugging Face27DenisaBumba /rfdetr-segmentation-leibniz-dataset Dataset Card for Leibniz's Manuscripts (Instance Segmentation Dataset) This dataset comprises instance segmentation annotations in raw COCO format, used to train an RF-DETR-Seg-nano model for the automatic recognition of textual, graphical, and mathematical expression zones within the manuscripts of the philosopher and mathematician Gottfried Wilhelm Leibniz (17th-early 18th c.). Dataset Details Uses Direct Use This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/DenisaBumba/rfdetr-segmentation-leibniz-dataset.imageimage-segmentation1K<n<10K0 likes963 downloads2mo agoHugging Face28andrematte /dam-segmentation Dam Segmentation Dataset Multispectral UAV Remote Sensing Data for Embankment Dam Segmentation Dataset Summary This dataset contains a series of multispectral image slices captured at the embankment dams and dikes of the Belo Monte Hydroelectric Complex, located in the state of Pará, northern Brazil. Each image is paired with its respective NDRE vegetation index values, binary segmentation mask and multiclass segmentation mask. The multispectral images were… See the full description on the dataset page: https://huggingface.co/datasets/andrematte/dam-segmentation.imageimage-segmentationn<1K2 likes934 downloads2y agoHugging Face29Novel-BioMedAI /Medical_Segmentation_Decathlon 🏆 Medical Segmentation Decathlon Dataset 📝 Overview The Medical Segmentation Decathlon (MSD) is a comprehensive benchmark dataset for validating algorithms in 3D medical image segmentation. It includes 10 distinct tasks, each with unique challenges like small data sizes, unbalanced labels, varying object scales, multi-class labels, and multimodal imaging. 🔗 Dataset Access 🌐 Website: Medical Decathlon 📂 Google Drive: MSD Google Drive 🧩 Task… See the full description on the dataset page: https://huggingface.co/datasets/Novel-BioMedAI/Medical_Segmentation_Decathlon.2 likes898 downloads1y agoHugging Face30Voxel51 /SpectralWaste-Segmentation SpectralWaste Segmentation → FiftyOne (Grouped RGB + Hyperspectral) The labeled split of SpectralWaste, rebuilt as a grouped FiftyOne dataset pairing each colour frame with its hyperspectral cube. The recordings come from a working waste-sorting plant, looking down at the conveyor as material passes. Each frame is captured twice over, once in colour and once by a shortwave infrared camera reading 224 bands from about 900 to 1700 nm. Material that looks identical in colour… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/SpectralWaste-Segmentation.imageimage-segmentation1K<n<10K2 likes893 downloads13d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.