datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Pleural-Line-Segmentation-Masks
Pleural-Line Masks with Stanford LUS Frames
Dataset Summary
This dataset contains pleural-line masks and their corresponding lung-ultrasound frames for anatomy-guided video classification.
The mask set includes:
masks created by four human annotators via sam2 model point promting and video aggregation;
masks predicted by a U-Net and subsequently reviewed and validated; and
the metadata required to reproduce the training pipeline.
The ultrasound frames originate… See the full description on the dataset page: https://huggingface.co/datasets/alyaalmsouti/Pleural-Line-Segmentation-Masks.fixed-tokenizer-segmentsimage-segmentation-toy-datauniversal-lesion-segmentation
Universal Lesion Segmentation Datasets
A collection of public medical imaging datasets for lesion segmentation in CT scans. These are the datasets exactly as downloaded from their original sources.
Datasets
This repository contains the following datasets:
CECT - Liver (primary). Luo J, Wang X, Zhang Y, et al. Comprehensive multi-phase three-dimensional contrast-enhanced CT imaging dataset for primary liver cancer. Scientific Data. 2025;12(1):768.… See the full description on the dataset page: https://huggingface.co/datasets/nielsRocholl/universal-lesion-segmentation.sslm-corpus-segmentsp1-segments
DR P1 speech segments
Dataset
Danish speech clips from DR P1, in mono 16 kHz OGG/Opus, with verbatim text, timing, and speaker metadata. Transcript text and speaker attribution may contain automated errors.
Source
The recordings cover roughly 2006–2022 and come from DR P1 recordings in kb.dk’s DR archive. Audio is sourced through the pinned syvai/p1 revision 449b9c2294026df6d0d37538f279fdec03f565ff. Transcripts were generated with ElevenLabs… See the full description on the dataset page: https://huggingface.co/datasets/syvai/p1-segments.Synthetic_Dataset_for_Stirrup_Rebar_SegmentationA Synthetic Dataset for Stirrup Rebar Segmentation
The dataset contains:
A synthetic training set of 12,000 images, and a synthetic validation set of 4,000.
A synthetic test set of 4,800 images (only top rebars are annotated).
A real-world test set of 233 images (only top rebars are annotated).
Diverse rebar specifications, stacking, lighting, distractors, and background conditions.
Before usage
mkdir -p train_syn/train2017
mv train_syn/train2017_sub{1,2,3}/*… See the full description on the dataset page: https://huggingface.co/datasets/tsrobcvai/Synthetic_Dataset_for_Stirrup_Rebar_Segmentation.Imbalanced_Segmentation_Datasetsmask-for-image-segmentation-testslgg-mri-segmentation-research
LGG Brain MRI Segmentation with Genomic Clusters
This repository provides a Patient-Centric version of the Lower-Grade Glioma (LGG) Segmentation dataset. While other versions of this data exist, they often treat slices as independent images. This version preserves the 3D patient volume and integrates all genomic/clinical labels directly into a multimodal-ready format.
🌟 Why This Version?
Developed for Multimodal AI Research, this dataset addresses several limitations… See the full description on the dataset page: https://huggingface.co/datasets/Ehsan-rmz/lgg-mri-segmentation-research.hnet-segmentsanimalclef2026-segmented
AnimalCLEF2026 Segmented
This Hugging Face dataset repo provides segmented and preprocessed derivatives of the official AnimalCLEF26 dataset released through the Kaggle competition:
AnimalCLEF26 @ CVPR & CLEF Kaggle Competition
The dataset was processed by applying animal segmentation to the original competition images in order to reduce background noise and improve downstream animal re-identification and classification experiments.
This repository is packaged in imagefolder… See the full description on the dataset page: https://huggingface.co/datasets/bassatbassat/animalclef2026-segmented.bert-base-uncased-refined-web-segment0
Dataset Card for "bert-base-uncased-refined-web-segment0"
More Information needed
warsh-segments-v3
Haitam03/warsh-v3
Warsh (Rewayat Warsh A'n Nafi') Quran recitation, segmented at waqf with
obadx/recitation-segmenter-v2.
Built with warsh-data.
Layout
path
what
data/<reciter>/<surah>.parquet
one file per source recording, audio embedded as 16 kHz mono FLAC
raw/<reciter>/<surah>.mp3
the source recording it came from
segment_params.json
the settings this corpus was produced with
One parquet per source recording, named after it, so re-running a… See the full description on the dataset page: https://huggingface.co/datasets/Haitam03/warsh-segments-v3.cv-v1.0-segment
CommonVoice v1 Phone-Segment Alignments
Phone-level time alignments for 10 languages of Mozilla Common Voice,
packaged in a canonical segmentation schema with embedded 16 kHz audio. The
phone boundaries come from the charsiu/cv_ali
release of MFA alignments; the audio and transcripts come from
Common Voice Corpus 13.0 (2023-03-09).
Dataset summary
lang
train rows
train hrs
val rows
val hrs
test rows
test hrs
en
1,008,669
1,354.0
3,537
4.9
1,285
1.7
rw… See the full description on the dataset page: https://huggingface.co/datasets/changelinglab/cv-v1.0-segment.hacs_segment_internvideo2_6b_w16_s8tmmluplus_CKIP_segmentedmedieval-segmentation
Dataset Card for CATMuS Medieval (Segmentation Version)
Join our Discord to ask questions about the dataset:
Dataset Details
CATMuS Medieval Segmentation (Consistent Approaches to Transcribing Manuscripts) is a specialized dataset designed for layout analysis of medieval manuscripts using the SegmOnto vocabulary for region and line classification. This dataset addresses the challenges associated with establishing consistent ground truth in layout analysis tasks… See the full description on the dataset page: https://huggingface.co/datasets/CATMuS/medieval-segmentation.Football-Player-Segmentation
Dataset Card for football-player-segmentation
This dataset is specifically designed for computer vision tasks related to player detection and segmentation in foot goalkeeperders, and forwards, captured from various angles and distances.
This is a FiftyOne dataset with 512 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
import fiftyone.utils.huggingface as fouh
# Load the dataset
#… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/Football-Player-Segmentation.segment_anything_video_subset51
Dataset Card for Segment Anything Video (Subset 51)
This is a FiftyOne dataset containing 917 video samples from the SA-V (Segment Anything Video) dataset. The videos are at 6 fps (matching the annotation cadence) and include both manual and automatic masklet (object mask tracklets) annotations for video object segmentation tasks.
These are the videos from Subset 51 of the full dataset.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/segment_anything_video_subset51.object-segmentationAVM_Segmentation_train
Dataset Card for AVM (Around View Monitoring) Semantic Segmentation Dataset
This repository provides a FiftyOne-compatible version of the AVM semantic segmentation dataset for autonomous parking systems, with enhanced metadata and visualization capabilities.
This is a FiftyOne dataset with 6763 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/AVM_Segmentation_train.RGB-D-SegmentEgocentricBodiesannotations_creators:
- other
language:
- en
language_creators:
- other
license:
- odc-by
multilinguality:
- monolingual
pretty_name: 'RGB-D-SegmentEgocentricBodies '
size_categories:
- 1K<n<10K
source_datasets:
- original
tags:
- egocentric segmentation
- extended reality
- xr
- human-body
- mixed-reality
- avatar
task_categories:
- image-segmentation
- depth-estimation
task_ids:
- semantic-segmentation
- features:
- name: image
dtype: image
- name: depth
dtype: image… See the full description on the dataset page: https://huggingface.co/datasets/ExtendedRealityLab/RGB-D-SegmentEgocentricBodies.synthetic_kidney_stone_segmentation_dataace-opencpop-segments
Citation Information
@misc{shi2024singingvoicedatascalingup,
title={Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing},
author={Jiatong Shi and Yueqian Lin and Xinyi Bai and Keyi Zhang and Yuning Wu and Yuxun Tang and Yifeng Yu and Qin Jin and Shinji Watanabe},
year={2024},
eprint={2401.17619},
archivePrefix={arXiv},
primaryClass={cs.SD},
url={https://arxiv.org/abs/2401.17619},
}
minne-apple-segmentationA version of MinneApple with coco annotations
rfdetr-segmentation-leibniz-dataset
Dataset Card for Leibniz's Manuscripts (Instance Segmentation Dataset)
This dataset comprises instance segmentation annotations in raw COCO format, used to train an RF-DETR-Seg-nano model for the automatic recognition of textual, graphical, and mathematical expression zones within the manuscripts of the philosopher and mathematician Gottfried Wilhelm Leibniz (17th-early 18th c.).
Dataset Details
Uses
Direct Use
This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/DenisaBumba/rfdetr-segmentation-leibniz-dataset.dam-segmentation
Dam Segmentation Dataset
Multispectral UAV Remote Sensing Data for Embankment Dam Segmentation
Dataset Summary
This dataset contains a series of multispectral image slices captured at the embankment dams and dikes
of the Belo Monte Hydroelectric Complex, located in the state of Pará, northern Brazil. Each image is
paired with its respective NDRE vegetation index values, binary segmentation mask and multiclass
segmentation mask.
The multispectral images were… See the full description on the dataset page: https://huggingface.co/datasets/andrematte/dam-segmentation.Medical_Segmentation_Decathlon
🏆 Medical Segmentation Decathlon Dataset
📝 Overview
The Medical Segmentation Decathlon (MSD) is a comprehensive benchmark dataset for validating algorithms in 3D medical image segmentation. It includes 10 distinct tasks, each with unique challenges like small data sizes, unbalanced labels, varying object scales, multi-class labels, and multimodal imaging.
🔗 Dataset Access
🌐 Website: Medical Decathlon
📂 Google Drive: MSD Google Drive
🧩 Task… See the full description on the dataset page: https://huggingface.co/datasets/Novel-BioMedAI/Medical_Segmentation_Decathlon.SpectralWaste-Segmentation
SpectralWaste Segmentation → FiftyOne (Grouped RGB + Hyperspectral)
The labeled split of SpectralWaste, rebuilt as a grouped FiftyOne dataset pairing each colour frame with its hyperspectral cube.
The recordings come from a working waste-sorting plant, looking down at the conveyor as material passes. Each frame is captured twice over, once in colour and once by a shortwave infrared camera reading 224 bands from about 900 to 1700 nm. Material that looks identical in colour… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/SpectralWaste-Segmentation.
