datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
strikerData
🎧 StrikerData
Overview
StrikerData is an audio dataset developed by Strikersoft for research and development in audio and speech technologies.It contains human speech, environmental noise, and other sound types. The dataset is available for non-commercial use only, except for the company Strikersoft.
Category
Percentage of Total Dataset
Clean human speech
20%
Distorted speech
15%
Human-made noise
15%
Non-human noise
50%
⚖️ License… See the full description on the dataset page: https://huggingface.co/datasets/strikersoft/strikerData.Streamer
GestureHYDRA: Semantic Co-speech Gesture Synthesis via Hybrid Modality Diffusion Transformer and Cascaded-Synchronized Retrieval-Augmented Generation.ICCV 2025
Dataset structure
The dataset contains three folders: train, test_seen, and test_unseen. Let's train train as an example for introduction.
-train/
├── audios/
│ ├── {anchor_id}/
│ │ └── {video_name_md5}/
│ │ └── {start_time}_{end_time}.wav
├── gestures/
│ ├── {anchor_id}/
│ │ └──… See the full description on the dataset page: https://huggingface.co/datasets/mumuwei/Streamer.stanislau-shumski-u-bitvakh-i-viaznitsakh-uspaminy-pra-1812-1848-gady-siargei-du
У бітвах і вязьніцах. Успаміны пра 1812–1848 гады
Metadata
Author: Станіслаў Шумскі
Title: У бітвах і вязьніцах. Успаміны пра 1812–1848 гады
Narrator: Сяргей Дубавец
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/stanislau-shumski-u-bitvakh-i-viaznitsakh-uspaminy-pra-1812-1848-gady-siargei-du.Pop-Rock-Hybrid-Stem-Dataset-cat001
Dataset Overview: Pop Rock Hybrid Stem Dataset (cat001)
This dataset contains a curated collection of original instrumental music designed for commercial and research applications in music analysis, audio modeling, and production workflows.
Every composition, arrangement, performance, sound design element, and production decision was created entirely through human musical and technical processes. All music contained in this dataset is 100% human-made (is_human_created: TRUE).… See the full description on the dataset page: https://huggingface.co/datasets/ToneCubeMedia/Pop-Rock-Hybrid-Stem-Dataset-cat001.video-interview-structural-segments
Video Interview Structural Segments Dataset
Опис датасету
Датасет містить анотації часових сегментів довгих відеоінтерв’ю та подкастів. Кожен сегмент має часові межі та одну з п’яти міток класу:
intro — вступна частина відео;
question — питання або репліка ведучого;
answer — відповідь гостя або основна змістова репліка;
ad — рекламний або службовий блок;
outro — завершення відео.
Датасет призначений для дослідження задач сегментації та мультимодальної… See the full description on the dataset page: https://huggingface.co/datasets/Cilach/video-interview-structural-segments.AudibleLight_Eigenmike32-5_DCASE-STARSS23_Dataset
AudibleLight Eigenmike32-5 DCASE-STARSS23 Dataset
This dataset contains 121 synthetic spatial audio scenes — 111 for training and 10 for evaluation — of 60 seconds each, generated with the AudibleLight dataset generator (DOI). Each scene is rendered as five independent simulated Eigenmike32 captures, with 32 channels per capture, resulting in 570 minutes of multichannel audio in total at 24 kHz.
Foreground Audio
Foreground events are sampled from ESC-50: Dataset for… See the full description on the dataset page: https://huggingface.co/datasets/PhilippXXY/AudibleLight_Eigenmike32-5_DCASE-STARSS23_Dataset.indicf5-nfe-9-10-11-stress-testds573-stat-xpriv-auds573-stat-s3sttivan-bunin-kazimir-stanislavavich-uladzimir-ragautsou
Казімір Станіслававіч
Metadata
Author: Іван Бунін
Title: Казімір Станіслававіч
Narrator: Уладзімір Рагаўцоў
Source Group: Аўдыёкнігі
Source: БЛР#аўдыякніга
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ivan-bunin-kazimir-stanislavavich-uladzimir-ragautsou.
