track1
Datasets
All datasets matching “track1”urgent26_track1_universal_seThe pre-simulated universal speech enhancement training and validation set of the ICASSP 2026 URGENT speech enhancement challenge, Track1.
Please check our GitHub Repo and webpage for more details.
How to use:
Use tar to decompress the dataset
cat ./urgent26_track2_se_dataset.tgz.* | tar xzv
The pre-simulated dataset can be loaded by the PreSimulatedDataset in the URGENT 2026 baseline code.
Directory structure:
.
├── data
│ ├── train_simulation # train set… See the full description on the dataset page: https://huggingface.co/datasets/lichenda/urgent26_track1_universal_se.mls_hq_urgent_track1ADD2023_track12_test_r1
ADD 2023 — Track 1.2 (FG-D Detection), Round 1 Test · labels only
Benchmark-ready packaging of the Round 1 (R1) evaluation partition of Track 1.2
(Fake Game Detection / FG-D) from the ADD 2023 challenge — the Second Audio
Deepfake Detection Challenge (arXiv 2305.13774).
Binary anti-spoofing over Mandarin audio: bonafide (genuine human speech) vs.
spoof (synthesized / fake speech). Track 1.2 is the defending side of an
attack-and-defense game; the fakes are produced by the… See the full description on the dataset page: https://huggingface.co/datasets/SpeechAntiSpoofingBenchmarks/ADD2023_track12_test_r1.ELSA1M_track1
ELSA - Multimedia use case
ELSA Multimedia is a large collection of Deep Fake images, generated using diffusion models
Dataset Summary
This dataset was developed as part of the EU project ELSA. Specifically for the Multimedia use-case.
Official webpage: https://benchmarks.elsa-ai.eu/
This dataset aims to develop effective solutions for detecting and mitigating the spread of deep fake images in multimedia content. Deep fake images, which are highly realistic and deceptive… See the full description on the dataset page: https://huggingface.co/datasets/elsaEU/ELSA1M_track1.AT-ADD-Track1
AT-ADD Track 1
This repository hosts Track 1 of the AT-ADD All-Type Audio Deepfake Detection Challenge. It contains the released audio splits and privacy-preserving sample-level metadata for non-commercial academic research and education.
Access
This is a gated dataset. Sign in to Hugging Face, review the access agreement, complete the short access form, and click Agree and access dataset. Access is granted automatically after acceptance.
Direct repository access… See the full description on the dataset page: https://huggingface.co/datasets/xieyuankun/AT-ADD-Track1.mls-hq-urgent-track1
Multilingual LibriSpeech HQ (MLS-HQ)
This is a mirror of the Multilingual LibriSpeech HQ (MLS-HQ) data used in URGENT 2025 Track 1.
The original files were converted from FLAC to Opus to reduce the size and accelerate streaming.
Sampling rate: 48 kHz (resampled from 44.1 kHz to support Opus format)
Channels: 1
Format: Opus
Splits:
spanish: 150 hours, 36031 utterances
german: 150 hours, 35890 utterances
french: 150 hours, 36078 utterances
License: CC0 1.0
Source:… See the full description on the dataset page: https://huggingface.co/datasets/philgzl/mls-hq-urgent-track1.
