datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Multitask-National-Speech-Corpus-v1-extendpretrain_mini_extractedExtremeDegradationBench
Extreme Degradation Bench
A benchmark for vocal restoration under extreme degradation.
See ./noisy for the raw, recorded audio files, ./predictions for all existing predictions, and ./pairwise-ranking.csv for all existing raw pairwise ranking data sourced in https://arxiv.org/abs/2510.21659.
See ./app.py for the Gradio application used for the pairwise rankings. Directly running app.py should automatically create a ./ratings directory for logging per-session vote information.
Please… See the full description on the dataset page: https://huggingface.co/datasets/smulelabs/ExtremeDegradationBench.dusha_extra_data
Dataset Card for "dusha_extra_data"
More Information needed
SALMon_Flow-SLM-1B-Extended
SALMon Normalized Dataset
This repo preserves the SALMon per-config folder layout while normalizing
mismatched schema details across model families.
torgoCRPIH_UVigo-GL-Voices_extended
CRPIH_UVigo-GL-Voices: Galician TTS dataset
CRPIH_UVigo-GL-Voices is a Galician TTS multi-speaker dataset containing audio recordings from four different speakers (two female and two male voices). The characteristics of each voice are detailed in the table below:
Voice name
Gender
Speaker
Recording
# Utts
Duration
Iago
Male
Amateur
Radio studio
1,316
1h 13min
Icía
Female
Amateur
Semi-professional studio
2,950
4h 5min
Paulo
Male
Amateur
Radio studio
1,316
1h 15min… See the full description on the dataset page: https://huggingface.co/datasets/proxectonos/CRPIH_UVigo-GL-Voices_extended.synthetic-parallel-external
Synthetic Parallel EN↔LG — external
Voice-controlled synthetic parallel speech dataset for Luganda-English
speech-to-speech translation, generated by the Hibiki-Zero fine-tuning pipeline.
Generation
Component
Model
Translation
Sunbird/translate-nllb-3.3b-salt
TTS
Sunbird/orpheus-3b-tts-multilingual
English speakers: salt_eng_0001, salt_eng_0002, salt_eng_0003
Luganda speakers: salt_lug_0001, waxal_lug_0001, waxal_lug_0002, waxal_lug_0003, waxal_lug_0004… See the full description on the dataset page: https://huggingface.co/datasets/yigagilbert/synthetic-parallel-external.ua-speechSTARSS23_extraIqra_Extra_IS26tibetan-english-8s_extend_speechEnvironmentalSoundClassification_ESC50-ExteriorAndUrbanNoises
Dataset Card for "environmental_sound_classification_exterior_and_urban_noises_ESC50"
More Information needed
haitian-creole-processed-extendeddindic_voices_only_extemporemozgach_multimodal_extraSpeech-Instructions-Extraextracted-id-subbed-video-v3nUrdu-TTS-test-3-extSALMon_Flow-SLM-1B-Extended-depSpeaking_Rate_Extremesindian_english_extendedEnvironmentalSoundClassification_ESC50-ExteriorAndUrbanNoises_TTSextracted-untestedextracted-id-subbed-video-v2Duration_Extremes_ExtractionEraX-WoW-dataset-EXTRA-84k-1Mar2025audio-prompt-coco-balanced-extendedextracted-id-subbed-video-v4nextreme_asr_pony
