CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AudioLLMs /Multitask-National-Speech-Corpus-v1-extendaudio10M<n<100M5 likes5.8k downloads1y agoHugging Face02lainlives /pretrain_mini_extractedaudio1K<n<10K0 likes1.1k downloads2mo agoHugging Face03smulelabs /ExtremeDegradationBench Extreme Degradation Bench A benchmark for vocal restoration under extreme degradation. See ./noisy for the raw, recorded audio files, ./predictions for all existing predictions, and ./pairwise-ranking.csv for all existing raw pairwise ranking data sourced in https://arxiv.org/abs/2510.21659. See ./app.py for the Gradio application used for the pairwise rankings. Directly running app.py should automatically create a ./ratings directory for logging per-session vote information. Please… See the full description on the dataset page: https://huggingface.co/datasets/smulelabs/ExtremeDegradationBench.audioaudio-to-audion<1K1 likes487 downloads11mo agoHugging Face04xbgoose /dusha_extra_data Dataset Card for "dusha_extra_data" More Information needed audio100K<n<1M0 likes325 downloads3y agoHugging Face05SpeechPPL /SALMon_Flow-SLM-1B-Extended SALMon Normalized Dataset This repo preserves the SALMon per-config folder layout while normalizing mismatched schema details across model families. audio1K<n<10K0 likes314 downloads5mo agoHugging Face06extraordinarylab /torgoaudio10K<n<100K1 likes313 downloads8mo agoHugging Face07proxectonos /CRPIH_UVigo-GL-Voices_extended CRPIH_UVigo-GL-Voices: Galician TTS dataset CRPIH_UVigo-GL-Voices is a Galician TTS multi-speaker dataset containing audio recordings from four different speakers (two female and two male voices). The characteristics of each voice are detailed in the table below: Voice name Gender Speaker Recording # Utts Duration Iago Male Amateur Radio studio 1,316 1h 13min Icía Female Amateur Semi-professional studio 2,950 4h 5min Paulo Male Amateur Radio studio 1,316 1h 15min… See the full description on the dataset page: https://huggingface.co/datasets/proxectonos/CRPIH_UVigo-GL-Voices_extended.audiotext-to-speech1 likes273 downloads4mo agoHugging Face08yigagilbert /synthetic-parallel-external Synthetic Parallel EN↔LG — external Voice-controlled synthetic parallel speech dataset for Luganda-English speech-to-speech translation, generated by the Hibiki-Zero fine-tuning pipeline. Generation Component Model Translation Sunbird/translate-nllb-3.3b-salt TTS Sunbird/orpheus-3b-tts-multilingual English speakers: salt_eng_0001, salt_eng_0002, salt_eng_0003 Luganda speakers: salt_lug_0001, waxal_lug_0001, waxal_lug_0002, waxal_lug_0003, waxal_lug_0004… See the full description on the dataset page: https://huggingface.co/datasets/yigagilbert/synthetic-parallel-external.audioautomatic-speech-recognition100K<n<1M0 likes171 downloads4mo agoHugging Face09extraordinarylab /ua-speechaudio10K<n<100K2 likes128 downloads8mo agoHugging Face10labhamlet /STARSS23_extraaudio0 likes123 downloads8mo agoHugging Face11IqraEval /Iqra_Extra_IS26gatedaudio1K<n<10K3 likes122 downloads9mo agoHugging Face12lilgoose777 /tibetan-english-8s_extend_speechaudio1K<n<10K0 likes94 downloads8mo agoHugging Face13DynamicSuperb /EnvironmentalSoundClassification_ESC50-ExteriorAndUrbanNoises Dataset Card for "environmental_sound_classification_exterior_and_urban_noises_ESC50" More Information needed audion<1K2 likes62 downloads3y agoHugging Face14EvgenyShivchenkoUIT /haitian-creole-processed-extendeddaudio1K<n<10K0 likes60 downloads5mo agoHugging Face15dianavdavidson /indic_voices_only_extemporeaudio100K<n<1M0 likes56 downloads23d agoHugging Face16nativemind /mozgach_multimodal_extraaudion<1K0 likes53 downloads1y agoHugging Face17malaysia-ai /Speech-Instructions-Extraaudio100K<n<1M0 likes36 downloads2y agoHugging Face18octava /extracted-id-subbed-video-v3naudio10K<n<100K1 likes36 downloads2y agoHugging Face19humair025 /Urdu-TTS-test-3-extaudio1K<n<10K0 likes36 downloads1y agoHugging Face20SpeechPPL /SALMon_Flow-SLM-1B-Extended-depaudio1K<n<10K0 likes36 downloads11mo agoHugging Face21MUGEN-Benchmark /Speaking_Rate_Extremesaudion<1K0 likes35 downloads8mo agoHugging Face22theothertom /indian_english_extendedaudio1K<n<10K1 likes32 downloads2y agoHugging Face23DynamicSuperbPrivate /EnvironmentalSoundClassification_ESC50-ExteriorAndUrbanNoises_TTSaudion<1K1 likes27 downloads2y agoHugging Face24octava /extracted-untestedaudio1K<n<10K0 likes27 downloads2y agoHugging Face25octava /extracted-id-subbed-video-v2audio10K<n<100K1 likes22 downloads2y agoHugging Face26MUGEN-Benchmark /Duration_Extremes_Extractionaudion<1K0 likes21 downloads8mo agoHugging Face27erax /EraX-WoW-dataset-EXTRA-84k-1Mar2025audio10K<n<100K0 likes20 downloads2y agoHugging Face28OscarGD6 /audio-prompt-coco-balanced-extendedaudio1K<n<10K0 likes19 downloads1y agoHugging Face29octava /extracted-id-subbed-video-v4naudio1K<n<10K0 likes18 downloads2y agoHugging Face30SpeechTest /extreme_asr_ponyaudion<1K0 likes17 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.