CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01laion /laions_got_talent_enhanced_no_metadataaudio10K<n<100K0 likes1.1k downloads2y agoHugging Face02DynamicSuperbPrivate /EnhancementDetection_LibrittsTrainClean360Wham Dataset Card for "EnhancementDetection_LibrittsTrainClean360Wham" More Information needed audio100K<n<1M0 likes379 downloads3y agoHugging Face03ebellob /voxforge_spanish_enhanced VoxForge Spanish Enhanced (CleanUNet + FlashSR) Dataset Summary This dataset is a processed and enhanced version of the Spanish subset of: VoxForge. Furthermore, as this is a personal project, we give no guarantees that the audio is completely clean from any artifacts or noise the CleanUNet model could not remove. However, we have personally tested the corpus via the fine-tuning of some SOTA speech models and the results have been satisfactory. It has been created to… See the full description on the dataset page: https://huggingface.co/datasets/ebellob/voxforge_spanish_enhanced.audio10K<n<100K1 likes358 downloads5mo agoHugging Face04FILM6912 /th-en-zh-tts-200k-enhanced TH-EN-ZH Multi-speaker TTS Dataset (200K, RE-USE Enhanced) Speech-enhanced variant of FILM6912/th-en-zh-tts-200k. Every clip has been processed through NVIDIA RE-USE (universal speech enhancement, SEMamba) at its native sample rate, then re-encoded losslessly as FLAC (PCM_16). Same schema, same row order, same 200,000 rows (th 100k / en 50k / zh 50k): Column Type Description text string Transcript (identical to the original dataset) audio Audio Enhanced audio, FLAC… See the full description on the dataset page: https://huggingface.co/datasets/FILM6912/th-en-zh-tts-200k-enhanced.audiotext-to-speech100K<n<1M0 likes323 downloads39m agoHugging Face05TTS-AGI /mls-enhanced-dacvae Multilingual LibriSpeech converted to DAC VAE latents Source facebook/multilingual_librispeech Format Each tar shard (~2GB) contains samples with three files per sample: {sample_key}.audio.flac # Original audio (FLAC, original sample rate) {sample_key}.dacvae.npy # DAC VAE latent [T_latent, 128] numpy float32 {sample_key}.metadata.json # All metadata + duration_seconds + chars_per_second DAC VAE Latent Format Model:… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/mls-enhanced-dacvae.audioautomatic-speech-recognition100K<n<1M0 likes223 downloads6mo agoHugging Face06kyutai /librispeech_test_clean_enhancedaudion<1K1 likes165 downloads9mo agoHugging Face07Cnam-LMSSC /vibravox_enhanced_by_EBEN Dataset Card Description This dataset features a speech-enhanced version of the test split from the speech_clean subset of the Vibravox Dataset. It is not intended for training. Enhancement procedure The Bandwidth extension task has been individually achieved for each sensor using configurable EBEN (arXiv link) models available at https://huggingface.co/Cnam-LMSSC/vibravox_EBEN_models. Ressources Results for speech-to-phoneme and speaker… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox_enhanced_by_EBEN.audioaudio-to-audio1K<n<10K1 likes62 downloads2y agoHugging Face08ebellob /annotated_catalan_common_voice_v17_cleaned_enhanced Processed Annotated Catalan Common Voice v17 (CleanUNet + FlashSR) Dataset Summary This dataset is a processed and enhanced version of: projecte-aina/annotated_catalan_common_voice_v17. Furthermore, as this is a personal project, we give no guarantees that the audio is completely clean from any artifacts or noise the CleanUNet model could not remove. However, we have personally tested the corpus via the fine-tuning of some SOTA speech models and the results have been… See the full description on the dataset page: https://huggingface.co/datasets/ebellob/annotated_catalan_common_voice_v17_cleaned_enhanced.audiotext-to-speech100K<n<1M2 likes50 downloads5mo agoHugging Face09laion /reference-voices-enhanced Reference Voices Enhanced 2,004 AI voice samples enhanced with ClearerVoice-Studio MossFormer2_SE_48K speech enhancement, annotated with Empathic Insight Voice Plus (59 quality + emotion scores). Dataset Summary Source: laion/ai-voices-deduplicated (2,004 speaker-deduplicated, quality-filtered AI voice samples) Speech Enhancement: ClearerVoice MossFormer2_SE_48K — background noise removal and speech clarity improvement Output Format: Enhanced WAV files at 48kHz… See the full description on the dataset page: https://huggingface.co/datasets/laion/reference-voices-enhanced.audioaudio-classification1K<n<10K0 likes46 downloads6mo agoHugging Face10ebellob /voxpopuli_spanish_enhanced VoxPopuli Spanish Enhanced (CleanUNet + FlashSR) Dataset Summary This dataset is a processed and enhanced version of the Spanish subset of: facebook/voxpopuli. Furthermore, as this is a personal project, we give no guarantees that the audio is completely clean from any artifacts or noise the CleanUNet model could not remove. However, we have personally tested the corpus via the fine-tuning of some SOTA speech models and the results have been satisfactory. In addition to… See the full description on the dataset page: https://huggingface.co/datasets/ebellob/voxpopuli_spanish_enhanced.audio10K<n<100K1 likes31 downloads5mo agoHugging Face11Bohemian-self /MCE_Mixed_Cantonese_English_Speech_Enhancedaudio10K<n<100K0 likes29 downloads4d agoHugging Face12rashid0784 /common_voice_audio_quality_enhancement_v3audio100K<n<1M1 likes26 downloads2y agoHugging Face13Cnam-LMSSC /french-mrt_enhanced_by_EBENSame dataset as Cnam-LMSSC/french-mrt but enhanced by EBEN models audion<1K0 likes23 downloads2y agoHugging Face14thaint /vi-speech-enhancementaudio100K<n<1M0 likes20 downloads5mo agoHugging Face15rashid0784 /common_voice_audio_quality_enhancementaudio10K<n<100K0 likes19 downloads2y agoHugging Face16Thanarit /TH-Speech-Enhancedaudion<1K0 likes19 downloads2y agoHugging Face17Trelis /eval-ursa-2-enhanced-eka-hard-20260408-1924 Evaluation Results: ursa-2-enhanced Evaluation results from Whisper model evaluation. Summary Model WER CER speechmatics/ursa-2-enhanced 34.09% 23.66% Source Data Evaluation Dataset: Trelis/eka-hard Model Evaluated: speechmatics/ursa-2-enhanced Columns Column Description audio Audio sample (if available from source dataset) reference Ground truth transcription prediction Model prediction wer Word Error Rate for this… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-ursa-2-enhanced-eka-hard-20260408-1924.audion<1K0 likes19 downloads6mo agoHugging Face18Bluebomber182 /Mara-Jade-Resemble-Enhance-Versionaudion<1K0 likes17 downloads3y agoHugging Face19DynamicSuperbPrivate /EnhancementDetection_LibriTTS-TestClean_WHAM_TTSaudion<1K0 likes17 downloads2y agoHugging Face20hamza11111 /jalandhary_asr_enhancedaudio10K<n<100K0 likes16 downloads1y agoHugging Face21DynamicSuperb /EnhancementDetection_LibriTTS-TestClean_WHAM Dataset Card for "EnhancementDetection_LibrittsTestCleanWham" More Information needed audion<1K0 likes15 downloads3y agoHugging Face22macabdul9 /EnhancementDetection_LibriTTS-TestClean_WHAMaudion<1K0 likes13 downloads2y agoHugging Face23Wikidepia /openslr_enhancedaudio100K<n<1M0 likes10 downloads7mo agoHugging Face24Hawat /speech-enhancementaudio1K<n<10K0 likes7 downloads3y agoHugging Face25iFaz /enhanced_facebook_voxpopulik_16k_Whisper_Compatibleaudio1K<n<10K0 likes7 downloads2y agoHugging Face26Trelis /eval-ursa-2-enhanced-medical-terms-2025-20260408-1928 Evaluation Results: ursa-2-enhanced Evaluation results from Whisper model evaluation. Summary Model WER CER speechmatics/ursa-2-enhanced 6.04% 3.29% Source Data Evaluation Dataset: Trelis/medical-terms-2025 Model Evaluated: speechmatics/ursa-2-enhanced Columns Column Description audio Audio sample (if available from source dataset) reference Ground truth transcription prediction Model prediction wer Word Error Rate… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-ursa-2-enhanced-medical-terms-2025-20260408-1928.audion<1K0 likes7 downloads6mo agoHugging Face27Arsen2453 /reuse-enhancedaudion<1K0 likes7 downloads4mo agoHugging Face28laion /eurospeech-enhanced-dacvae EuroSpeech parliamentary speech converted to DAC VAE latents Source disco-eth/EuroSpeech Format Each tar shard (~2GB) contains samples with three files per sample: {sample_key}.audio.flac # Original audio (FLAC, original sample rate) {sample_key}.dacvae.npy # DAC VAE latent [T_latent, 128] numpy float32 {sample_key}.metadata.json # All metadata + duration_seconds + chars_per_second DAC VAE Latent Format Model:… See the full description on the dataset page: https://huggingface.co/datasets/laion/eurospeech-enhanced-dacvae.audioautomatic-speech-recognition1M<n<10M0 likes4 downloads5mo agoHugging Face29Trelis /eval-ursa-2-enhanced-multimed-hard-20260408-1933 Evaluation Results: ursa-2-enhanced Evaluation results from Whisper model evaluation. Summary Model WER CER speechmatics/ursa-2-enhanced 10.46% 6.03% Source Data Evaluation Dataset: Trelis/multimed-hard Model Evaluated: speechmatics/ursa-2-enhanced Columns Column Description audio Audio sample (if available from source dataset) reference Ground truth transcription prediction Model prediction wer Word Error Rate for… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-ursa-2-enhanced-multimed-hard-20260408-1933.audion<1K0 likes4 downloads6mo agoHugging Face30sleeping-ai /enhanced-vocal-burstaudion<1K0 likes3 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.