CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RidheshBhati /Codemixed_New Codemixed ASR Dataset Unified collection of code-mixed ASR datasets. audioautomatic-speech-recognition100K<n<1M2 likes5.2k downloads5mo agoHugging Face02Codec-SUPERB /fluent_speech_commands_synth Dataset Card for "fluent_speech_commands_synth" More Information needed audio100K<n<1M1 likes1.4k downloads3y agoHugging Face03rogertseng /CodecFake CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems Paper, Code, Project Page Interspeech 2024 TL;DR: We show that better detection of deepfake speech from codec-based TTS systems can be achieved by training models on speech re-synthesized with neural audio codecs. This dataset is released for this purpose. See our paper and Github for more details on using our dataset. Acknowledgement… See the full description on the dataset page: https://huggingface.co/datasets/rogertseng/CodecFake.audio100K<n<1M5 likes1.3k downloads2y agoHugging Face04code-lover-ai /WSC-Evalaudio0 likes1.2k downloads3mo agoHugging Face05ajaykarthick /codecfake-audio Codecfake Dataset Overview The Codecfake dataset is a large-scale dataset designed for the detection of Audio Language Model (ALM)-based deepfake audio. This dataset includes millions of audio samples across two languages and various test conditions, tailored specifically for ALM-based audio detection. Conversion The original dataset was downloaded from Zenodo and converted to FLAC format to maintain audio quality while reducing file size. The dataset has been… See the full description on the dataset page: https://huggingface.co/datasets/ajaykarthick/codecfake-audio.audioaudio-classification100K<n<1M1 likes977 downloads2y agoHugging Face06Codec-SUPERB /librispeech_synth Dataset Card for "librispeech_synth" More Information needed audio1M<n<10M1 likes936 downloads3y agoHugging Face07Perle-ai /ASR_Code_Switch ASR Code-Switching Benchmark A curated benchmark of 1,200 code-switching utterances (300 per language pair) for evaluating commercial ASR systems on multilingual speech with intra-sentential language switching. Paper Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German arXiv link Language pairs Split Language pair Samples Scripts egyptian_arabic_english Egyptian Arabic–English 300 Arabic + Latin… See the full description on the dataset page: https://huggingface.co/datasets/Perle-ai/ASR_Code_Switch.audioautomatic-speech-recognition1K<n<10K12 likes752 downloads4mo agoHugging Face08besimple-ai /voice-code-bench VoiceCodeBench VoiceCodeBench is a test-only benchmark for evaluating whether automatic speech recognition (ASR) systems preserve exact structured values in English workplace speech. Paper: VoiceCodeBench: Evaluating Exact Structured-Token Recovery in Automatic Speech Recognition The benchmark targets cases where a transcript is software input: callback numbers, email addresses, command-line flags, file paths, URLs, account identifiers, dates, measurements, and similar values… See the full description on the dataset page: https://huggingface.co/datasets/besimple-ai/voice-code-bench.audioautomatic-speech-recognitionn<1K13 likes729 downloads10d agoHugging Face09Codec-SUPERB /voxceleb1_synthaudio10K<n<100K3 likes606 downloads3y agoHugging Face10Codec-SUPERB /vocal_imitation_synth Dataset Card for "vocal_imitation_synth" More Information needed audio10K<n<100K1 likes584 downloads3y agoHugging Face11Codec-SUPERB /maestro_synth Dataset Card for "maestro_synth" More Information needed audio1K<n<10K0 likes548 downloads3y agoHugging Face12Codec-SUPERB /crema_d_synth Dataset Card for "crema_d_synth" More Information needed audio100K<n<1M0 likes544 downloads3y agoHugging Face13Codec-SUPERB /vocalset_synth Dataset Card for "vocalset_synth" More Information needed audio10K<n<100K0 likes486 downloads3y agoHugging Face14CodecSR /librispeech_asr_test_48k_synthaudio100K<n<1M0 likes478 downloads2y agoHugging Face15CodecSR /vocalset_synthaudio10K<n<100K0 likes473 downloads3y agoHugging Face16CodecSR /vox_lingua_top10_synthaudio10K<n<100K0 likes469 downloads3y agoHugging Face17CodecSR /torgo_synthaudio100K<n<1M0 likes461 downloads2y agoHugging Face18CodecSR /speech_accent_archive_synthaudio10K<n<100K0 likes454 downloads2y agoHugging Face19Codec-SUPERB /opensinger_synthaudio10K<n<100K0 likes435 downloads3y agoHugging Face20CodecSR /librispeech_asr_test_synthaudio100K<n<1M0 likes428 downloads3y agoHugging Face21CodecSR /fluent_speech_commands_femaleaudio10K<n<100K1 likes412 downloads2y agoHugging Face22CodecSR /voxceleb1_synthaudio100K<n<1M0 likes373 downloads3y agoHugging Face23Codec-SUPERB /Nsynth-test_synthaudio10K<n<100K0 likes357 downloads3y agoHugging Face24Codec-SUPERB /noisy_vctk_16k_synth Dataset Card for "noisy_vctk_16k_synth" More Information needed audio100K<n<1M0 likes352 downloads3y agoHugging Face25Codec-SUPERB /quesst14_all_synth Dataset Card for "quesst14_all_synth" More Information needed audio100K<n<1M0 likes329 downloads3y agoHugging Face26CodecSR /easycall_synthaudio100K<n<1M0 likes316 downloads2y agoHugging Face27Codec-SUPERB /quesst_synth Dataset Card for "quesst_synth" More Information needed audio100K<n<1M0 likes314 downloads3y agoHugging Face28CodecSR /vox_lingua_top10_16k_synthaudio10K<n<100K0 likes314 downloads3y agoHugging Face29Codec-SUPERB /audioset_synth Dataset Card for "audioset_synth" More Information needed audio100K<n<1M0 likes307 downloads3y agoHugging Face30CodecSR /esc50_synthaudio10K<n<100K0 likes306 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.