CoolFace
9 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01rogertseng /CodecFake CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems Paper, Code, Project Page Interspeech 2024 TL;DR: We show that better detection of deepfake speech from codec-based TTS systems can be achieved by training models on speech re-synthesized with neural audio codecs. This dataset is released for this purpose. See our paper and Github for more details on using our dataset. Acknowledgement… See the full description on the dataset page: https://huggingface.co/datasets/rogertseng/CodecFake.audio100K<n<1M5 likes1k downloads2y agoHugging Face02ajaykarthick /codecfake-audio Codecfake Dataset Overview The Codecfake dataset is a large-scale dataset designed for the detection of Audio Language Model (ALM)-based deepfake audio. This dataset includes millions of audio samples across two languages and various test conditions, tailored specifically for ALM-based audio detection. Conversion The original dataset was downloaded from Zenodo and converted to FLAC format to maintain audio quality while reducing file size. The dataset has been… See the full description on the dataset page: https://huggingface.co/datasets/ajaykarthick/codecfake-audio.audioaudio-classification100K<n<1M1 likes735 downloads2y agoHugging Face03CodecFake /CodecFake_Plus_Dataset Codecfake+ Dataset Overview This is the official dataset repository for CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset.It stores and provides access to the dataset (CoRS and CoSG), including audio samples and accompanying protocol/label files. News [2025.10] — CoRS and CoSG dataset and corresponding label files have been uploaded. [2025.09] — Released public audio samples of the CoRS subset. Download We provide two… See the full description on the dataset page: https://huggingface.co/datasets/CodecFake/CodecFake_Plus_Dataset.text1M<n<10M4 likes293 downloads8mo agoHugging Face04rogertseng /CodecFake_wavs CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems Paper, Code, Project Page Interspeech 2024 TL;DR: We show that better detection of deepfake speech from codec-based TTS systems can be achieved by training models on speech re-synthesized with neural audio codecs. This dataset is released for this purpose. See our paper and Github for more details on using our dataset. Acknowledgement… See the full description on the dataset page: https://huggingface.co/datasets/rogertseng/CodecFake_wavs.audio0 likes137 downloads1y agoHugging Face05CodecFake /CodecFakeaudio100K<n<1M1 likes84 downloads2y agoHugging Face06singleman /Codecfake0 likes67 downloads1y agoHugging Face07ggirishg /Expressive_CodecFake Expressive CodecFake Expressive CodecFake is a codec-fake expressive speech dataset for audio deepfake detection research. The dataset contains codec-generated expressive speech and nonverbal vocalization samples organized into verified TAR shards. Current Dataset Structure Expressive_CodecFake/ ├── Verbal speech CF/ │ ├── emodb_2.0_CF/ │ │ ├── emodb_2.0_CF-0000.tar │ │ └── ... │ ├── EMOVO_CF/ │ │ ├── EMOVO_CF-0000.tar │ │ └── ... │ └──… See the full description on the dataset page: https://huggingface.co/datasets/ggirishg/Expressive_CodecFake.audioaudio-classification0 likes22 downloads3mo agoHugging Face08DarwinKxg /Union-Codecfake-Dataset0 likes2 downloads1y agoHugging Face09Milind392 /SEA-Codecfake-Datasetaudion<1K0 likes2 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.