CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01XRXRX /X-Voice-Dataset-Train X-Voice Training Dataset Overview The X-Voice training dataset is a large-scale multilingual speech corpus curated for high-performance speech models. It provides a robust foundation for cross-lingual phonetic and prosodic modeling. Also the train set of X-Voice Model. Core Statistics Total Speech Duration: 420K hours 30 languages European: bg (Bulgarian), cs (Czech), da (Danish), de (German), el (Greek), en (English), es (Spanish), et (Estonian), fi… See the full description on the dataset page: https://huggingface.co/datasets/XRXRX/X-Voice-Dataset-Train.audiotext-to-speech10M<n<100M11 likes4.5k downloads5mo agoHugging Face02mispeech /xares_llm_dataaudioaudio-classification1M<n<10M4 likes356 downloads11mo agoHugging Face03shahink /xvoxaudio1M<n<10M0 likes248 downloads2y agoHugging Face04XRXRX /X-Voice-TestsetX-Voice Multilingual Test Set High-Fidelity Test Set for Multilingual Text-to-Speech across 30 Languages This test set is built as part of the research: X-Voice: One Speaker, 30+ Languages with Zero-Shot Voice Cloning, serving as the evaluation benchmark for our model. Dataset Summary 30 languages European: bg (Bulgarian), cs (Czech), da (Danish), de (German), el (Greek), en (English), es (Spanish), et (Estonian), fi (Finnish), fr (French), hr (Croatian), hu (Hungarian), it… See the full description on the dataset page: https://huggingface.co/datasets/XRXRX/X-Voice-Testset.audiotext-to-speech10K<n<100K4 likes60 downloads5mo agoHugging Face05dalietng /dataset_asv_x2audio100K<n<1M0 likes16 downloads8mo agoHugging Face06clatter-1 /XF-DenoiseFar-field speech enhancement plays a crucial role in speech signal processing, primarily aimed at improving speech intelligibility and other speech-based applications in real-world scenarios. However, models trained on simulated data or existing real-world far-field datasets often exhibit limited performance in extreme far-field conditions. To address these challenges, we present XF-Denoise, a real-world dataset specifically designed for extreme far-field speech enhancement. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/clatter-1/XF-Denoise.audio100K<n<1M1 likes13 downloads1y agoHugging Face07xiaheyan /hualonghua-webdatasetaudio1K<n<10K0 likes1 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.