CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01amphion /Emilia-Datasetgated Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation This is the official repository 👑 for the Emilia dataset and the source code for the Emilia-Pipe speech data preprocessing pipeline. News 🔥 2025/02/26: The Emilia-Large dataset, featuring over 200,000 hours of data, is now available!!! Emilia-Large combines the original 101k-hour Emilia dataset (licensed under CC BY-NC 4.0) with the brand-new 114k-hour Emilia-YODAS… See the full description on the dataset page: https://huggingface.co/datasets/amphion/Emilia-Dataset.audiotext-to-speech10M<n<100M489 likes46k downloads2y agoHugging Face02Vyvo-Research /Emilia-YODAS-ENaudio10M<n<100M4 likes4.3k downloads11mo agoHugging Face03TTS-AGI /emilia-yodasA mirror of the Emilia-YODAS dataset. Only includes the YODAS subset from the original dataset. https://huggingface.co/datasets/amphion/Emilia-Dataset audiotext-to-speech10M<n<100M5 likes3.2k downloads2y agoHugging Face04mesolitica /Malaysian-Emilia-annotated Malaysian Emilia Annotated Annotate Malaysian-Emilia using Data-Speech pipeline. Malaysian Youtube Originally from malaysia-ai/crawl-youtube Total 3168.8 hours. Gender prediction, filtered-24k_processed_24k_gender.zip Language prediction, filtered-24k_processed_language.zip Force alignment. Post cleaned to 24k and 44k sampling rates, 24k, filtered-24k_processed_24k.zip 44k, filtered-24k_processed_44k.zip Synthetic description… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Malaysian-Emilia-annotated.tabulartext-to-speech1M<n<10M2 likes2.1k downloads1y agoHugging Face05Vyvo-Research /Emilia-ENaudio10M<n<100M2 likes1.6k downloads11mo agoHugging Face06laion /Emilia-with-Emotion-Annotations4audio10M<n<100M1 likes894 downloads1y agoHugging Face07abbuibuibui /emilia-Dataset Emilia (Re:Zero) — Illustrious SDXL Character LoRA Training Dataset 中文说明 | English (current) This is the actual training subset used to train abbuibuibui/emilia-Lora, an unofficial Emilia character LoRA on waiIllustriousSDXL v17. The files here are a copy of the directory named in dataset.toml (image_dir = .../03_captioned/main). Nothing was added from unused candidates, and captions were not rewritten for this release. This is a fan-made derivative, not an official product.… See the full description on the dataset page: https://huggingface.co/datasets/abbuibuibui/emilia-Dataset.imagetext-to-imagen<1K0 likes867 downloads12d agoHugging Face08duplexio /emilia-yodas-en-speaker-embeddings Emilia-YODAS English Qwen3-TTS Speaker Embeddings This dataset contains precomputed speaker embeddings for the English subset of Emilia-YODAS. Each row maps an Emilia-YODAS sample ID to one speaker embedding extracted from the corresponding audio. Dataset Details Source dataset: amphion/Emilia-Dataset Source subset: Emilia-YODAS English Embedding model: Qwen/Qwen3-TTS-12Hz-1.7B-Base Embedding shape: (2048,) Embedding dtype: float16 Rows: 4,516,833 Split: train Additional… See the full description on the dataset page: https://huggingface.co/datasets/duplexio/emilia-yodas-en-speaker-embeddings.textfeature-extraction1M<n<10M0 likes866 downloads4mo agoHugging Face09laion /Emilia-with-Emotion-Annotations5audio10M<n<100M3 likes738 downloads1y agoHugging Face10Scicom-intl /YouTube-Cantonese-Emilia YouTube Cantonese — Emilia 2,064,679 speaker-homogeneous Cantonese speech segments — 5,312.6 hours — produced by running alvanlii/cantonese-youtube through the Emilia speech-data pipeline (source separation → diarization → VAD segmentation → ASR → MOS filtering). Each row is one clean, single-speaker segment of 3–30 s with a transcript, a speaker turn label and a DNSMOS quality score. Audio is shipped separately as MP3s inside zip parts, in both an original and a silence-trimmed… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/YouTube-Cantonese-Emilia.tabularautomatic-speech-recognition1M<n<10M1 likes731 downloads1mo agoHugging Face11AdrienB134 /Emilia-dataset-french-with-gender Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/AdrienB134/Emilia-dataset-french-with-gender.audioautomatic-speech-recognition100K<n<1M1 likes639 downloads2y agoHugging Face12Scicom-intl /Malaysian-Emilia Malaysian Emilia Gather Malaysian Emilia from, https://huggingface.co/datasets/mesolitica/Malaysian-Emilia-v2 https://huggingface.co/datasets/Scicom-intl/Malaysian-Chinese-Emilia https://huggingface.co/datasets/mesolitica/Malaysian-Emilia#malaysian-dialect And do, Trim silent. Permutation for Voice Conversion include post-filtering during permutation. Convert to Neucodec speech tokens. tabular10M<n<100M1 likes636 downloads8mo agoHugging Face13Scicom-intl /Malaysian-Chinese-Emilia Malaysian-Chinese-Emilia Use https://github.com/mesolitica/Emilia to pseudo-label Malaysian Chinese audio. Total rows: 605169 Total hours: 1857.611445057867 hours Permutation for Voice Conversion Also we already calculated speaker permutation to prepare for voice conversion. tabular10M<n<100M1 likes604 downloads8mo agoHugging Face14nytopop /emilia-en-snac Stats (EN) Emilia: 46,349 hours Emilia-YODAS: 87,258 hours Total: 133,607 hours License The Emilia subset is licensed under CC BY-NC 4.0. The Emilia-YODAS subset is licensed under CC BY 4.0. Reference @inproceedings{emilialarge, author={He, Haorui and Shang, Zengqiang and Wang, Chaoren and Li, Xuyuan and Gu, Yicheng and Hua, Hua and Liu, Liwei and Yang, Chen and Li, Jiaqi and Shi, Peiyang and Wang, Yuancheng and Chen, Kai and Zhang, Pengyuan and Wu… See the full description on the dataset page: https://huggingface.co/datasets/nytopop/emilia-en-snac.texttext-to-speech100M<n<1B2 likes597 downloads1y agoHugging Face15Vyvo /Emilia-DE Emilia - DE Clean version with only text and audio from the Emilia Dataset. Samples: 653,109 Language: DE Usage from datasets import load_dataset dataset = load_dataset("Vyvo/Emilia-DE") sample = dataset['train'][0] text = sample['text'] audio = sample['audio']['array'] sampling_rate = sample['audio']['sampling_rate'] audio100K<n<1M0 likes520 downloads1y agoHugging Face16mrfakename /emilia-yodas-fr-parquetaudio1M<n<10M0 likes445 downloads1y agoHugging Face17AdrienB134 /Emilia-dataset-french-splitaudio100K<n<1M4 likes432 downloads2y agoHugging Face18Scicom-intl /Emilia-YODAS-Voice-Conversion Emilia-YODAS-Voice-Conversion We sample https://huggingface.co/datasets/amphion/Emilia-Dataset YODAS set for voice conversion. Filter transcriptions based on character repetitiveness and word ngrams. Filter speaker similarity using https://huggingface.co/nvidia/speakerverification_en_titanet_large during speaker permutation. Convert audio to speech tokens using https://huggingface.co/neuphonic/neucodec We also upload the full permutation as zip files. Speech Tokenizer… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/Emilia-YODAS-Voice-Conversion.audio10M<n<100M4 likes408 downloads8mo agoHugging Face19theodorr /emilia_hifitts_fulltext10M<n<100M0 likes408 downloads6mo agoHugging Face20laion /Emilia-with-Emotion-Annotations3audio10M<n<100M1 likes382 downloads1y agoHugging Face21laion /Emilia-Annotated-WIPStill a WIP, full dataset is still being annotated audio1M<n<10M3 likes368 downloads1y agoHugging Face22jspaulsen /emilia-yodas-en-mimitabular10M<n<100M0 likes337 downloads1y agoHugging Face23laion /Emilia-with-Emotion-Annotations2audio10M<n<100M1 likes316 downloads1y agoHugging Face24mesolitica /Malaysian-Emilia-v2 Malaysian Emilia v2 This version 2 should fixed https://github.com/open-mmlab/Amphion/issues/436, an Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Malaysian and Singaporean Speech Generation. Replicating Emilia on, Dataset Clone and Extract We upload as split zip files so you can clone and extract distributedly, huggingface-cli download --repo-type dataset \ --include '*.zip' \ --local-dir './' \ --max-workers 20 \… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Malaysian-Emilia-v2.tabular1M<n<10M2 likes315 downloads1y agoHugging Face25Vyvo /Emilia-YODAS-DE Emilia-YODAS - DE Clean version with only text and audio from the Emilia Dataset. Samples: 2,005,364 Language: DE Usage from datasets import load_dataset dataset = load_dataset("Vyvo/Emilia-YODAS-DE") sample = dataset['train'][0] text = sample['text'] audio = sample['audio']['array'] sampling_rate = sample['audio']['sampling_rate'] audio1M<n<10M1 likes302 downloads1y agoHugging Face26humairawan /emilia_mfa_correctaudio1M<n<10M1 likes296 downloads9mo agoHugging Face27mesolitica /Extra-Emilia Extra Emilia Extra dataset to extend Tamil and Mandarin capability for Malaysian-Emilia. Tamil Total length is 891 hours. Mandarin Total length is 301 hours. text1M<n<10M0 likes295 downloads1y agoHugging Face28AstraMindAI /Emilia-R-POSTPROCESS-44d9db35text100K<n<1M0 likes280 downloads1y agoHugging Face29amphion /Emilia-NVgated NVSpeech Dataset Overview The NVSpeech dataset provides extensive annotations of paralinguistic vocalizations for Mandarin Chinese speech, aimed at enhancing the capabilities of automatic speech recognition (ASR) and text-to-speech (TTS) systems. The dataset features explicit word-level annotations for 18 categories of paralinguistic vocalizations, including non-verbal sounds like laughter and breathing, as well as lexicalized interjections like "uhm" and "oh."… See the full description on the dataset page: https://huggingface.co/datasets/amphion/Emilia-NV.audiotext-to-speech100K<n<1M52 likes269 downloads1y agoHugging Face30vvwangvv /emilia-captions-v3text10M<n<100M0 likes252 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.