CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01FreedomIntelligence /ExpressiveSpeech ExpressiveSpeech Dataset Project Webpage 中文版 (Chinese Version) About The Dataset ExpressiveSpeech is a high-quality, expressive, and bilingual (Chinese-English) speech dataset created to address the common lack of consistent vocal expressiveness in existing dialogue datasets. This dataset is meticulously curated from five renowned open-source emotional dialogue datasets: Expresso, NCSSD, M3ED, MultiDialog, and IEMOCAP. Through a rigorous processing and selection pipeline… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/ExpressiveSpeech.audio10K<n<100K11 likes422 downloads11mo agoHugging Face02Scicom-intl /ExpressiveSpeech ExpressiveSpeech Expressive Speech dataset, Default, we build our own by combining multiple classifier models and use LLM to generate synthetic description, https://github.com/Scicom-AI-Enterprise-Organization/Multilingual-TTS/issues/2 gigaspeech, from https://speechcraft2024.github.io/speechcraft2024/ libritts_r, from https://speechcraft2024.github.io/speechcraft2024/ Data source for Default You can follow… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/ExpressiveSpeech.tabular1M<n<10M1 likes157 downloads7mo agoHugging Face03SpeechPPL /SALMon_Spirit-LM-Expressive SALMon Normalized Dataset This repo preserves the SALMon per-config folder layout while normalizing mismatched schema details across model families. audio1K<n<10K0 likes59 downloads6mo agoHugging Face04Aynursusuz /tts-en-zonos2-expressive ZONOS2 Expressive — English voice cloning English Mandarin-pipeline counterpart: a diverse reference voice is generated with Qwen3-TTS VoiceDesign, then cloned with Zyphra/ZONOS2 in expressive mode (accurate_mode=false). Conversational, human-sounding texts (some emotional, mixed lengths); deliberately not anime/cartoon-style voices. Content is verified with Qwen/Qwen3-ASR-1.7B: every clone is transcribed and compared to its target text (asr_wer). Columns… See the full description on the dataset page: https://huggingface.co/datasets/Aynursusuz/tts-en-zonos2-expressive.audiotext-to-speechn<1K0 likes22 downloads3mo agoHugging Face05Aynursusuz /tts-zh-zonos2-expressive ZONOS2 — Accurate vs Expressive (Mandarin voice cloning) Side-by-side A/B comparison of Zyphra/ZONOS2 accurate mode (accurate_mode=true) vs expressive mode (accurate_mode=false). Same reference voice and same target text per row, cloned twice — one per mode — so each can be heard back to back. Reference voices are clean Qwen3 generations. Columns column meaning index row id ref_text text of the reference voice ref_audio reference voice (cloning… See the full description on the dataset page: https://huggingface.co/datasets/Aynursusuz/tts-zh-zonos2-expressive.audiotext-to-speechn<1K0 likes19 downloads3mo agoHugging Face06r-labs /expressive-eng-tts r-labs/expressive-eng-tts Expressive synthetic Ugandan English speech dataset for conversational Text-to-Speech (TTS) fine-tuning. Dataset Summary r-labs/expressive-eng-tts is a fully synthetic expressive Ugandan English TTS dataset designed for fine-tuning conversational speech models with authentic Ugandan English accent, prosody, and expressive speaking behaviors. The dataset contains speech generated from 3 synthetic speakers: 2 Female speakers 1 Male… See the full description on the dataset page: https://huggingface.co/datasets/r-labs/expressive-eng-tts.audiotext-to-speechn<1K0 likes8 downloads4mo agoHugging Face07rayzox57 /Youtube_Expressive_Artiststabularn<1K0 likes4 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.