expressive
emilia-expressive-zh
Emilia Expressive — Phase-1 Filtered Subset
Auto-generated by emilia_pipeline.scoring.phase1_hf. This is the Phase-1
filtered view: every clip that survived the S0+S1 acoustic funnel, physically
partitioned into quality tiers so you can download exactly the strictness
level you want -- before Phase-2 emotion labeling.
Derived from amphion/Emilia-Dataset
(CC-BY-NC-4.0); the same license and usage restrictions apply.
Pipeline version: voxsift-emilia-v1.3-full (schema 1.3)
Clips:… See the full description on the dataset page: https://huggingface.co/datasets/leeoxiang/emilia-expressive-zh.ExpressiveSpeech
ExpressiveSpeech Dataset
Project Webpage
中文版 (Chinese Version)
About The Dataset
ExpressiveSpeech is a high-quality, expressive, and bilingual (Chinese-English) speech dataset created to address the common lack of consistent vocal expressiveness in existing dialogue datasets.
This dataset is meticulously curated from five renowned open-source emotional dialogue datasets: Expresso, NCSSD, M3ED, MultiDialog, and IEMOCAP. Through a rigorous processing and selection pipeline… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/ExpressiveSpeech.ExpressiveSpeech
ExpressiveSpeech
Expressive Speech dataset,
Default, we build our own by combining multiple classifier models and use LLM to generate synthetic description, https://github.com/Scicom-AI-Enterprise-Organization/Multilingual-TTS/issues/2
gigaspeech, from https://speechcraft2024.github.io/speechcraft2024/
libritts_r, from https://speechcraft2024.github.io/speechcraft2024/
Data source for Default
You can follow… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/ExpressiveSpeech.SALMon_Spirit-LM-Expressive-depPrahaTTS-ML-Expressive-DatasetSALMon_Spirit-LM-Expressive
SALMon Normalized Dataset
This repo preserves the SALMon per-config folder layout while normalizing
mismatched schema details across model families.
