CoolFace
Datasetpublicgated

amphion/Emilia-NV

NVSpeech Dataset Overview The NVSpeech dataset provides extensive annotations of paralinguistic vocalizations for Mandarin Chinese speech, aimed at enhancing the capabilities of automatic speech recognition (ASR) and text-to-speech (TTS) systems. The dataset features explicit word-level annotations for 18 categories of paralinguistic vocalizations, including non-verbal sounds like laughter and breathing, as well as lexicalized interjections like "uhm" and "oh."… See the full description on the dataset page: https://huggingface.co/datasets/amphion/Emilia-NV.

sourceHugging Facecc-by-nc-sa-4.0updated 1y agoView on Hugging Face
52likes202downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.