CoolFace
Datasetpublicgated

alvanlii/cantonese-youtube-tts

Cantonese Audio TTS Dataset This dataset contains alvanlii/cantonese-radio, alvanlii/cantonese-youtube, plus a dataset of equal size. It is catered towards TTS (text-to-speech) use cases, more than the 2 previously published datasets, as there is more extensive filtering and audio enhancement. For speaker labelling, you can use speaker embedding models like Nvidia's TitaNet Filtered out: Overlapped voices, detected using pyannote/speaker-diarization-3.1 Music, detected using a… See the full description on the dataset page: https://huggingface.co/datasets/alvanlii/cantonese-youtube-tts.

sourceHugging Faceupdated 6mo agoView on Hugging Face
3likes873downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.