CoolFace
Datasetpublic

Scicom-intl/YouTube-Cantonese-Emilia

YouTube Cantonese — Emilia 2,064,679 speaker-homogeneous Cantonese speech segments — 5,312.6 hours — produced by running alvanlii/cantonese-youtube through the Emilia speech-data pipeline (source separation → diarization → VAD segmentation → ASR → MOS filtering). Each row is one clean, single-speaker segment of 3–30 s with a transcript, a speaker turn label and a DNSMOS quality score. Audio is shipped separately as MP3s inside zip parts, in both an original and a silence-trimmed… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/YouTube-Cantonese-Emilia.

sourceHugging Faceupdated 1mo agoView on Hugging Face
1likes419downloads
settings

This repository belongs to Scicom-intl on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameYouTube-Cantonese-Emilia
visibilitypublic
licencenot set
gatedno
ownerScicom-intl
Account settings
Scicom-intl/YouTube-Cantonese-Emilia · CoolFace