CoolFace
Datasetpublic

myandev/rfa_shan_language_voices

RFA Shan Language Voices This dataset contains 20.58 hours of audio in the Shan (Tai-Yai) language, sourced from news broadcasts by Radio Free Asia (RFA) Burmese. This is one of the largest publicly accessible audio resources for the Shan language, designed to support research in low-resource automatic speech recognition (ASR), voice activity detection, and other speech-related tasks. The audio has been automatically segmented into 5,047 manageable chunks and prepared in the… See the full description on the dataset page: https://huggingface.co/datasets/myandev/rfa_shan_language_voices.

sourceHugging Faceotherupdated 9d agoView on Hugging Face
0likes42downloads

myandev/rfa_shan_language_voices · main · files are served by the source, never re-hosted here