k2speech/FeruzaSpeech
Dataset Card for Dataset Name FeruzaSpeech is a read speech dataset of the Uzbek language, transcribed in both Cyrillic and Latin alphabets, freely available for academic research purposes. It includes 60 hours of high-quality recordings from a single native female speaker from Tashkent, Uzbekistan. ICNLSPConference: https://www.youtube.com/watch?v=9whj9yzI_s4&ab_channel=ICNLSPConference Paper: https://arxiv.org/abs/2410.00035 Example test.tsv: audio text_latin… See the full description on the dataset page: https://huggingface.co/datasets/k2speech/FeruzaSpeech.
This repository belongs to k2speech on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
