k2speech/FeruzaSpeech
Dataset Card for Dataset Name FeruzaSpeech is a read speech dataset of the Uzbek language, transcribed in both Cyrillic and Latin alphabets, freely available for academic research purposes. It includes 60 hours of high-quality recordings from a single native female speaker from Tashkent, Uzbekistan. ICNLSPConference: https://www.youtube.com/watch?v=9whj9yzI_s4&ab_channel=ICNLSPConference Paper: https://arxiv.org/abs/2410.00035 Example test.tsv: audio text_latin… See the full description on the dataset page: https://huggingface.co/datasets/k2speech/FeruzaSpeech.
5382
No card is published for this repository, or it could not be fetched from Hugging Face right now.
