Lindo20/lwazi-asr-corpus-compressed
Lwazi ASR Corpus Collection This repository contains a curated collection of the Lwazi Automatic Speech Recognition (ASR) Corpus for several low-resourced South African languages. These datasets are designed for use in speech recognition research and development, particularly for underrepresented languages. Corpus Overview Each corpus consists of scripted telephonic speech recordings collected from native speakers, along with corresponding transcriptions. The… See the full description on the dataset page: https://huggingface.co/datasets/Lindo20/lwazi-asr-corpus-compressed.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face