oddadmix/arabic-audio-collection-sudanese-nuuar
Nuuar Sudanese Arabic Speech Dataset Dataset Summary The Nuuar Sudanese Arabic Speech Dataset is a single-speaker Sudanese Arabic speech corpus containing approximately 75 hours of speech recordings and corresponding transcripts. Sudanese Arabic remains one of the most underrepresented Arabic varieties in speech technology. This dataset directly addresses that gap by providing long-form, natural, dialectal Sudanese speech from a single consistent speaker, making… See the full description on the dataset page: https://huggingface.co/datasets/oddadmix/arabic-audio-collection-sudanese-nuuar.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face