CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01dsfsi-anv /za-african-next-voices-compressedgatedNote: This dataset is a compressed version of za-african-next-voices. It was compressed to .opus format using a 32k bitrate. Swivuriso: ZA-African Next Voices-Compressed Swivuriso is a large-scale multilingual speech dataset targeting over 3000 hours of audio across 7 South African languages. The dataset is developed to support Automatic Speech Recognition (ASR) and inclusive speech technologies for low-resource African languages. It combines both scripted and unscripted speech… See the full description on the dataset page: https://huggingface.co/datasets/dsfsi-anv/za-african-next-voices-compressed.audioautomatic-speech-recognition100K<n<1M1 likes100 downloads8mo agoHugging Face02dsfsi /lwazi-asr-corpus-compressed Lwazi ASR Corpus Collection This repository contains a curated collection of the Lwazi Automatic Speech Recognition (ASR) Corpus for several low-resourced South African languages. These datasets are designed for use in speech recognition research and development, particularly for underrepresented languages. Corpus Overview Each corpus consists of scripted telephonic speech recordings collected from native speakers, along with corresponding transcriptions. The… See the full description on the dataset page: https://huggingface.co/datasets/dsfsi/lwazi-asr-corpus-compressed.automatic-speech-recognition2 likes30 downloads1y agoHugging Face03Lindo20 /lwazi-asr-corpus-compressed Lwazi ASR Corpus Collection This repository contains a curated collection of the Lwazi Automatic Speech Recognition (ASR) Corpus for several low-resourced South African languages. These datasets are designed for use in speech recognition research and development, particularly for underrepresented languages. Corpus Overview Each corpus consists of scripted telephonic speech recordings collected from native speakers, along with corresponding transcriptions. The… See the full description on the dataset page: https://huggingface.co/datasets/Lindo20/lwazi-asr-corpus-compressed.automatic-speech-recognition0 likes9 downloads7mo agoHugging Face04rareRabbit /lwazi-asr-corpus-compressed Lwazi ASR Corpus Collection This repository contains a curated collection of the Lwazi Automatic Speech Recognition (ASR) Corpus for several low-resourced South African languages. These datasets are designed for use in speech recognition research and development, particularly for underrepresented languages. Corpus Overview Each corpus consists of scripted telephonic speech recordings collected from native speakers, along with corresponding transcriptions. The… See the full description on the dataset page: https://huggingface.co/datasets/rareRabbit/lwazi-asr-corpus-compressed.automatic-speech-recognition0 likes7 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.