CoolFace
Datasetpublic

michsethowusu/swahili-words-speech-text-parallel

Swahili Words Speech-Text Parallel Dataset Dataset Description This dataset contains 411048 parallel speech-text pairs for Swahili, a widely spoken language in East Africa. The dataset consists of audio recordings paired with corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Dataset Summary Language: Swahili - sw Task: Speech Recognition, Text-to-Speech Size: 411048 audio… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/swahili-words-speech-text-parallel.

sourceHugging Facecc-by-4.0updated 1y agoView on Hugging Face
1likes209downloads

michsethowusu/swahili-words-speech-text-parallel · main · files are served by the source, never re-hosted here