datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
asr-ckb
ASR-CKB Dataset Card
Dataset Summary
The ASR-CKB dataset is a comprehensive collection of audio recordings and their corresponding transcriptions in Central Kurdish (Sorani). It is designed to facilitate research and development in automatic speech recognition (ASR) for the Central Kurdish language.
Dataset Structure
Features
audio: Audio recordings with a sampling rate of 16,000 Hz.
sentence: Textual transcriptions of the audio recordings.… See the full description on the dataset page: https://huggingface.co/datasets/PawanKrd/asr-ckb.asr-ckb-v2
Dataset Summary
The ASR-CKB-V2 dataset is a comprehensive collection of audio recordings and their corresponding transcriptions in Central Kurdish (Sorani). It is designed to facilitate research and development in automatic speech recognition (ASR) for the Central Kurdish language.
Dataset Structure
Features
audio: Audio recordings with a sampling rate of 16,000 Hz.
sentence: Textual transcriptions of the audio recordings.
Splits
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/PawanKrd/asr-ckb-v2.
