CoolFace
20 results

shona

realtime-speech /shona1audio10K<n<100K1 likes407 downloads2y agoHugging Facemanassehzw /shona-bible-bdsc-aligned Shona Bible Speech Alignment Dataset Lossless, verse-aligned Shona Bible speech dataset derived from the BDSC source audio made available by Biblica, Inc. through Open.Bible. This release contains the complete Bible: 66 books, 1,189 chapters, and 31,284 speech segments covering approximately 75.55 hours. Dataset summary Language: Shona (sna) Speaker: narrator 1 Speaker sex: male Books: 66 Clips: 31,284 Audio: approximately 75.55 hours Audio format: mono 16 kHz… See the full description on the dataset page: https://huggingface.co/datasets/manassehzw/shona-bible-bdsc-aligned.audioautomatic-speech-recognition10K<n<100K3 likes267 downloads17d agoHugging Facemanassehzw /shona-waxal-pseudo-labeled Shona WAXAL pseudo-labelled speech This release contains 90,253 Shona speech clips, totalling 441.585 hours. Each clip keeps its original FLAC audio and a Sunbird Whisper pseudo-transcript. These are model outputs, not human reference transcriptions. What this release contains The source is the unlabeled Shona ASR split from WAXAL NLP, preserved in the operational checkpoint manassehzw/sna-waxal-annotated-unlabeled. The source checkpoint has no transcripts. This… See the full description on the dataset page: https://huggingface.co/datasets/manassehzw/shona-waxal-pseudo-labeled.audioautomatic-speech-recognition10K<n<100K3 likes190 downloads24d agoHugging Facebadrex /shona-speechaudio10K<n<100K8 likes175 downloads11mo agoHugging FacetakuM23 /shona-vibevoice-corpus Shona VibeVoice Corpus Unified 24 kHz mono Shona speech corpus in the VibeVoice segment schema, with a mix of single- and multi-utterance rows (short clips concatenated with silence gaps and multi-segment labels). Compiled by build_and_push.py from: realtime-speech/shona1 google/WaxalNLP (sna ASR) google/fleurs (sn_zw) teeofftechnologies/badrex-shona-whisper-cleaned-16k (train + test) Kittech/mixed_shona_dataset realtime-speech/shona2 (train + test + validation)… See the full description on the dataset page: https://huggingface.co/datasets/takuM23/shona-vibevoice-corpus.audioautomatic-speech-recognition10K<n<100K0 likes135 downloads3mo agoHugging FaceKittech /mixed_shona_datasetaudiotext-generationn<1K4 likes83 downloads2y agoHugging Face