nonstandard
en_whisper_nonstandard_mediumwhisper-large-v3_finetuned_rwandan_english_nonstandard_speech_v1.0whisper-large-v3_finetuned_kinyarwanda_nonstandard_speech_v1.0whisper-small_finetuned_kenyan_english_nonstandard_speech_v1.0whisper-large-v3_finetuned_kenyan_swahili_nonstandard_speech_v1.0whisper-small-kenyan-english-nonstandardwhisper-large-v3_finetuned_ugandan_english_nonstandard_speech_v1.0sw_whisper_nonstandard_medium
kenyan_swahili_nonstandard_speech_v1.0This dataset provides 32.5 hours of Swahili speech recordings (5,535 samples) from 52 Kenyan speakers living with speech impairments. The participants represent a diversity of progressive, acquired and congenital aetiologies, including cerebral palsy, Parkinson's disease, multiple sclerosis, autism spectrum disorder, Down syndrome, stroke and stuttering.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker or… See the full description on the dataset page: https://huggingface.co/datasets/cdli/kenyan_swahili_nonstandard_speech_v1.0.rwandan_kinyarwanda_nonstandard_speech_v1.0This dataset provides 61.7 hours of Kinyarwanda speech recordings (14,739 samples) from 61 Rwandan speakers living with speech impairments. The participants represent a limited diversity of speech patterns, mostly stuttering, and a few examples of Dysarthria, Dysphonia, and Phonological disorders.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker or phrase level. All speech recordings of this datasets have been… See the full description on the dataset page: https://huggingface.co/datasets/cdli/rwandan_kinyarwanda_nonstandard_speech_v1.0.ugandan_english_nonstandard_speech_v1.0This dataset provides 42.4 hours of Ugandan English speech recordings (7,251 samples) from 59 Ugandan speakers living with speech impairments. The participants represent a diversity of progressive, acquired and congenital aetiologies, including cerebral palsy, Parkinson's disease, multiple sclerosis, autism spectrum disorder, Down syndrome, stroke and stuttering.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker… See the full description on the dataset page: https://huggingface.co/datasets/cdli/ugandan_english_nonstandard_speech_v1.0.rwandan_english_nonstandard_speech_v1.0This dataset provides 32.7 hours of English speech recordings (7,592 samples) from 44 Rwandan speakers living with speech impairments. The participants represent a limited diversity of speech patterns, mostly stuttering, and a few examples of Dysarthria, Dysphonia, and Phonological disorders.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker or phrase level. All speech recordings of this datasets have been… See the full description on the dataset page: https://huggingface.co/datasets/cdli/rwandan_english_nonstandard_speech_v1.0.kenyan_english_nonstandard_speech_v1.0This dataset provides 32.3 hours of English speech recordings (5,998 samples) from 52 Kenyan speakers living with speech impairments. The participants represent a diversity of progressive, acquired and congenital aetiologies, including cerebral palsy, Parkinson's disease, multiple sclerosis, autism spectrum disorder, Down syndrome, stroke and stuttering.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker or… See the full description on the dataset page: https://huggingface.co/datasets/cdli/kenyan_english_nonstandard_speech_v1.0.ghanian_ga_nonstandard_speech_v1.0This dataset provides 6.28 hours of Ga nonstandard speech recordings (12,160 samples) from 21 Ga speakers living with speech impairments. The participants represent a diversity of progressive, acquired and congenital aetiologies, including cerebral palsy, Parkinson's disease, multiple sclerosis, autism spectrum disorder, Down syndrome, stroke and stuttering.
This dataset includes a split into a training, test and development set. The splits were created avoiding any overlap on the speaker or… See the full description on the dataset page: https://huggingface.co/datasets/cdli/ghanian_ga_nonstandard_speech_v1.0.
