CoolFace
Datasetpublic

ghanaopenai/ghana-english-speech-ipa

Ghanaian English Speech — Audio with IPA Transcripts Speech with both transcript forms: the original orthography and the IPA phoneme sequence read off the audio by ASR. Each language is a subset, with real train/validation splits. from datasets import load_dataset ds = load_dataset("ghanaopendata/ghana-english-speech-ipa", "English_eng", split="train") ds[0]["audio"] # decoded waveform, 16 kHz ds[0]["text"] # original orthography ds[0]["ipa"] # IPA phonemes 52,855… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-english-speech-ipa.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes732downloads
3 commits on main
439d03c2mo ago

Upload README.md with huggingface_hub

michsethowusu
50928ae2mo ago

Add English_eng

michsethowusu
2444bdd2mo ago

initial commit

michsethowusu