CoolFace
Datasetpublic

mirfan899/phoneme_asr

This dataset contains the phonetic transcriptions of audios as well as English transcripts. Phonetic transcriptions are based on the g2p model. It can be used to train phoneme recognition model using wav2vec2.

sourceHugging Facebsdupdated 3y agoView on Hugging Face
4likes30downloads

mirfan899/phoneme_asr · main · files are served by the source, never re-hosted here