CoolFace
Datasetpublic

mirfan899/phoneme_asr

This dataset contains the phonetic transcriptions of audios as well as English transcripts. Phonetic transcriptions are based on the g2p model. It can be used to train phoneme recognition model using wav2vec2.

sourceHugging Facebsdupdated 3y agoView on Hugging Face
4likes28downloads
Dataset Card

This dataset contains the phonetic transcriptions of audios as well as English transcripts. Phonetic transcriptions are based on the g2p model. It can be used to train phoneme recognition model using wav2vec2.