pronunciation
wav2vec2-xls-r-300m-Korean-children-pronunciation-jamo-based-semi-supervised_V2lexide-pronunciationwav2vec2-large-xlsr-53-english-pronunciation-evaluation-aod-cut-balancelexide-pronunciation-mergedwav2vec2-finetuned-pronunciation-correctionwhisper-small-korean-pronunciation-scorer-sampledataw2v2_pronunciation_score_modelpronunciation_accuracy
portuguese-unified-pronunciation-lexicon
Portuguese Unified Pronunciation Lexicon
A flat, single-row-per-pronunciation dataset merging Portuguese IPA transcriptions from three authoritative sources. Each row is a word × region × POS tuple with both broad phonemic (ipa_broad) and narrow phonetic (ipa_narrow) transcriptions normalized across sources.
Source
Words
Convention
Description
Infopédia (Porto Editora)
102,685
Broad phonemic
European Portuguese dictionary IPA
Wiktionary (pt.wiktionary.org)
15,720… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/portuguese-unified-pronunciation-lexicon.english-pronunciation-audiowiktionary_pronunciations-finalPronunciation-dictionary-malayalam
Malayalam Pronunciation Dictionary
This Dataset has an alternate name of Malayalam Phonetic Lexicon. It is curated from the original source here
It gives Phonemic transcription of Malayalam words in IPA format.
Dataset Details
Dataset Description
This is a collection of Malayalam words and their pronunciation described in IPA format. The pronunciations has been automatically generated using [Mlphon]
(https://pypi.org/project/mlphon/) Python library.
Curated… See the full description on the dataset page: https://huggingface.co/datasets/kavyamanohar/Pronunciation-dictionary-malayalam.wiktionary_pronunciations-backupPronunciation-boldvoice
Pronunciation Assessment Dataset (BoldVoice + speechocean762)
Dataset for fine-tuning multimodal models on English pronunciation assessment.
Overview
Source
Samples
Audio Duration
Description
BoldVoice
38,182
10-20s
Non-native English learners, BoldVoice API annotations
speechocean762
5,000
1.6-20s
Public dataset, 5-expert scored, Mandarin speakers
Total
43,182
Schema
Column
Type
Description
audio
Audio (16kHz mono)
Speech… See the full description on the dataset page: https://huggingface.co/datasets/aigc-x/Pronunciation-boldvoice.
