aigc-x/Pronunciation-boldvoice
Pronunciation Assessment Dataset (BoldVoice + speechocean762) Dataset for fine-tuning multimodal models on English pronunciation assessment. Overview Source Samples Audio Duration Description BoldVoice 38,182 10-20s Non-native English learners, BoldVoice API annotations speechocean762 5,000 1.6-20s Public dataset, 5-expert scored, Mandarin speakers Total 43,182 Schema Column Type Description audio Audio (16kHz mono)… See the full description on the dataset page: https://huggingface.co/datasets/aigc-x/Pronunciation-boldvoice.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face