Scicom-intl/Evaluation-Multilingual-VC
Evaluation-Multilingual-VC We use dataset https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon, Filter languages that support by Whisper Large V3 to evaluate WER automatically, Only take test set, sort by up votes. Because VC required to source text, source audio, target text, we make sure the target text is not same as source text, target text we take from other rows. Only build first 500 rows for each language Github issue at… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/Evaluation-Multilingual-VC.
Evaluation-Multilingual-VC
We use dataset https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon,
- Filter languages that support by Whisper Large V3 to evaluate WER automatically,
- Only take test set, sort by up votes.
- Because VC required to source text, source audio, target text, we make sure the target text is not same as source text, target text we take from other rows.
- Only build first 500 rows for each language
Github issue at https://github.com/Scicom-AI-Enterprise-Organization/Multilingual-TTS/issues/4
