CoolFace
Datasetpublic

Scicom-intl/Evaluation-Multilingual-VC

Evaluation-Multilingual-VC We use dataset https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon, Filter languages that support by Whisper Large V3 to evaluate WER automatically, Only take test set, sort by up votes. Because VC required to source text, source audio, target text, we make sure the target text is not same as source text, target text we take from other rows. Only build first 500 rows for each language Github issue at… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/Evaluation-Multilingual-VC.

sourceHugging Faceupdated 6mo agoView on Hugging Face
0likes258downloads
Dataset Card

Evaluation-Multilingual-VC

We use dataset https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon,

  1. 1.Filter languages that support by Whisper Large V3 to evaluate WER automatically,
  2. 2.Only take test set, sort by up votes.
  3. 3.Because VC required to source text, source audio, target text, we make sure the target text is not same as source text, target text we take from other rows.
  4. 4.Only build first 500 rows for each language

Github issue at https://github.com/Scicom-AI-Enterprise-Organization/Multilingual-TTS/issues/4