CoolFace
Modelpublic

saattrupdan/wav2vec2-xls-r-300m-ftspeech

sourceHugging Faceotherupdated 3y agoView on Hugging Face
0likes918kdownloads
Model Card

XLS-R-300m-FTSpeech

Model description

This model is a fine-tuned version of facebook/wav2vec2-xls-r-300m on the FTSpeech dataset, being a dataset of 1,800 hours of transcribed speeches from the Danish parliament.

Performance

The model achieves the following WER scores (lower is better):

**Dataset****WER without LM****WER with 5-gram LM**
Danish part of Common Voice 8.020.4817.91
Alvenir test set15.4613.84

License

The use of this model needs to adhere to this license from the Danish Parliament.