CoolFace
Modelpublic

sammy786/wav2vec2-xlsr-georgian

sourceHugging Faceapache-2.0updated 5y agoView on Hugging Face
1likes79downloads
Model Card

sammy786/wav2vec2-xlsr-georgian

This model is a fine-tuned version of facebook/wav2vec2-xls-r-1b on the MOZILLA-FOUNDATION/COMMONVOICE8_0 - ka dataset. It achieves the following results on evaluation set (which is 10 percent of train data set merged with other and dev datasets):

  • Loss: 10.54
  • Wer: 27.53

Model description

"facebook/wav2vec2-xls-r-1b" was finetuned.

Intended uses & limitations

More information needed

Training and evaluation data

Training data - Common voice Finnish train.tsv, dev.tsv and other.tsv

Training procedure

For creating the train dataset, all possible datasets were appended and 90-10 split was used.

Training hyperparameters

The following hyperparameters were used during training:

  • learning_rate: 0.000045637994662983496
  • trainbatchsize: 8
  • evalbatchsize: 16
  • seed: 13
  • gradientaccumulationsteps: 4
  • totaltrainbatch_size: 32
  • optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • lrschedulertype: cosinewithrestarts
  • lrschedulerwarmup_steps: 500
  • num_epochs: 30
  • mixedprecisiontraining: Native AMP

Training results

StepTraining LossValidation LossWer
2004.1521000.8236720.967814
4000.8895000.1967400.444792
6000.4937000.1556590.366115
8000.3280000.1380660.358069
10000.2606000.1192360.324989
12000.2172000.1140500.313366
14000.1888000.1126000.302190
16000.1669000.1111540.295485
18000.1555000.1099630.286544
20000.1404000.1075870.277604
22000.1426000.1056620.277157
24000.1354000.1054140.275369

Framework versions

  • Transformers 4.16.0.dev0
  • Pytorch 1.10.0+cu102
  • Datasets 1.17.1.dev0
  • Tokenizers 0.10.3
Evaluation Commands
  1. 1.To evaluate on mozilla-foundation/common_voice_8_0 with split test
bash
python eval.py --model_id sammy786/wav2vec2-xlsr-georgian --dataset mozilla-foundation/common_voice_8_0 --config ka --split test