CoolFace
Modelpublic

sethilakshay/wav2vec2-base-finetune-flickr8k

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
0likes5downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

wav2vec2-base-finetune-flickr8k

This model is a fine-tuned version of facebook/wav2vec2-base-960h on the None dataset. It achieves the following results on the evaluation set:

  • Loss: 3.0580
  • Wer: 1.0

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • learning_rate: 0.0001
  • trainbatchsize: 12
  • evalbatchsize: 2
  • seed: 42
  • optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • lrschedulertype: linear
  • lrschedulerwarmup_steps: 500
  • num_epochs: 15
  • mixedprecisiontraining: Native AMP

Training results

Training LossEpochStepValidation LossWer
3.22520.85003.07871.0
3.19581.610003.06981.0
3.15122.415003.05711.0
3.14113.220003.05781.0
3.11864.025003.05531.0
3.08524.830003.06161.0
3.10635.635003.06211.0
3.06786.440003.05501.0
3.02397.245003.05921.0
3.03218.050003.06521.0
3.03968.855003.05421.0
3.07239.660003.05521.0
3.056710.465003.05451.0
3.043811.270003.06111.0
3.07312.075003.05481.0
3.029312.880003.05851.0
3.044413.685003.05711.0
3.021314.490003.05801.0

Framework versions

  • Transformers 4.27.4
  • Pytorch 2.0.0
  • Datasets 2.11.0
  • Tokenizers 0.13.2