CoolFace
Modelpublic

Marialab/finetuned-whisper-large-v3-turbo-1000-v2-step

sourceHugging Facemitupdated 2y agoView on Hugging Face
0likes4downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

Finetuned Whisper large-v3-turbo for darija speech translation

This model is a fine-tuned version of openai/whisper-large-v3-turbo on the Darija-C dataset. It achieves the following results on the evaluation set:

  • —Loss: 0.0004
  • —Bleu: 0.8080

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 1e-05
  • —trainbatchsize: 16
  • —evalbatchsize: 8
  • —seed: 42
  • —optimizer: Use adamwtorch with betas=(0.9,0.999) and epsilon=1e-08 and optimizerargs=No additional optimizer arguments
  • —lrschedulertype: linear
  • —lrschedulerwarmup_steps: 100
  • —training_steps: 1000
  • —mixedprecisiontraining: Native AMP

Training results

Training LossEpochStepValidation LossBleu
2.46882.2727502.10490.0694
0.84844.54551000.99470.1871
0.33586.81821500.25790.5827
0.13959.09092000.09360.6840
0.066911.36362500.03830.7778
0.042113.63643000.02140.7793
0.029315.90913500.01950.8053
0.022818.18184000.01020.8019
0.013220.45454500.00650.8014
0.01122.72735000.00500.8053
0.009425.05500.00250.8080
0.004827.27276000.00090.8080
0.00229.54556500.00070.8080
0.001131.81827000.00050.8080
0.000734.09097500.00050.8080
0.000636.36368000.00040.8080
0.000438.63648500.00040.8080
0.000340.90919000.00040.8080
0.000243.18189500.00040.8080
0.000245.454510000.00040.8080

Framework versions

  • —Transformers 4.47.1
  • —Pytorch 2.5.1+cu121
  • —Datasets 2.19.2
  • —Tokenizers 0.21.0