deepdml/whisper-large-v3-turbo-ta-mix-norm
020
<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->
Whisper Turbo ta
This model is a fine-tuned version of openai/whisper-large-v3-turbo on the Common Voice 17.0 dataset. It achieves the following results on the evaluation set:
- Loss: 0.1246
- Wer: 28.7424
- Cer: 4.7569
Model description
More information needed
Intended uses & limitations
More information needed
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
- learning_rate: 1e-05
- trainbatchsize: 16
- evalbatchsize: 16
- seed: 42
- optimizer: Use adamwtorch with betas=(0.9,0.999) and epsilon=1e-08 and optimizerargs=No additional optimizer arguments
- lrschedulertype: linear
- lrschedulerwarmup_ratio: 0.04
- training_steps: 18000
Training results
Framework versions
- Transformers 4.48.0.dev0
- Pytorch 2.5.1+cu121
- Datasets 3.6.0
- Tokenizers 0.21.0
Citation
Please cite the model using the following BibTeX entry:
@misc{deepdml/whisper-large-v3-turbo-ta-mix-norm,
title={Fine-tuned Whisper turbo ASR model for speech recognition in Tamil},
author={Jimenez, David},
howpublished={\url{https://huggingface.co/deepdml/whisper-large-v3-turbo-ta-mix-norm}},
year={2026}
}