CoolFace
Modelpublic

ishmamzarif/augmented_normal_v2_extra_dataset_bangla-whisper-epoch-6

sourceHugging Facemitupdated 10mo agoView on Hugging Face
0likes12downloads
Model Card

Model Details

Bangla dialect ASR (Automatic Speech Recognition) Model

Training Procedure

Bangla ASR Model trained for the contest https://www.kaggle.com/competitions/shobdotori The original dataset was augmented and then trained with an extra dataset

Model Sources:

  • —Repository: https://huggingface.co/zarifmahir21/finetuned-modelv6
Training Hyperparameters
  • —Batch Size: 4
  • —Learning rate: 2e-5
  • —Warmup Steps: 200
  • —Epochs: 8
  • —Gradient Accumulation Steps: 4
  • —LR Scheduler: cosinewithrestarts

Evaluation

EpochTraining LossValidation LossWERNormalized Levenshtein Similarity
01.4806001.4565146.94037194.997286
21.4458001.4328584.83871096.562330
41.4285001.4285822.93255197.937398
61.4220001.4251592.98142797.792654

Language and Environment

  • —Python version: 3.12.12
  • —PyTorch version: 2.8.0+cu126
  • —Transformers version: 4.44.2
  • —NumPy version: 1.26.4
  • —

CUDA Info

  • —GPU: Tesla T4
  • —GPU Memory: 15.83 GB