CoolFace
Modelpublic

dhasmana/Kumaoni-Bhojpuri-Chhattisgarhi-Angika-Maithili-Magadhi-w2v-bert-2.0

sourceHugging Facemitupdated 8mo agoView on Hugging Face
0likes14downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

Kumaoni-Bhojpuri-Chhattisgarhi-Angika-Maithili-Magadhi-w2v-bert-2.0

This model is a fine-tuned version of facebook/w2v-bert-2.0 on the None dataset. It achieves the following results on the evaluation set:

  • —Loss: 2.1846
  • —Cer: 0.1794

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 5e-05
  • —trainbatchsize: 16
  • —evalbatchsize: 8
  • —seed: 42
  • —gradientaccumulationsteps: 2
  • —totaltrainbatch_size: 32
  • —optimizer: Use OptimizerNames.ADAMWTORCHFUSED with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
  • —lrschedulertype: linear
  • —lrschedulerwarmup_steps: 500
  • —num_epochs: 10
  • —mixedprecisiontraining: Native AMP

Training results

Training LossEpochStepValidation LossCer
5.28160.50933001.07180.2254
1.68651.01876000.87800.1970
1.26081.52809000.86680.1891
1.02942.037412000.85830.1882
0.80382.546715000.87480.1825
0.65993.056018001.02730.1829
0.47873.565421001.02210.1824
0.37974.074724001.18660.1818
0.27424.584027001.18580.1850
0.21745.093430001.33130.1824
0.15325.602733001.35050.1844
0.12416.112136001.50080.1840
0.09016.621439001.66860.1844
0.06667.130742001.77210.1844
0.05037.640145001.87640.1801
0.03758.149448001.94880.1823
0.02188.658751001.96710.1817
0.01739.168154002.09470.1812
0.00999.677457002.18460.1794

Framework versions

  • —Transformers 5.0.0
  • —Pytorch 2.9.0+cu126
  • —Datasets 4.0.0
  • —Tokenizers 0.22.2