CoolFace
Modelpublic

HouraMor/wh-loraft-lr5e6-dtstf5-adm-ga1ba16-st15k-v2-evalstp50-pat20-trainvalch

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
0likes11downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

wh-loraft-lr5e6-dtstf5-adm-ga1ba16-st15k-v2-evalstp50-pat20-trainvalch

This model is a fine-tuned version of HouraMor/wh-ft-lr5e6-dtstf5-adm-ga1ba16-st15k-v2-evalstp500-pat5 on an unknown dataset. It achieves the following results on the evaluation set:

  • —Loss: 0.5748

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 5e-06
  • —trainbatchsize: 16
  • —evalbatchsize: 8
  • —seed: 42
  • —optimizer: Use adamwtorch with betas=(0.9,0.999) and epsilon=1e-08 and optimizerargs=No additional optimizer arguments
  • —lrschedulertype: linear
  • —lrschedulerwarmup_steps: 250
  • —training_steps: 5000
  • —mixedprecisiontraining: Native AMP

Training results

Training LossEpochStepValidation Loss
0.30460.1004500.5668
0.34360.20081000.5666
0.36020.30121500.5670
0.29370.40162000.5672
0.30960.50202500.5671
0.26710.60243000.5678
0.34610.70283500.5690
0.22730.80324000.5701
0.27730.90364500.5710
0.44841.00405000.5714
0.20351.10445500.5712
0.18351.20486000.5720
0.2781.30526500.5729
0.29291.40567000.5741
0.32851.50607500.5742
0.31621.60648000.5747
0.26851.70688500.5748

Framework versions

  • —PEFT 0.15.2
  • —Transformers 4.52.3
  • —Pytorch 2.7.0+cu118
  • —Datasets 3.6.0
  • —Tokenizers 0.21.1