CoolFace
Modelpublic

soikit/chinese-bert-wwm-chinese_bert_wwm3

sourceHugging Faceapache-2.0updated 5y agoView on Hugging Face
0likes105downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

chinese-bert-wwm-chinesebertwwm3

This model is a fine-tuned version of hfl/chinese-bert-wwm on the None dataset. It achieves the following results on the evaluation set:

  • —Loss: 0.0000

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 2e-05
  • —trainbatchsize: 64
  • —evalbatchsize: 64
  • —seed: 42
  • —optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • —lrschedulertype: linear
  • —num_epochs: 30.0

Training results

Training LossEpochStepValidation Loss
No log1.0720.4251
No log2.01440.0282
No log3.02160.0048
No log4.02880.0018
No log5.03600.0011
No log6.04320.0006
0.4837.05040.0004
0.4838.05760.0004
0.4839.06480.0002
0.48310.07200.0002
0.48311.07920.0002
0.48312.08640.0001
0.48313.09360.0001
0.003114.010080.0001
0.003115.010800.0001
0.003116.011520.0001
0.003117.012240.0001
0.003118.012960.0001
0.003119.013680.0001
0.003120.014400.0001
0.001521.015120.0001
0.001522.015840.0001
0.001523.016560.0001
0.001524.017280.0001
0.001525.018000.0000
0.001526.018720.0001
0.001527.019440.0000
0.00128.020160.0000
0.00129.020880.0000
0.00130.021600.0000

Framework versions

  • —Transformers 4.11.3
  • —Pytorch 1.9.1
  • —Datasets 1.13.3
  • —Tokenizers 0.10.3