CoolFace
Modelpublic

arjunpatel/distilgpt2-finetuned-pokemon-moves

sourceHugging Faceapache-2.0updated 4y agoView on Hugging Face
0likes27downloads
Model Card

<!-- This model card has been generated automatically according to the information Keras had access to. You should probably proofread and complete it, then remove this comment. -->

arjunpatel/distilgpt2-finetuned-pokemon-moves

This model is a fine-tuned version of distilgpt2 on an unknown dataset. It achieves the following results on the evaluation set:

  • —Train Loss: 1.8709
  • —Validation Loss: 2.3512
  • —Epoch: 14

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —optimizer: {'name': 'AdamWeightDecay', 'learningrate': 2e-05, 'decay': 0.0, 'beta1': 0.9, 'beta2': 0.999, 'epsilon': 1e-07, 'amsgrad': False, 'weightdecay_rate': 0.01}
  • —training_precision: float32

Training results

Train LossValidation LossEpoch
3.71463.22880
3.11592.89611
2.85922.73882
2.66842.64233
2.53582.57094
2.43302.51375
2.33082.47366
2.24992.44447
2.18432.41158
2.13222.39319
2.06832.382910
2.01222.366911
1.96762.359612
1.90872.359113
1.87092.351214

Framework versions

  • —Transformers 4.18.0
  • —TensorFlow 2.8.0
  • —Datasets 2.1.0
  • —Tokenizers 0.11.0