CoolFace
Modelpublic

Fatoumataa/mt5-bambara-resumer-final

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes14downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

mt5-bambara-resumer-final

This model is a fine-tuned version of Fatoumataa/mt5-bambara-phase0-pro on the None dataset. It achieves the following results on the evaluation set:

  • Loss: 3.3116
  • Rouge1: 0.509
  • Rouge2: 0.2569
  • Rougel: 0.354
  • Rougelsum: 0.3542

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • learning_rate: 1e-05
  • trainbatchsize: 2
  • evalbatchsize: 2
  • seed: 42
  • gradientaccumulationsteps: 8
  • totaltrainbatch_size: 16
  • optimizer: Use OptimizerNames.ADAFACTOR and the args are: No additional optimizer arguments
  • lrschedulertype: constant
  • num_epochs: 8
  • labelsmoothingfactor: 0.1

Training results

Training LossEpochStepValidation LossRouge1Rouge2RougelRougelsum
35.30401.011283.83560.45620.21510.30570.3056
32.92912.022563.69690.47560.23040.32370.3239
31.61703.033843.58950.48490.23720.33140.3314
30.45124.045123.49880.48630.23880.33430.3344
29.92365.056403.44310.49150.24430.34020.3403
29.37646.067683.39600.50220.25130.34590.3459
29.02247.078963.35210.50720.25530.35170.3517
28.68978.090243.31160.5090.25690.3540.3542

Framework versions

  • Transformers 5.0.0
  • Pytorch 2.9.0+cu128
  • Datasets 4.0.0
  • Tokenizers 0.22.2