CoolFace
Modelpublic

toanduz/PRIMERA-multinews-lora-finetuned

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes4downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

PRIMERA-multinews-lora-finetuned

This model is a fine-tuned version of allenai/PRIMERA-multinews on an unknown dataset. It achieves the following results on the evaluation set:

  • —Loss: 1.6767
  • —Rouge1: 13.1661
  • —Rouge2: 6.075
  • —Rougel: 11.1948
  • —Rougelsum: 12.1382
  • —Gen Len: 20.0

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 2e-05
  • —trainbatchsize: 2
  • —evalbatchsize: 2
  • —seed: 42
  • —optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • —lrschedulertype: linear
  • —num_epochs: 16
  • —mixedprecisiontraining: Native AMP

Training results

Training LossEpochStepValidation LossRouge1Rouge2RougelRougelsumGen Len
2.23851.047251.918315.83087.603912.662414.186620.0
2.14192.094501.893314.85456.889312.14413.462320.0
2.12863.0141751.861916.25858.143113.422614.865320.0
2.06694.0189001.812915.96247.529313.380914.776520.0
2.04485.0236251.763616.28018.04513.617814.99620.0
1.98316.0283501.703713.87355.995610.825112.254520.0
1.99267.0330751.762313.95915.886111.11212.434920.0
1.998.0378001.724713.14415.256510.711711.85120.0
1.94959.0425251.706512.48634.644410.015511.387420.0
1.978210.0472501.691911.83944.00689.455410.642120.0
1.908711.0519751.691013.0115.564410.725511.853220.0
1.969312.0567001.687213.26785.796611.053712.110320.0
1.944513.0614251.708413.27575.933711.08412.235420.0
1.946714.0661501.672912.92025.42410.531511.666120.0
1.958215.0708751.678613.28516.080611.29112.251820.0
1.918616.0756001.676713.16616.07511.194812.138220.0

Framework versions

  • —PEFT 0.12.0
  • —Transformers 4.43.2
  • —Pytorch 2.2.1+cu121
  • —Datasets 2.20.0
  • —Tokenizers 0.19.1