CoolFace
Modelpublic

langtest/falcon-7b-sharded-bf16-finetuned-mental-health-hf-plus-dsm5mistral

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes3downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

falcon-7b-sharded-bf16-finetuned-mental-health-hf-plus-dsm5mistral

This model is a fine-tuned version of ybelkada/falcon-7b-sharded-bf16 on the None dataset. It achieves the following results on the evaluation set:

  • —Loss: 1.6730

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 0.0002
  • —trainbatchsize: 16
  • —evalbatchsize: 16
  • —seed: 42
  • —gradientaccumulationsteps: 4
  • —totaltrainbatch_size: 64
  • —optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • —lrschedulertype: cosine
  • —lrschedulerwarmup_ratio: 0.03
  • —training_steps: 200

Training results

Training LossEpochStepValidation Loss
1.65740.1003101.7478
1.45190.2005201.7755
1.58230.3008301.7614
1.56330.4010401.7620
1.38160.5013502.0733
1.750.6015601.7332
1.34080.7018701.7491
1.64360.8020801.7197
1.4390.9023901.7320
1.37551.00251001.7057
1.57511.10281101.7190
1.16491.20301201.7603
1.4811.30331301.6892
1.27691.40351401.6977
1.25641.50381501.7133
1.44861.60401601.6711
1.08681.70431701.6710
1.48921.80451801.6721
1.19031.90481901.6727
1.15772.00502001.6730

Framework versions

  • —PEFT 0.13.1.dev0
  • —Transformers 4.45.1
  • —Pytorch 2.4.1+cu121
  • —Datasets 3.0.1
  • —Tokenizers 0.20.0