CoolFace
Modelpublic

TARARARAK/HGU_rulebook-Llama3.2-Bllossom-5B_fine-tuning-QLoRA-32_64_3

sourceHugging Facellama3.2updated 1y agoView on Hugging Face
0likes3downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

HGUrulebook-Llama3.2-Bllossom-5Bfine-tuning-QLoRA-32643

This model is a fine-tuned version of Bllossom/llama-3.2-Korean-Bllossom-AICA-5B on an unknown dataset. It achieves the following results on the evaluation set:

  • —Loss: 5.6952

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 1e-05
  • —trainbatchsize: 2
  • —evalbatchsize: 2
  • —seed: 42
  • —gradientaccumulationsteps: 8
  • —totaltrainbatch_size: 16
  • —optimizer: Use adamwbnb8bit with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
  • —lrschedulertype: cosinewithrestarts
  • —lrschedulerwarmup_ratio: 0.1
  • —training_steps: 942
  • —mixedprecisiontraining: Native AMP

Training results

Training LossEpochStepValidation Loss
12.4590.29914711.8428
8.95270.5982948.3608
6.84380.89741416.6652
6.12721.19651886.0176
5.8581.49562355.8316
5.77071.79472825.7618
5.74352.09393295.7331
5.72152.39303765.7194
5.71182.69214235.7114
5.70472.99124705.7063
5.7033.29045175.7027
5.69733.58955645.7001
5.69593.88866115.6985
5.69214.18776585.6972
5.69684.48697055.6965
5.69544.78607525.6959
5.69585.08517995.6955
5.69375.38428465.6953
5.69155.68348935.6952
5.69515.98259405.6952

Framework versions

  • —PEFT 0.12.0
  • —Transformers 4.46.2
  • —Pytorch 2.0.1+cu118
  • —Datasets 3.0.0
  • —Tokenizers 0.20.1