CoolFace
Modelpublic

RichardErkhov/Heejindo_-_feedback_model_e10_save5000-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes1.7kdownloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

feedbackmodele10_save5000 - GGUF

  • Model creator: https://huggingface.co/Heejindo/
  • Original model: https://huggingface.co/Heejindo/feedbackmodele10_save5000/

Original model description: --- libraryname: transformers license: llama3.2 basemodel: meta-llama/Llama-3.2-1B tags:

  • trl
  • sft
  • generatedfromtrainer model-index:
  • name: feedbackmodele10_save5000 results: [] ---

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

feedbackmodele10_save5000

This model is a fine-tuned version of meta-llama/Llama-3.2-1B on an unknown dataset. It achieves the following results on the evaluation set:

  • Loss: 2.0187

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • learning_rate: 1e-05
  • trainbatchsize: 4
  • evalbatchsize: 4
  • seed: 42
  • optimizer: Use adamwtorch with betas=(0.9,0.999) and epsilon=1e-08 and optimizerargs=No additional optimizer arguments
  • lrschedulertype: linear
  • num_epochs: 10

Training results

Training LossEpochStepValidation Loss
1.04960.476950002.0187
0.34480.9538100002.5590
0.20271.4308150002.8900
0.17461.9077200002.9743
0.14882.3846250003.1316
0.13952.8615300003.2616
0.11893.3384350003.3299
0.1163.8153400003.4260
0.09444.2923450003.5361
0.09374.7692500003.6240
0.07315.2461550003.7380
0.07375.7230600003.8218
0.05496.1999650003.9090
0.05566.6768700004.0614
0.04447.1538750004.1715
0.04297.6307800004.2847
0.03698.1076850004.4924
0.03648.5845900004.5050
0.0349.0614950004.6572
0.03459.53831000004.6768

Framework versions

  • Transformers 4.46.3
  • Pytorch 2.3.0
  • Datasets 2.14.4
  • Tokenizers 0.20.3