CoolFace
Modelpublic

afrideva/refusal-GGUF

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes429downloads
Model Card

refusal-GGUF

Quantized GGUF model files for refusal from mrfakename

Original Model Card:

I messed up on the previous model. This is a fixed version.

A tiny 1B model that refuses basically anything you ask it! Trained on the refusal dataset. Prompt format is ChatML.

Training results:

Training LossEpochStepValidation Loss
2.43520.058012.4462
1.57410.521791.4304
1.52041.0435181.3701
1.07941.5217271.3505
1.12752.0435361.3344
0.66522.5217451.4360
0.62483.0435541.4313
0.61423.5072631.4934

Training hyperparemeters:

The following hyperparameters were used during training:

  • learning_rate: 0.0002
  • trainbatchsize: 2
  • evalbatchsize: 2
  • seed: 42
  • gradientaccumulationsteps: 4
  • totaltrainbatch_size: 8
  • optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • lrschedulertype: cosine
  • lrschedulerwarmup_steps: 10
  • num_epochs: 4

Base model: https://huggingface.co/TinyLlama/TinyLlama-1.1B-intermediate-step-1431k-3T