Chat-Error/Rose-Kimiko-20B
28
<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->
qlora-out
This model is a fine-tuned version of tavtav/Rose-20B on the Kimiko dataset.
Model description
The prompt formats used is ShareGPT/Vicuna format.
Intended uses & limitations
Per many people requests, this LoRA is intended to fix spelling from Rose 20B.
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
- learning_rate: 0.0002
- trainbatchsize: 2
- evalbatchsize: 2
- seed: 42
- gradientaccumulationsteps: 4
- totaltrainbatch_size: 8
- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
- lrschedulertype: cosine
- lrschedulerwarmup_steps: 10
- num_epochs: 2
Training results
Framework versions
- Transformers 4.36.0.dev0
- Pytorch 2.0.1+cu118
- Datasets 2.15.0
- Tokenizers 0.15.0
Training procedure
The following bitsandbytes quantization config was used during training:
- quant_method: bitsandbytes
- loadin8bit: False
- loadin4bit: True
- llmint8threshold: 6.0
- llmint8skip_modules: None
- llmint8enablefp32cpu_offload: False
- llmint8hasfp16weight: False
- bnb4bitquant_type: nf4
- bnb4bitusedoublequant: True
- bnb4bitcompute_dtype: bfloat16
Framework versions
- PEFT 0.6.0
