CoolFace
Modelpublic

khanhdhq/finetune_vietcuna_3b_qlora_e1_lr0.0002

sourceHugging Faceotherupdated 3y agoView on Hugging Face
0likes116downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

khanhdhq/finetunevietcuna3bqlorae1_lr0.0002

This model is a fine-tuned version of vilm/vietcuna-3b on an unknown dataset.

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

The following bitsandbytes quantization config was used during training:

  • —loadin8bit: False
  • —loadin4bit: True
  • —llmint8threshold: 6.0
  • —llmint8skip_modules: None
  • —llmint8enablefp32cpu_offload: False
  • —llmint8hasfp16weight: False
  • —bnb4bitquant_type: nf4
  • —bnb4bitusedoublequant: True
  • —bnb4bitcompute_dtype: bfloat16

Training hyperparameters

The following hyperparameters were used during training:

Training results

Framework versions

  • —PEFT 0.4.0.dev0
  • —Transformers 4.31.0.dev0
  • —Pytorch 2.0.1+cu118
  • —Datasets 2.13.0
  • —Tokenizers 0.13.3