CoolFace
Modelpublic

D1zzYzz/BOOLQ-QLORA-llama-3.2-3B

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
0likes9downloads
Model Card

meta-llama/Llama-3.2-3B Fine-tuned with QLora

This model is a fine-tuned version of meta-llama/Llama-3.2-3B using the LoRA on the google/boolq dataset.

๐Ÿš€ Training Details

Fine-tuning Configuration

  • โ€”Base Model: meta-llama/Llama-3.2-3B
  • โ€”Quantization: 4-bit compute.
  • โ€”LoRA Rank: 16
  • โ€”LoRA Alpha: 32
  • โ€”Batch Size: 8 (per device)
  • โ€”Gradient Accumulation: 4
  • โ€”Learning Rate: 2e-5
  • โ€”Sequence Length: 1024 tokens
  • โ€”Gradient Checkpointing: Enabled

๐Ÿ“Š Training Metrics

  • โ€”Total Steps: 295
  • โ€”Final Loss: 1.618368478548729
  • โ€”Trainable Params: 24,313,856

## โš–๏ธ License
This model inherits the Apache 2.0 license.