CoolFace
Modelpublic

tianzl66/Llama-3.1-8B-Instruct-CommonSense170K-LoRA-Epoch1

sourceHugging Faceupdated 22d agoView on Hugging Face
0likes36downloads
Model Card

Llama-3.1-8B-Instruct + Commonsense170K — LoRA (Epoch 1)

LoRA adapter for Llama-3.1-8B-Instruct fine-tuned on Commonsense170K.

Adapter

  • —Dataset: Commonsense170K
  • —Training epochs: 2
  • —LoRA rank: 16
  • —LoRA alpha: 32
  • —Target modules: qproj, kproj, vproj, oproj, gateproj, upproj, down_proj

Evaluation

TaskLoRA+ Spectral Surgery (o_proj + down_proj, 8+2)
BoolQ87.0336%86.9725%
PIQA87.3232%87.3232%
SocialIQA79.6315%79.6315%
HellaSwag91.6550%90.3107%
WinoGrande87.6085%86.8193%
ARC-Easy93.8131%93.7290%
ARC-Challenge83.9590%84.7270%
OpenBookQA90.6000%90.6000%
Macro87.7030%87.5141%
Micro89.1521%88.5276%
Correct19,987 / 22,41919,847 / 22,419

Evaluation uses the Llama-3.1-Instruct tokenizer chat template, greedy decoding, max_new_tokens=8, the vLLM backend, max model length 2048, and seed 42.

Files

  • —adapter_model.safetensors: PEFT LoRA weights
  • —adapter_config.json: PEFT configuration
  • —eval-commonsense8/summary.json: eight-task aggregate metrics
  • —eval-commonsense8/summary.csv: compact task metrics