CoolFace
Modelpublic

tianzl66/Llama-3.1-8B-Instruct-CommonSense170K-LoRA-Epoch2

sourceHugging Faceupdated 22d agoView on Hugging Face
0likes25downloads
Model Card

Llama-3.1-8B-Instruct + Commonsense170K — LoRA (Epoch 2)

LoRA adapter for Llama-3.1-8B-Instruct fine-tuned on Commonsense170K.

Adapter

  • —Dataset: Commonsense170K
  • —Training epochs: 2
  • —LoRA rank: 16
  • —LoRA alpha: 32
  • —Target modules: qproj, kproj, vproj, oproj, gateproj, upproj, down_proj

Evaluation

TaskLoRA+ Spectral Surgery (o_proj + down_proj, 8+2)
BoolQ88.0122%88.1346%
PIQA89.6083%89.4450%
SocialIQA82.0880%81.4739%
HellaSwag93.6566%93.2484%
WinoGrande88.7924%88.3189%
ARC-Easy93.8552%93.8973%
ARC-Challenge85.3242%85.5802%
OpenBookQA90.4000%90.6000%
Macro88.9671%88.8373%
Micro90.7311%90.4947%
Correct20,341 / 22,41920,288 / 22,419

Evaluation uses the Llama-3.1-Instruct tokenizer chat template, greedy decoding, max_new_tokens=8, the vLLM backend, max model length 2048, and seed 42.

Files

  • —adapter_model.safetensors: PEFT LoRA weights
  • —adapter_config.json: PEFT configuration
  • —eval-commonsense8/summary.json: eight-task aggregate metrics
  • —eval-commonsense8/summary.csv: compact task metrics