tianzl66/Llama-3.1-8B-Instruct-CommonSense170K-LoRA-Epoch2
025
Llama-3.1-8B-Instruct + Commonsense170K — LoRA (Epoch 2)
LoRA adapter for Llama-3.1-8B-Instruct fine-tuned on Commonsense170K.
Adapter
- Dataset: Commonsense170K
- Training epochs: 2
- LoRA rank: 16
- LoRA alpha: 32
- Target modules: qproj, kproj, vproj, oproj, gateproj, upproj, down_proj
Evaluation
Evaluation uses the Llama-3.1-Instruct tokenizer chat template, greedy decoding, max_new_tokens=8, the vLLM backend, max model length 2048, and seed 42.
Files
adapter_model.safetensors: PEFT LoRA weightsadapter_config.json: PEFT configurationeval-commonsense8/summary.json: eight-task aggregate metricseval-commonsense8/summary.csv: compact task metrics
