CoolFace
Modelpublic

Undi95/Llama2-13B-no_robots-alpaca-lora

sourceHugging Facecc-by-nc-4.0updated 3y agoView on Hugging Face
11likes130downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

<img src="https://raw.githubusercontent.com/OpenAccess-AI-Collective/axolotl/main/image/axolotl-badge-web.png" alt="Built with Axolotl" width="200" height="32"/>

no_robots-alpaca

This lora was trained with Doctor-Shotgun/no-robots-sharegpt dataset on TheBloke/Llama-2-13B-fp16. It achieves the following results on the evaluation set:

  • Loss: 1.6087

Model description

The LoRA was trained on Doctor-Shotgun/no-robots-sharegpt, a ShareGPT converted dataset from the OG HuggingFaceH4/no_robots but with Alpaca prompting.

Prompt template: Alpaca

Below is an instruction that describes a task. Write a response that appropriately completes the request.

### Instruction:
{prompt}

### Response:

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • learning_rate: 0.00065
  • trainbatchsize: 2
  • evalbatchsize: 2
  • seed: 42
  • optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • lrschedulertype: constant
  • lrschedulerwarmup_steps: 10
  • num_epochs: 2

Training results

Training LossEpochStepValidation Loss
1.55230.011.5476
1.21390.1421.5008
1.63480.2841.4968
1.64980.31261.4962
1.56450.41681.4983
1.64870.52101.4981
1.61470.62521.4965
1.30480.72941.4973
1.62050.83361.5007
1.60450.93781.5003
1.57811.04201.5013
1.48071.094621.5492
1.05411.195041.5596
1.23371.295461.5789
0.97191.395881.5859
1.21891.496301.5959
1.25661.596721.5968
0.70491.697141.5987
1.21331.797561.5907
1.03271.897981.6087

Framework versions

  • Transformers 4.34.1
  • Pytorch 2.0.1+cu117
  • Datasets 2.14.6
  • Tokenizers 0.14.1

If you want to support me, you can here.

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

MetricValue
Avg.46.55
ARC (25-shot)58.87
HellaSwag (10-shot)82.43
MMLU (5-shot)53.11
TruthfulQA (0-shot)40.46
Winogrande (5-shot)75.3
GSM8K (5-shot)6.44
DROP (3-shot)9.26