CoolFace
Modelpublic

avrecum/mistral7b-v0.3-alpaca-cleaned

sourceHugging Facemitupdated 2y agoView on Hugging Face
0likes18downloads
Model Card

Model Card for mistral7b-v0.3-alpaca-cleaned

<!-- Provide a quick summary of what the model is/does. -->

Mistral 7B v0.3 finetuned on cleaned Stanford Alpaca dataset using LoRA

Model was finetuned on for 1 epoch using pagedadamw8bit optimizer with these params: perdevicetrainbatchsize = 10, gradientaccumulationsteps = 4, warmupsteps = 5, numtrainepochs=1, learningrate = 2e-4, optim = "pagedadamw8bit", weightdecay = 0.01, lrscheduler_type = "linear", seed = 3407