CoolFace
Modelpublic

sardukar/physiology-8k-llama3-8b-qlora

sourceHugging Facemitupdated 2y agoView on Hugging Face
0likes5downloads
Model Card

Model Card for Model ID

<!-- Provide a quick summary of what the model is/does. --> This model is a 1 epoch training with ORPO Trainer on the sardukar/physiology-mcqa-8k dataset

Base model is NousResearch/Meta-Llama-3-8B-Instruct

Training results [image]