CoolFace
Modelpublic

ASethi04/meta-llama-Llama-3.1-8B-Instruct-dpo-HumanLLMs-Human-Like-DPO-Dataset-second-2-5e-06

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes
fileadapter_model.safetensors160.1 MBdownload
filetraining_args.bin6 KBdownload

ASethi04/meta-llama-Llama-3.1-8B-Instruct-dpo-HumanLLMs-Human-Like-DPO-Dataset-second-2-5e-06 · main · files are served by the source, never re-hosted here