CoolFace
Modelpublic

RichardErkhov/etri-xainlp_-_llama3-8b-dpo_v1-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes532downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

llama3-8b-dpo_v1 - GGUF

  • —Model creator: https://huggingface.co/etri-xainlp/
  • —Original model: https://huggingface.co/etri-xainlp/llama3-8b-dpo_v1/

Original model description: --- license: apache-2.0 ---

etri-xainlp/llama3-8b-dpo_v1

Model Details

Model Developers ETRI xainlp team

Input text only.

Output text only.

Model Architecture

Base Model meta-llama/Llama-8b-hf

Training Dataset

  • —sft+lora: 1,821 k instruction-following set
  • —dpo+lora: 221 k user preference set
  • —We use A100 GPU 80GB * 8, when training.