CoolFace
Modelpublic

hfl/chinese-alpaca-2-7b-rlhf-gguf

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
4likes339downloads
Model Card

Chinese-Alpaca-2-7B-RLHF-GGUF

This repository contains GGUF-v3 version (llama.cpp compatible) of Chinese-Alpaca-2-7B-RLHF, which is tuned on Chinese-Alpaca-2-7B with RLHF using DeepSpeed-Chat.

Performance

Metric: PPL, lower is better

Quantoriginalimatrix (`-im`)
Q2_K10.5211 +/- 0.1413911.9331 +/- 0.16168
Q3_K8.9748 +/- 0.120438.8238 +/- 0.11850
Q4_08.7843 +/- 0.11854-
Q4_K8.4643 +/- 0.113418.4226 +/- 0.11302
Q5_08.4563 +/- 0.11353-
Q5_K8.3722 +/- 0.112368.3336 +/- 0.11192
Q6_K8.3207 +/- 0.111848.3047 +/- 0.11159
Q8_08.3100 +/- 0.11173-
F168.3112 +/- 0.11173-

The model with `-im` suffix is generated with important matrix, which has generally better performance (not always though).

Others

For full model in HuggingFace format, please see: https://huggingface.co/hfl/chinese-alpaca-2-7b-rlhf

Please refer to https://github.com/ymcui/Chinese-LLaMA-Alpaca-2/ for more details.