RichardErkhov/georgesung_-_llama2_7b_chat_uncensored-8bits
013
Quantization made by Richard Erkhov.
llama27bchat_uncensored - bnb 8bits
- Model creator: https://huggingface.co/georgesung/
- Original model: https://huggingface.co/georgesung/llama27bchat_uncensored/
Original model description: --- license: other datasets:
- georgesung/wizardvicuna70k_unfiltered ---
Overview
Fine-tuned Llama-2 7B with an uncensored/unfiltered Wizard-Vicuna conversation dataset (originally from ehartford/wizard_vicuna_70k_unfiltered). Used QLoRA for fine-tuning. Trained for one epoch on a 24GB GPU (NVIDIA A10G) instance, took ~19 hours to train.
The version here is the fp16 HuggingFace model.
GGML & GPTQ versions
Thanks to TheBloke, he has created the GGML and GPTQ versions:
- https://huggingface.co/TheBloke/llama27bchat_uncensored-GGML
- https://huggingface.co/TheBloke/llama27bchat_uncensored-GPTQ
Prompt style
The model was trained with the following prompt style:
### HUMAN:
Hello
### RESPONSE:
Hi, how are you?
### HUMAN:
I'm fine.
### RESPONSE:
How can I help you?
...Training code
Code used to train the model is available here.
To reproduce the results:
git clone https://github.com/georgesung/llm_qlora
cd llm_qlora
pip install -r requirements.txt
python train.py configs/llama2_7b_chat_uncensored.yamlFine-tuning guide
https://georgesung.github.io/ai/qlora-ift/
Open LLM Leaderboard Evaluation Results
Detailed results can be found here
