brittlewis12/stablelm-2-zephyr-1_6b-GGUF
StableLM-2-Zephyr-1.6B GGUF
Original model: StableLM 2 Zephyr 1.6B Model creator: Stability AI
This repo contains GGUF format model files for Stability AI’s StableLM 2 Zephyr 1.6B.
Stable LM 2 Zephyr 1.6B is a 1.6 billion parameter instruction tuned language model inspired by HugginFaceH4's Zephyr 7B training pipeline. The model is trained on a mix of publicly available datasets and synthetic datasets, utilizing Direct Preference Optimization (DPO).
What is GGUF?
GGUF is a file format for representing AI models. It is the third version of the format, introduced by the llama.cpp team on August 21st 2023. It is a replacement for GGML, which is no longer supported by llama.cpp. Converted using an proposed version of llama.cpp (PR #5052)
Prompt template: Zephyr
<|system|>
{{system_message}}<|endoftext|>
<|user|>
{{prompt}}<|endoftext|>
<|assistant|>Download & run with cnvrs on iPhone, iPad, and Mac!

cnvrs is the best app for private, local AI on your device:
- create & save Characters with custom system prompts & temperature settings
- download and experiment with any GGUF model you can find on HuggingFace!
- make it your own with custom Theme colors
- powered by Metal ⚡️ & Llama.cpp, with haptics during response streaming!
- try it out yourself today, on Testflight!
- follow cnvrs on twitter to stay up to date
Original Model Evaluations:
| Model | Size | MT-Bench | |-------------------------|------|----------| | Mistral-7B-Instruct-v0.2| 7B | 7.61 | | Llama2-Chat | 70B | 6.86 | | stablelm-zephyr-3b | 3B | 6.64 | | MPT-30B-Chat | 30B | 6.39 | | stablelm-2-zephyr-1.6b | 1.6B | 5.42 | | Falcon-40B-Instruct | 40B | 5.17 | | Qwen-1.8B-Chat | 1.8B | 4.95 | | dolphin-2.6-phi-2 | 2.7B | 4.93 | | phi-2 | 2.7B | 4.29 | | TinyLlama-1.1B-Chat-v1.0| 1.1B | 3.46 |
