RichardErkhov/double7_-_vicuna-160m-4bits
019
Quantization made by Richard Erkhov.
vicuna-160m - bnb 4bits
- Model creator: https://huggingface.co/double7/
- Original model: https://huggingface.co/double7/vicuna-160m/
Original model description: --- license: apache-2.0 datasets:
- anon8231489123/ShareGPTVicunaunfiltered language:
- en pipeline_tag: text-generation ---
Model description
This is a Vicuna-like model with only 160M parameters, which is fine-tuned from LLaMA-160m on ShareGPT data.
The training setup follows the Vicuna suite.
The model is mainly developed as a base Small Speculative Model in MCSD paper. As a comparison, it can be better aligned to the Vicuna models than LLaMA-160m with little loss of alignment to the LLaMA models.
