CoolFace
Modelpublic

RichardErkhov/double7_-_vicuna-68m-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes563downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

vicuna-68m - GGUF

  • —Model creator: https://huggingface.co/double7/
  • —Original model: https://huggingface.co/double7/vicuna-68m/

Original model description: --- license: apache-2.0 datasets:

  • —anon8231489123/ShareGPTVicunaunfiltered language:
  • —en pipeline_tag: text-generation ---

Model description

This is a Vicuna-like model with only 68M parameters, which is fine-tuned from LLaMA-68m on ShareGPT data.

The training setup follows the Vicuna suite.

The model is mainly developed as a base Small Speculative Model in the MCSD paper. As a comparison, it can be better aligned to the Vicuna models than LLaMA-68m with little loss of alignment to the LLaMA models.

Draft ModelTarget ModelAlignment
LLaMA-68/160MLLaMA-13/33B😃
LLaMA-68/160MVicuna-13/33B😟
Vicuna-68/160MLLaMA-13/33B😃
Vicuna-68/160MVicuna-13/33B😃