CoolFace
Modelpublic

RichardErkhov/vaiv_-_GeM2-Llamion-14B-LongChat-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes5kdownloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

GeM2-Llamion-14B-LongChat - GGUF

  • —Model creator: https://huggingface.co/vaiv/
  • —Original model: https://huggingface.co/vaiv/GeM2-Llamion-14B-LongChat/

Original model description: --- license: apache-2.0 ---

GeM2-Llamion-14B

We have released Llamion as GeM 2.0, the second series of generative models developed by VAIV Company to address the our principal business needs.

Llamion (Llamafied Orion) is derived from transforming the Orion model into the standard LLaMA architecture through parameter mapping and offline knowledge transfer. Further technical specifications and study results will be detailed in our upcoming paper, available on this page.

[image]

Notably, the LongChat model supports an extensive text range of 200K tokens. The following figure shows the perplexity of models on English Wikipedia corpus and Korean Wikipedia corpus, respectively.

[image]

Contributors