CoolFace
Modelpublic

afrideva/smol_llama-220M-GQA-GGUF

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
0likes408downloads
Model Card

BEE-spoke-data/smol_llama-220M-GQA-GGUF

Quantized GGUF model files for smol_llama-220M-GQA from BEE-spoke-data

Original Model Card:

smol_llama: 220M GQA

model card WIP, more details to come

A small 220M param (total) decoder model. This is the first version of the model.

  • —1024 hidden size, 10 layers
  • —GQA (32 heads, 8 key-value), context length 2048
  • —train-from-scratch on one GPU :)