CoolFace
Modelpublic

bartowski/bigstral-12b-32k-8xMoE-GGUF

sourceHugging Faceupdated 3y agoView on Hugging Face
2likes448downloads
Model Card

Llamacpp Quantizations of bigstral-12b-32k-8xMoE

Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b2354">b2354</a> for quantization.

Original model: https://huggingface.co/bartowski/bigstral-12b-32k-8xMoE

Download a file (not the whole branch) from below:

FilenameQuant typeFile SizeDescription
bigstral-12b-32k-8xMoE-Q8_0.ggufQ8_086.63GBExtremely high quality, generally unneeded but max available quant.
bigstral-12b-32k-8xMoE-Q6_K.ggufQ6_K67.00GBVery high quality, near perfect, recommended.
bigstral-12b-32k-8xMoE-Q5_K_M.ggufQ5KM58.00GBHigh quality, very usable.
bigstral-12b-32k-8xMoE-Q5_K_S.ggufQ5KS56.25GBHigh quality, very usable.
bigstral-12b-32k-8xMoE-Q5_0.ggufQ5_056.25GBHigh quality, older format, generally not recommended.
bigstral-12b-32k-8xMoE-Q4_K_M.ggufQ4KM49.60GBGood quality, similar to 4.25 bpw.
bigstral-12b-32k-8xMoE-Q4_K_S.ggufQ4KS46.70GBSlightly lower quality with small space savings.
bigstral-12b-32k-8xMoE-Q4_0.ggufQ4_046.13GBDecent quality, older format, generally not recommended.
bigstral-12b-32k-8xMoE-Q3_K_L.ggufQ3KL42.16GBLower quality but usable, good for low RAM availability.
bigstral-12b-32k-8xMoE-Q3_K_M.ggufQ3KM39.30GBEven lower quality.
bigstral-12b-32k-8xMoE-Q3_K_S.ggufQ3KS35.62GBLow quality, not recommended.
bigstral-12b-32k-8xMoE-Q2_K.ggufQ2_K30.17GBExtremely low quality, not recommended.

Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski