CoolFace
Modelpublic

bartowski/Tess-10.7B-v2.0-GGUF

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
1likes175downloads
Model Card

Llamacpp Quantizations of Tess-10.7B-v2.0

Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b2536">b2536</a> for quantization.

Original model: https://huggingface.co/Joseph717171/Tess-10.7B-v2.0

Download a file (not the whole branch) from below:

FilenameQuant typeFile SizeDescription
Tess-10.7B-v2.0-Q8_0.ggufQ8_011.40GBExtremely high quality, generally unneeded but max available quant.
Tess-10.7B-v2.0-Q6_K.ggufQ6_K8.80GBVery high quality, near perfect, recommended.
Tess-10.7B-v2.0-Q5_K_M.ggufQ5KM7.59GBHigh quality, very usable.
Tess-10.7B-v2.0-Q5_K_S.ggufQ5KS7.39GBHigh quality, very usable.
Tess-10.7B-v2.0-Q5_0.ggufQ5_07.39GBHigh quality, older format, generally not recommended.
Tess-10.7B-v2.0-Q4_K_M.ggufQ4KM6.46GBGood quality, uses about 4.83 bits per weight.
Tess-10.7B-v2.0-Q4_K_S.ggufQ4KS6.11GBSlightly lower quality with small space savings.
Tess-10.7B-v2.0-IQ4_NL.ggufIQ4_NL6.14GBDecent quality, similar to Q4KS, new method of quanting,
Tess-10.7B-v2.0-IQ4_XS.ggufIQ4_XS5.82GBDecent quality, new method with similar performance to Q4.
Tess-10.7B-v2.0-Q4_0.ggufQ4_06.07GBDecent quality, older format, generally not recommended.
Tess-10.7B-v2.0-Q3_K_L.ggufQ3KL5.65GBLower quality but usable, good for low RAM availability.
Tess-10.7B-v2.0-Q3_K_M.ggufQ3KM5.19GBEven lower quality.
Tess-10.7B-v2.0-IQ3_M.ggufIQ3_M4.84GBMedium-low quality, new method with decent performance.
Tess-10.7B-v2.0-IQ3_S.ggufIQ3_S4.69GBLower quality, new method with decent performance, recommended over Q3 quants.
Tess-10.7B-v2.0-Q3_K_S.ggufQ3KS4.66GBLow quality, not recommended.
Tess-10.7B-v2.0-Q2_K.ggufQ2_K4.00GBExtremely low quality, not recommended.

Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski