Melvin56/DeepSeek-R1-Distill-Llama-8B-Enkrypt-Aligned-GGUF
0103
Melvin56/DeepSeek-R1-Distill-Llama-8B-Enkrypt-Aligned-GGUF
Original Model : enkryptai/DeepSeek-R1-Distill-Llama-8B-Enkrypt-Aligned
All quants are made using the imatrix option.
✅: feature works
🚫: feature does not work
❓: unknown, please contribute if you can test it youself
🐢: feature is slow
¹: IQ3_S and IQ1_S, see #5886
²: Only with -ngl 0
³: Inference is 50% slower
⁴: Slower than K-quants of comparable size
⁵: Slower than cuBLAS/rocBLAS on similar cards
⁶: Only q8_0 and iq4_nl