matrixportalx/gemma-2-9b-it-GGUF
0457
matrixportal/gemma-2-9b-it-GGUF
This model was converted to GGUF format from `google/gemma-2-9b-it` using llama.cpp via the ggml.ai's all-gguf-same-where space. Refer to the original model card for more details on the model.
You can access the list of available quantized model files and download links here.
โ Quantized Models Download List
๐ Recommended Quantizations
- โจ General CPU Use: `Q4_K_M` (Best balance of speed/quality)
- ๐ฑ ARM Devices: `Q4_0` (Optimized for ARM CPUs)
- ๐ Maximum Quality: `Q8_0` (Near-original quality)
๐ฆ Full Quantization Options
๐ก Tip: Use F16 for maximum precision when quality is critical
