spiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF
01.6k
About
weighted/imatrix quants of https://huggingface.co/nbeerbower/Gemma4-Gutenberg-26B-A4B
Benchmarked on [CaliperBench](https://caliperbench.com) — a creative-writing benchmark scoring prose craft, roleplay and willingness rather than general intelligence. See this model's scores: caliperbench.com/m/gemma4-gutenberg-26b-a4b.
Note: Gemma4-Gutenberg-26B-A4B is published as a LoRA adapter, not a full model. These GGUFs are that adapter merged onto its base google/gemma-4-26B-A4B-it (the text LM extracted from the multimodal base), then quantized.
imatrix (importance matrix) computed over bartowski calibration_datav3. Weighted quants spend precision where it matters most — the low bit-rates (IQ1–IQ4) are meaningfully better than same-size static quants; at Q5/Q6 the difference is negligible.
Static (non-imatrix) quants of this model are at spiritfather/Gemma4-Gutenberg-26B-A4B-GGUF.
<!-- provided-files -->
Usage
If you are unsure how to use GGUF files, refer to one of TheBloke's READMEs for more details, including on how to concatenate multi-part files.
Provided Quants
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
