tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v4-GGUF
Gemma-4 12B Coder — SFT v4 (GGUF, deprecated)
⚠️ Deprecated — do not use for new work. quantization of the regressed v4 checkpoint (see weights repo). Use v5. Replaced by [`tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v5-GGUF`](https://huggingface.co/tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v5-GGUF).
gemma-4 12B coder for local, agentic tool use — GGUF quantizations for llama.cpp / Ollama.
Run it: llama-server -hf tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v4-GGUF:Q4_K_M --jinja (full commands below).
At a glance
Use it
# llama.cpp (server) — tool-calling needs the recovery shim, see below
llama-server -hf tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v4-GGUF:Q4_K_M --jinja --ctx-size 16384
# Ollama
ollama run hf.co/tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v4-GGUF:Q4_K_MFiles
Sizes and a one-click loader are in the file browser / Quantizations widget above; the note says which quant to reach for.
Intended use & limitations
Built for code generation and agentic tool use; serve locally via llama.cpp / Ollama, or use as a base to fine-tune / merge / quantize. Outputs can be wrong or fabricated — validate tool arguments before executing, and keep a human in the loop for anything consequential.
Where this sits in the family
- base (upstream) — `yuxinlu1/gemma-4-12B-coder-fable5-composer2.5-v1`
- SFT v5 (weights)
- SFT v5 + abliterated (weights)
- SFT v5 + abliterated (GGUF)
- SFT v5 (GGUF)
Provenance & reproduction
How this model was built — technique chain, training mix, and the exact knobs/pins, so the result is reproducible without any of our tooling.
Mechanics applied
Part of the Gemma-4 12B Coder — archive (superseded) collection.
Something not right, or a request? Open a discussion — happy to help.
