Abiray/gemma-4-E4B-it-The-DECKARD-HERETIC-UNCENSORED-Thinking-GGUF
111.8k
Gemma 4 E4B Instruct - The DECKARD HERETIC UNCENSORED Thinking (GGUF)
This repository contains GGUF format quantizations for DavidAU's Gemma-4-E4B-it-The-DECKARD-HERETIC-UNCENSORED-Thinking.
These files allow for highly efficient local inference on CPUs and Apple Silicon, as well as VRAM-constrained GPUs, utilizing llama.cpp and compatible frontends.
๐ง About the Base Model
The base model is a specialized, heavily fine-tuned version of Google's Gemma 4 E4B (Effective 4B parameters).
- HERETIC / DECKARD Fine-Tune: This model has undergone the "HERETIC" abliteration and fine-tuning process by DavidAU. It is completely uncensored, unrestricted, and heavily calibrated for complex creative writing, roleplay, and deep fiction generation.
- Reasoning & Thinking: It includes augmented reasoning/thinking capabilities, allowing the model to internally chain logic before outputting text, making its responses deeper and more contextually accurate.
- Context Window: The Gemma 4 architecture supports extended context windows (up to 256k in base), making it highly capable for long-form conversations.
๐ฆ Available Quantization Formats
This repository provides multiple levels of quantization to help you balance VRAM/RAM usage, generation speed, and model fidelity.
