Abiray/Gemory-26B-A4B-GGUF
11.2k
Gemory-26B-A4B - GGUF
This repository contains GGUF quantizations of UltimateIntent/Gemory-26B-A4B-GGUF.
Gemory-26B-A4B is a 26-billion parameter Mixture of Experts (MoE) model built on the Gemma architecture foundation, fine-tuned for high-fidelity creative writing, unrestricted roleplay, and complex multi-turn conversational depth. All quantizations were generated directly from the uncompressed 50.5 GB BF16 base model and sanity-tested for generation integrity.
Quantization Breakdown
Usage Instructions
1. llama.cpp
Command-Line Inference (`llama-cli`):
./llama-cli \
-hf Abiray/Gemory-26B-A4B-GGUF:Q4_K_M \
-p "Write an opening scene set in a rain-soaked neon alleyway." \
-n 512 \
--temp 0.7