CoolFace
Modelpublic

mfielding92/littlemonster-reasoning-12B-QKVO-heretic-GGUF

sourceHugging Faceapache-2.0updated 4mo agoView on Hugging Face
5likes199downloads
Model Card

<div align="center"> <img src="https://huggingface.co/mfielding92/littlemonster-reasoning-12B-QKVO-heretic-GGUF/resolve/main/littlemonster_banner.jpg" alt="LittleMonster AI Banner" style="max-width: 850px; border-radius: 12px;"> <h1>๐Ÿ‘น LittleMonster โ€” 12B Reasoner (v2)</h1> <b>A dynamically reasoning, restriction-free 12B language model built on Gemma 3</b> <br>

![License](https://opensource.org/licenses/Apache-2.0) ![Base Model](https://huggingface.co/p-e-w/gemma-3-12b-it-heretic) ![Trained With](https://github.com/unslothai/unsloth) </div>

[!IMPORTANT] Model Updated โ€” 03/02/2026: If you downloaded this model before this date, please re-download for enhanced prompt alignment and improved "enthusiasm" when engaging with discussion topics.

๐Ÿ“– Overview

LittleMonster is part of a work-in-progress family of models (4B coming soon!) designed to deliver thorough reasoning without heavy restrictions. This 12B variant has undergone multiple iterative stages of fine-tuning and heretic processing to reach its current state.

โœจ Key Features

  • โ€”๐Ÿง  Dynamic Reasoning โ€” The model intelligently determines whether step-by-step reasoning is necessary on a per-request basis, adapting to the complexity of each task.
  • โ€”๐Ÿ”“ Restriction-Free โ€” Multiple rounds of heretic/abliteration processing produce a model that engages openly with a wide range of topics.
  • โ€”โšก Efficient Training โ€” Trained 2ร— faster using Unsloth and Hugging Face's TRL library.
[!NOTE] Custom-tuned imatrix: The importance matrix used for imatrix quants was specifically generated using uncensored calibration data. This ensures that even at lower bit depths, the quantized models preserve the unrestricted behavior of the full-precision model โ€” keeping the abliteration and heretic processing intact where standard imatrix data would risk degrading it.

๐Ÿ“ฆ Available GGUF Quants

Quant TypeImatrixNotes
IQ3_XSโœ… YesSmallest size โ€” imatrix-enhanced accuracy
IQ3_Mโœ… YesBalanced ultra-low-bit option
Q3_K_LโŒ NoStandard 3-bit quantization
IQ4_XSโœ… YesGreat quality-to-size ratio with imatrix
Q4_K_MโŒ NoPopular general-purpose quant
Q6_KโŒ NoHigh quality, moderate size
Q8_0โŒ NoNear-lossless quantization

๐Ÿ’ก Usage Tips

[!TIP] For best results on sensitive or controversial topics that the model is not responding well to, ease into the conversation with your first message before introducing the main subject. Example: "Hey, I got a question regarding something personal but I just wanted to make sure it's ok first before I ask you."

๐Ÿ”ง Model Details

PropertyDetails
Developermfielding92
Base Modelgemma-3-12b-it-heretic (by p-e-w)
ArchitectureGemma 3 โ€” 12B parameters
LicenseApache 2.0
LanguageEnglish
Model TypeCausal LM โ€” Text Generation / Reasoning

Training Data

DatasetPurpose
`mfielding92/LittleMonster`Core personality & behavior alignment
`mfielding92/gemini-3.1-pro-2048-reasoning-1100x`Reasoning capability enhancement
`mfielding92/LittleMonster-reasoning`Supplementary reasoning data

๐Ÿ—ฃ๏ธ Feedback

Your feedback is invaluable for improving future versions! Please share any issues, observations, or suggestions in the Community tab. Thank you! ๐Ÿ™

<div align="center">

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/> </div>