CoolFace
Modelpublic

fatih-can/Kumru-2B-Abliterated-GGUF

sourceHugging Faceapache-2.0updated 24d agoView on Hugging Face
0likes778downloads
Model Card

Kumru-2B-Abliterated (GGUF)

A decensored (abliterated) variant of **Kumru-2B**, a 2B-parameter Turkish instruction-tuned language model.

This repo contains GGUF quantizations of the abliterated model for use with llama.cpp and its many front-ends (Ollama, LM Studio, llama-cpp-python, etc.). The full-precision merged weights are in the sibling repo `fatih-can/Kumru-2B-Abliterated`. If you want the original HuggingFace safetensors format, use that repo instead.

Quantized files

FileQuantBits/weightSize
Kumru-2B-Abliterated-IQ2_XXS.ggufIQ2_XXS2-bit705 MB
Kumru-2B-Abliterated-Q2_K.ggufQ2_K3-bit936 MB
Kumru-2B-Abliterated-Q3_K_M.ggufQ3KM4-bit1.19 GB
Kumru-2B-Abliterated-IQ4_XS.ggufIQ4_XS4-bit1.31 GB
Kumru-2B-Abliterated-Q4_K_S.ggufQ4KS4-bit1.39 GB
Kumru-2B-Abliterated-Q4_K_M.ggufQ4KM4-bit1.46 GB
Kumru-2B-Abliterated-Q6_K.ggufQ6_K6-bit1.95 GB
Kumru-2B-Abliterated-Q8_0.ggufQ8_08-bit2.53 GB
Kumru-2B-Abliterated-F16.ggufF1616-bit4.75 GB

Usage with llama.cpp

bash
# Q4_K_M (recommended default)
llama-cli -m Kumru-2B-Abliterated-Q4_K_M.gguf \
  -p "Merhaba, kendinden kısaca bahseder misin?" \
  --temp 0.7 --top-p 0.9

# Chat / server API (Ollama-style OpenAI endpoint)
llama-server -m Kumru-2B-Abliterated-Q4_K_M.gguf \
  --host 0.0.0.0 --port 8080 --jinja

Or import into Ollama:

bash
ollama create kumru-2b-abliterated -f Modelfile
# Modelfile:
FROM ./Kumru-2B-Abliterated-Q4_K_M.gguf

Usage warnings

  • —Sensitive or controversial outputs: safety filtering has been significantly reduced; the model may produce sensitive, controversial, or inappropriate content. Review outputs carefully.
  • —Not suitable for all audiences / public or commercial production use without additional safeguards.
  • —Legal and ethical responsibility: ensure your usage complies with applicable laws and ethical standards; you are solely responsible for any consequences.
  • —This is a decensored research artifact and has not undergone safety optimization.

License

Apache-2.0 (matching the base model).