CoolFace
Modelpublic

kabachuha/G4-MeroMero-v2-31B-Heretic-ARA-LoRA

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
2likes31downloads
Model Card

G4-MeroMero-v2-31B - Heretic - LoRA

91% fewer refusals (8/100 This vs 97/100 Original) while preserving model quality (0.0785 KL divergence).

Made with the new 4bit ARA LoRA (usearalora) technique at home.

FAQ - How to use this model?

Because I don't want to spam 60+ GB uploads on huggingface, this model (natively created as lora) is distributed as a lora file. You can merge it into the original model using PEFT or, if you are using llama.cpp, you can use the convenient GGUF file (convert lora to gguf script) which is located in the "quantized" section.

llama.cpp example command:

./llama-server <your current command> --lora-scaled "/media/kabachuha/doc/G4-MeroMero-v2-31B-Heretic-ARA-LoRA.gguf:1.0"

Performance

MetricThis modelOriginal model ([G4-MeroMero-v2-31B](https://huggingface.co/zerofata/G4-MeroMero-v2-31B))
KL divergence0.07850 (by definition)
Refusals8/10097/100

Lower refusals indicate fewer content restrictions, while lower KL divergence indicates more closeness to the original model's baseline. Higher refusals cause more rejections, objections, pushbacks, lecturing, censorship, softening and deflections.

Abliteration parameters

ParameterValue
start_layer_index26
end_layer_index52
preserve_good_behavior_weight0.8949
steer_bad_behavior_weight0.0025
overcorrect_relative_weight0.9535
neighbor_count15

Targeted components

  • —attn.o_proj
  • —mlp.down_proj