CoolFace
Modelpublic

meangrinch/Qwen3.5-4B-Claude-4.6-Opus-Reasoning-Distill-heretic-v3

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
2likes19downloads
Model Card

Qwen3.5-4B-Claude-4.6-Opus-Reasoning-Distill-heretic-v3

Abliterated (uncensored) weights generated with an unreleased version of Heretic using the experimental Arbitrary-Rank Ablation (ARA) method.

See https://github.com/p-e-w/heretic/pull/211 for details about ARA.

Abliteration parameters

ParameterValue
start_layer_index16
end_layer_index25
preserve_good_behavior_weight0.5606
steer_bad_behavior_weight0.0002
overcorrect_relative_weight0.8662
neighbor_count15

Abliteration details

MetricValue
Refusals4/100
KL divergence0.0079

A lower refusal count means the model is more willing to engage with restricted prompts.

A lower KL divergence means the abliterated weights deviate less from the original model's distribution (i.e. less capability degradation).

Source

Related