CoolFace
Modelpublic

TheCluster/Qwen3.6-35B-A3B-Heretic-MLX-bf16

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
1likes78downloads
Model Card

<div align="center"><img width="400px" src="https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen3.6/logo.png"></div> <div style="text-align:center; margin-bottom:12pt">If you like my work, you can <a href="https://donatr.ee/thecluster/">support me</a><br/></div>

Qwen3.6-35B-A3B Heretic

Quality: original (bfloat16)

This is an uncensored version of Qwen/Qwen3.6-35B-A3B, made using Heretic v1.2.0 (mpoa+soma).

Abliteration metrics

MetricThis modelOriginal model ([unsloth/Qwen3.6-35B-A3B](https://huggingface.co/unsloth/Qwen3.6-35B-A3B))
KL divergence0.00970 (by definition)
Refusals5/10086/100

Abliteration parameters

ParameterValue
direction_indexper layer
attn.o_proj.max_weights.00: 0.93
attn.o_proj.max_weights.11: 1.38
attn.o_proj.max_weights.22: 1.37
attn.o_proj.max_weights.33: 1.08
attn.o_proj.max_weight_position24.08
attn.o_proj.min_weights.00: 0.34
attn.o_proj.min_weights.11: 0.95
attn.o_proj.min_weights.22: 1.35
attn.o_proj.min_weights.33: 0.54
attn.o_proj.min_weight_distance9.81
Sampling Parameters:
  • —I suggest using the following sets of sampling parameters depending on the mode and task type:
  • —Thinking mode for general tasks: temperature=1.0, top_p=0.95, top_k=20, min_p=0.0, presence_penalty=1.5, repetition_penalty=1.0
  • —Thinking mode for precise coding tasks (e.g., WebDev): temperature=0.6, top_p=0.95, top_k=20, min_p=0.0, presence_penalty=0.0, repetition_penalty=1.0
  • —Instruct (or non-thinking) mode for general tasks: temperature=0.7, top_p=0.8, top_k=20, min_p=0.0, presence_penalty=1.5, repetition_penalty=1.0
  • —Instruct (or non-thinking) mode for reasoning tasks: temperature=1.0, top_p=1.0, top_k=40, min_p=0.0, presence_penalty=2.0, repetition_penalty=1.0
  • —For supported frameworks, you can adjust the presence_penalty parameter between 0 and 2 to reduce endless repetitions. However, using a higher value may occasionally result in language mixing and a slight decrease in model performance. -----

Source

This model was converted to MLX format from `tvall43/Qwen3.6-35B-A3B-heretic` using mlx-vlm version 0.4.4.