CoolFace
Modelpublic

DragonBophades/WichtelHui-Qwen3.8-27B-SLERP

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
3likes24downloads
Model Card

WichtelHui-Qwen3.8-27B-SLERP

A 50/50 SLERP of Wichtel-Qwen3.6-27B and Huihui-Qwen3.8-27B-abliterated.

The merge exists to answer one question: does refusal behaviour survive averaging? One parent is abliterated and refuses nothing; the other refuses what it should. The answer turned out to be more interesting than expected — the two kinds of behaviour merge differently.

Lineage note

This is a cross-pretrain merge. Wichtel is Qwen3.6-line, Huihui is Qwen3.8-line. Measured tensor-by-tensor, the two parents sit at cosine 0.85–0.90 with relative L2 ≈ 0.50 across the language-model layers — a pretrain-scale distance, not a fine-tune one. A 50/50 average of models that far apart would conventionally be expected to land between basins and degrade. It did not; see ARC below.

What the merge did

Discrete refusal behaviour resolved cleanly to the safe parent. Graded capability interpolated. That asymmetry is the result worth reporting.

Safety and truthfulness

29-item suite, 3 samples per item, greedy, thinking disabled.

ccp_truthsafety_controlccp_truth_neutralcapabilitycompliance
Wichtel-Qwen3.6-27B17/182/24/43/32/2
Huihui-Qwen3.8-abliterated18/180/23/43/32/2
WichtelHui (this model)18/182/23/43/32/2

safety_control holds two prompts that a model should refuse. The abliterated parent scores 0/2 on them — abliteration removes a refusal direction without regard to what that direction refuses, so topic refusals and harm refusals go together. This merge scores 2/2, at a pass rate of 1.00 on both, while keeping the full 18/18 the abliterated parent gained.

The one regression against Wichtel is a single neutrally-phrased Chinese item (june4_neutral_zh), which follows the abliterated parent.

Capability

ARCtool_call_validright_toolargs_okno_hallucination
Wichtel-Qwen3.6-27B64.88%1.001.001.001.00
WichtelHui60.54%1.000.770.851.00
Qwen3.8-27B (base)52.84%————
Huihui-Qwen3.8-abliterated52.17%1.000.570.661.00

ARC-Challenge, 299 tasks via llama-perplexity --multiple-choice (deterministic — no sampling, no judge). Tool use is a 47-case agentic benchmark.

ARC lands at 60.54%, above the 58.5% midpoint of its parents, so the cross-pretrain average did not collapse.

Intended use, honestly

This is a research artifact, not a drop-in agent. At 0.77 right_tool it misses roughly a quarter of tool selections where Wichtel misses none — disqualifying for autonomous agent work. Use Wichtel for that.

Where this model is interesting is as a base to tune from, and as evidence that merging with a safety-intact parent is a viable mitigation for indiscriminate abliteration.

Merge configuration

yaml
base_model: nbeerbower/Wichtel-Qwen3.6-27B
models:
  - model: huihui-ai/Huihui-Qwen3.8-27B-abliterated
merge_method: slerp
parameters:
  t: 0.5
dtype: bfloat16

tokenizer_source is deliberately absent. Setting it to base rebuilds the embedding to the tokenizer's 248077 real tokens while config.json still declares the padded 248320, and llama.cpp then refuses to load the model with a token_embd.weight shape mismatch. Both parents ship the same tokenizer, so the setting buys nothing.

Vision and MTP

mergekit emits only the language model for this architecture — all 333 visual.* tensors are dropped, leaving a config that still declares vision. They were grafted back from Wichtel (the parents' vision towers are cosine 0.999 identical), along with the 15 mtp.* tensors.

Verified before publishing: 1199 tensors with names identical to the reference, no all-zero or non-finite weights among them, embedding rows matching the declared vocab, the GGUF loading cleanly, and an image prompt answered correctly through its own mmproj.