CoolFace
Modelpublic

cesarsal1nas/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M-GGUF

sourceHugging Faceapache-2.0updated 6mo agoView on Hugging Face
10likes31kdownloads
Model Card

Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated — Q4KM GGUF

  • —This model has been tested by myself in Hermes Agent, it does everything it is supposed to, text, coding, vision, thinking, tools usage, everything works perfectly well so far there. I have also found that this very specific Opus 4.6 Distillation does not use as many emojis as the non-Opus one, which is wave:great.

This is a Q4KM GGUF quantization of huihui-ai/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated.

Refer to the original model card for full details, usage warnings, and licensing information.

This is a Qwen3.5-35B-A3B MoE model that has been:

  • —Fine-tuned in Claude 4.6 Opus style — targets Claude Opus-level instruction following, reasoning depth, and structured output quality
  • —Abliterated — refusal mechanisms removed; no alignment restrictions
  • —Vision-capable — the original weights include a full vision encoder (preprocessor_config.json, video_preprocessor_config.json); an mmproj file can be generated from the source safetensors for multimodal use

Files

FileSizeDescription
Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M.gguf~20 GBMain text model (Q4KM)
Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-mmproj-F16.gguf~858 MBVision projector for image input

Details

Source modelhuihui-ai/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated
Architectureqwen35moe (35B total params, ~3B active; 256 experts, 8 active per token)
QuantizationQ4KM (~4.88 BPW)
File size~20 GB + 858 MB mmproj
Quantized withllama.cpp b8352
Context length262,144 tokens (trained)

Usage with llama.cpp

bash
llama-cli \
  --hf-repo cesarsal1nas/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M-GGUF \
  --hf-file Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M.gguf \
  -p "Tell me about the universe"
bash
llama-server \
  --hf-repo cesarsal1nas/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M-GGUF \
  --hf-file Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M.gguf \
  -c 131072

With vision (image input)

bash
llama-server \
  --hf-repo cesarsal1nas/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M-GGUF \
  --hf-file Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M.gguf \
  --mmproj Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-mmproj-F16.gguf \
  -c 131072

Usage with Ollama

bash
ollama run hf.co/cesarsal1nas/Huihui-Qwen3.5-35B-A3B-Claude-4.6-Opus-abliterated-Q4_K_M-GGUF

Credits