CoolFace
Modelpublic

mlx-community/Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated-4bit-msq

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
4likes100downloads
Model Card

Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated — MLX 4.4 BPW

Mixed-precision MLX quantization of `huihui-ai/Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated`, quantized with MLX Smart Quantize (MSQ) — my own sensitivity-based mixed-precision quantization method for Apple Silicon. It measures per-layer NMSE and assigns optimal bit widths automatically, combining architecture knowledge with measured data.

Details

  • —Type: Vision (VLM)
  • —Average: 4.45 bits per weight
  • —Method: MLX Smart Quantize (MSQ)
  • —AWQ scaling: applied to 96 groups