mlx-community/Qwen3.8-27B-OBLITERATED-OptiQ-4bit
63.4k
Update kv_config.json: re-measured with `optiq kv-cache` float32 sensitivity (the previous config was ranked on bfloat16 rounding noise). 8-bit layers [15, 51] -> [19, 35].
Add kv_config.json: per-layer KV cache precision from `optiq kv-cache` (target 4.5 bits). Use with `optiq serve --kv-config kv_config.json`.
OptiQ mixed-precision 4-bit quant
initial commit
