CoolFace
Modelpublic

Freaksterz/Qwen3.8-27B-SmoothQuant-W8A8-INT8

sourceHugging Faceapache-2.0updated 23d agoView on Hugging Face
22likes36kdownloads
6 commits on main
2df4e3b23d ago

v3: add MTP head (model-mtp.safetensors, BF16, `re:.*mtp.*` ignored) + index/config; card: group naming fix, MTP AL

Freaksterz
7a2bd1d23d ago

v3: SmoothQuant + activation-aware GPTQ, mixed W8A8/W8A16 (288/112), no rotation — KLD 0.00556, DFlash2-compatible (AL 4.34)

Freaksterz
68746dd1mo ago

Remove v1-era quantization scripts superseded by the rotation pipeline

Freaksterz
29f2b291mo ago

Rotation + SmoothQuant + GPTQ W8A8 (KLD 0.01414 -> 0.01098); fix missing processor/tokenizer files (discussion #1)

Freaksterz
417ede11mo ago

Upload folder using huggingface_hub

Freaksterz
d8704041mo ago

initial commit

Freaksterz