CoolFace
Modelpublic

Mantrah/Qwen3.8-27B-NVFP4-GDN

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
2likes234downloads
4 commits on main
53097a41mo ago

Model card: long-context RULER-lite vs unsloth (needle 64/64 both, VT-120K 10 vs 13), gsm8k with official sampling (95.5 vs 96.4), greedy note, 460 W power limit

Mantrah
1591d0d1mo ago

Model card: correct the novelty claim — GDN-in-NVFP4 uploads existed since 08-14; what this adds is the FP8 lm_head + side-by-side measurements

Mantrah
773af1f1mo ago

Qwen3.8-27B NVFP4, linear attention included

Mantrah
b3e2e471mo ago

initial commit

Mantrah