CoolFace
Modelpublic

philbert440/Qwen3.8-27B-W4A16-AWQ

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
29likes497kdownloads
8 commits on main
7908d421mo ago

Fix tokenizer: drop leftover truncation block, restore upstream tokenizer.json/tokenizer_config.json (fixes vision >1024 tokens; see discussion #3)

philbert440
61078f71mo ago

Measured performance v2: warm multi-pass methodology, draft-mode comparison, instruct + concurrency regimes

philbert440
00ed8dc1mo ago

Update README.md

philbert440
8bcc88b1mo ago

Model card v3: base-model benchmarks, official sampling params, YaRN long-context, citation

philbert440
7bdc9171mo ago

Model card v2: badges, V100 throughput chart, full validation matrix, serving recipes

philbert440
3242f381mo ago

Model card: recipe, thinking-mode calibration, V100 validation matrix w/ MTP

philbert440
b4168c11mo ago

W4A16-AWQ g128 asym mse, June-proven hybrid smoothing, Magpie thinking calib, bf16 MTP graft

philbert440
f4845f41mo ago

initial commit

philbert440