CoolFace
Modelpublic

devan-carlin/Qwen3.8-Flash-Next-W4A16

sourceHugging Faceotherupdated 28d agoView on Hugging Face
1likes1.4kdownloads
11 commits on main
40b8f1828d ago

Reword serving tip: describe failure mode as repetition/loops, not language

devan-carlin
f9f37b128d ago

Document temperature guidance: run at 0.7, not the recommended 1.0; pin via override-generation-config

devan-carlin
6e3232e28d ago

Note served arch is Qwen4ExpForConditionalGeneration (multimodal)

devan-carlin
2f27b4b28d ago

Revert text-only note: repo ships model.visual.* weights (333 tensors), model is multimodal

devan-carlin
af3f11928d ago

Fix model card: text-only (no visual.* weights shipped), correct MTP note

devan-carlin
aba5c5a28d ago

Upload README.md with huggingface_hub

devan-carlin
a2c629528d ago

Add real links: fork branch xpu-qwen4exp, setup script, patch, ops guide

devan-carlin
342c82e28d ago

Remove owner-only publishing section from model card

devan-carlin
1de421728d ago

Polish model card language (Gemma4-assisted edit)

devan-carlin
26ce17228d ago

Qwen3.8-Flash-Next W4A16 (int4 g128) + PLE table, XPU serving recipe

devan-carlin
5eb1de328d ago

initial commit

devan-carlin