CoolFace
Modelpublic

ThakiCloud/Qwen3-Coder-30B-A3B-W4A16

sourceHugging Faceapache-2.0updated 20d agoView on Hugging Face
1likes503downloads
14 commits on main
de1daf620d ago

docs: 카드 메타데이터 보강 (base_model_relation/library_name/datasets/language)

thaki-AI
25ca7841mo ago

Measure the Hopper case directly: W4A16 loses to FP8 there too (0.81-0.84x)

thaki-AI
95e5e622mo ago

Answer the open kernel-tuning hedge with the four-way NVFP4 vs FP8 result; note the FP8 size discrepancy

thaki-AI
9c97c3d2mo ago

card: NVFP4 sibling — 4-bit slowness is W4A16's, not 4-bit's

thaki-AI
0f2a55d2mo ago

add model.safetensors

thaki-AI
a9545ac2mo ago

add tokenizer.json

thaki-AI
36c6d722mo ago

add config.json

thaki-AI
e4b63c22mo ago

add qwen3coder_tool_parser.py

thaki-AI
ef5824e2mo ago

add chat_template.jinja

thaki-AI
b5a2eb52mo ago

add tokenizer_config.json

thaki-AI
1b41e602mo ago

add generation_config.json

thaki-AI
f24f9c12mo ago

add recipe.yaml

thaki-AI
c27bd082mo ago

card: measured results before weights

thaki-AI
36ce3d42mo ago

initial commit

thaki-AI