CoolFace
Modelpublic

ThakiCloud/Qwen3-Coder-30B-A3B-Prune3-W4A16

sourceHugging Faceapache-2.0updated 21d agoView on Hugging Face
0likes337downloads
14 commits on main
03f323c21d ago

docs: 카드 메타데이터 보강 (base_model_relation/library_name/datasets/language)

thaki-AI
9625d2e1mo ago

Measure the Hopper case directly instead of inferring it (0.81-0.84x vs FP8)

thaki-AI
a9c76fa2mo ago

Answer the open kernel-tuning hedge with the four-way NVFP4 vs FP8 result; note the FP8 size discrepancy

thaki-AI
eb395ac2mo ago

card: NVFP4 sibling — 4-bit slowness is W4A16's, not 4-bit's

thaki-AI
774ffb82mo ago

add model.safetensors

thaki-AI
ab53bc72mo ago

add tokenizer.json

thaki-AI
4b69d062mo ago

add config.json

thaki-AI
5cab4e52mo ago

add qwen3coder_tool_parser.py

thaki-AI
ae543ab2mo ago

add chat_template.jinja

thaki-AI
ad7121d2mo ago

add tokenizer_config.json

thaki-AI
dd702222mo ago

add generation_config.json

thaki-AI
83b6dce2mo ago

add recipe.yaml

thaki-AI
f1ab7fd2mo ago

card: measured results before weights

thaki-AI
a54e0092mo ago

initial commit

thaki-AI