CoolFace
Modelpublic

weddle/Qwen3.8-27B-AutoRound-W4A16-G128

sourceHugging Faceapache-2.0updated 18d agoView on Hugging Face
1likes114downloads
7 commits on main
9598bf418d ago

Clarify decode timing in headline caption

weddle
3c7106c18d ago

Lead with quantization purpose and single-B65 performance highlights

weddle
f5fb2e018d ago

Add measured short and long context prefill rates

weddle
0dedc6918d ago

Document single-B65 throughput hardware and context capacity

weddle
4e70ddb18d ago

Improve model card layout and explain packed tensor display

weddle
769776818d ago

Add AutoRound W4A16 G128 artifact and model card

weddle
18b153b18d ago

initial commit

weddle