CoolFace
Modelpublic

jnjj/gemma-3-1b-it-qat-int4-quantized-inference-unrestricted-pruned-sf

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes9downloads
3 commits on main
72ed6121y ago

Create README.md

jnjj
02782971y ago

Upload INT4 quantized Gemma‑3‑1B‑IT QAT with bfloat16 compute, 90% magnitude pruning, extensive unconventional modifications including instruction conversion and GPTQ/AutoGPTQ flags, and only bfloat16 .weight tensors saved as safetensors (bfloat16 compute)

jnjj
1f6ded61y ago

initial commit

jnjj