CoolFace
Modelpublic

jnjj/gemma-3-1b-it-qat-int4-quantized-inference-unrestricted-weights-only-sf

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes14downloads
4 commits on main
1afe5561y ago

Upload INT4 quantized Gemma‑3‑1B‑IT QAT with bfloat16 compute, extensive unconventional modifications including instruction conversion flag, and only bfloat16 .weight tensors saved as safetensors (bfloat16 compute)

jnjj
541df7f1y ago

Create README.md

jnjj
0728af21y ago

Upload INT4 quantized Gemma‑3‑1B‑IT QAT with bfloat16 compute, extensive unconventional modifications, and only bfloat16 .weight tensors saved as safetensors (bfloat16 compute)

jnjj
6cb66871y ago

initial commit

jnjj