CoolFace
Modelpublic

tensorblock/gemma-3-4b-it-GGUF

sourceHugging Facegemmaupdated 8mo agoView on Hugging Face
0likes172downloads
Model Card

<div style="width: auto; margin-left: auto; margin-right: auto"> <img src="https://i.imgur.com/jC7kdl8.jpeg" alt="TensorBlock" style="width: 100%; min-width: 400px; display: block; margin: auto;"> </div>

![Website](https://tensorblock.co) ![Twitter](https://twitter.com/tensorblock_aoi) ![Discord](https://discord.gg/Ej5NmeHFf2) ![GitHub](https://github.com/TensorBlock) ![Telegram](https://t.me/TensorBlock)

google/gemma-3-4b-it - GGUF

This repo contains GGUF format model files for google/gemma-3-4b-it.

The files were quantized using machines provided by TensorBlock, and they are compatible with llama.cpp as of commit b4882.

Our projects

<table border="1" cellspacing="0" cellpadding="10"> <tr> <th colspan="2" style="font-size: 25px;">Forge</th> </tr> <tr> <th colspan="2"> <img src="https://imgur.com/faI5UKh.jpeg" alt="Forge Project" width="900"/> </th> </tr> <tr> <th colspan="2">An OpenAI-compatible multi-provider routing layer.</th> </tr> <tr> <th colspan="2"> <a href="https://github.com/TensorBlock/forge" target="_blank" style=" display: inline-block; padding: 8px 16px; background-color: #FF7F50; color: white; text-decoration: none; border-radius: 6px; font-weight: bold; font-family: sans-serif; ">๐Ÿš€ Try it now! ๐Ÿš€</a> </th> </tr>

<tr> <th style="font-size: 25px;">Awesome MCP Servers</th> <th style="font-size: 25px;">TensorBlock Studio</th> </tr> <tr> <th><img src="https://imgur.com/2Xov7B7.jpeg" alt="MCP Servers" width="450"/></th> <th><img src="https://imgur.com/pJcmF5u.jpeg" alt="Studio" width="450"/></th> </tr> <tr> <th>A comprehensive collection of Model Context Protocol (MCP) servers.</th> <th>A lightweight, open, and extensible multi-LLM interaction studio.</th> </tr> <tr> <th> <a href="https://github.com/TensorBlock/awesome-mcp-servers" target="blank" style=" display: inline-block; padding: 8px 16px; background-color: #FF7F50; color: white; text-decoration: none; border-radius: 6px; font-weight: bold; font-family: sans-serif; ">๐Ÿ‘€ See what we built ๐Ÿ‘€</a> </th> <th> <a href="https://github.com/TensorBlock/TensorBlock-Studio" target="blank" style=" display: inline-block; padding: 8px 16px; background-color: #FF7F50; color: white; text-decoration: none; border-radius: 6px; font-weight: bold; font-family: sans-serif; ">๐Ÿ‘€ See what we built ๐Ÿ‘€</a> </th> </tr> </table>

Prompt template

<bos><start_of_turn>user
{system_prompt}

{prompt}<end_of_turn>
<start_of_turn>model

Model file specification

FilenameQuant typeFile SizeDescription
gemma-3-4b-it-Q2_K.ggufQ2_K1.729 GBsmallest, significant quality loss - not recommended for most purposes
gemma-3-4b-it-Q3_K_S.ggufQ3KS1.937 GBvery small, high quality loss
gemma-3-4b-it-Q3_K_M.ggufQ3KM2.098 GBvery small, high quality loss
gemma-3-4b-it-Q3_K_L.ggufQ3KL2.236 GBsmall, substantial quality loss
gemma-3-4b-it-Q4_0.ggufQ4_02.363 GBlegacy; small, very high quality loss - prefer using Q3KM
gemma-3-4b-it-Q4_K_S.ggufQ4KS2.378 GBsmall, greater quality loss
gemma-3-4b-it-Q4_K_M.ggufQ4KM2.490 GBmedium, balanced quality - recommended
gemma-3-4b-it-Q5_0.ggufQ5_02.764 GBlegacy; medium, balanced quality - prefer using Q4KM
gemma-3-4b-it-Q5_K_S.ggufQ5KS2.764 GBlarge, low quality loss - recommended
gemma-3-4b-it-Q5_K_M.ggufQ5KM2.830 GBlarge, very low quality loss - recommended
gemma-3-4b-it-Q6_K.ggufQ6_K3.191 GBvery large, extremely low quality loss
gemma-3-4b-it-Q8_0.ggufQ8_04.130 GBvery large, extremely low quality loss - not recommended

Downloading instruction

Command line

Firstly, install Huggingface Client

shell
pip install -U "huggingface_hub[cli]"

Then, downoad the individual model file the a local directory

shell
huggingface-cli download tensorblock/gemma-3-4b-it-GGUF --include "gemma-3-4b-it-Q2_K.gguf" --local-dir MY_LOCAL_DIR

If you wanna download multiple model files with a pattern (e.g., *Q4_K*gguf), you can try:

shell
huggingface-cli download tensorblock/gemma-3-4b-it-GGUF --local-dir MY_LOCAL_DIR --local-dir-use-symlinks False --include='*Q4_K*gguf'