CoolFace
Modelpublic

sunil-pathak/gemma-4-E4B-it-GGUF

sourceHugging Faceotherupdated 6mo agoView on Hugging Face
0likes29downloads
Model Card

Gemma 4 E4B IT โ€“ GGUF (Q8 Quantized)

Model Size Quantization Format Runtime

๐Ÿ”ท Model Overview

This repository provides a GGUF-format quantized version of the original:

  • โ€”Base Model: google/gemma-4-E4B-it
  • โ€”Developed by: Google
  • โ€”Format: GGUF (for llama.cpp)
  • โ€”Quantization: Q8 (8-bit)
  • โ€”Conversion Tooling: llama.cpp

This model enables efficient CPU-based inference.

โš ๏ธ License & Usage Notice

This is a converted derivative model.

๐Ÿ‘‰ You MUST comply with: https://huggingface.co/google/gemma-4-E4B-it

  • โ€”โŒ No new rights
  • โ€”โŒ Not official
  • โ€”โœ… Ownership remains with Google

๐Ÿ“ฆ Files

FileDescription
gemma-4-E4B-it.Q8.gguf~8GB quantized model

โš™๏ธ Technical Specs

ParameterValue
ArchitectureGemma
FormatGGUF
QuantizationQ8
Runtimellama.cpp

๐Ÿš€ Quick Start

bash
./llama-simple -m gemma-4-E4B-it.Q8.gguf -p "Explain AI simply."