glogwa68/granite-4.0-h-1b-DISTILL-glm-4.7-GGUF
3447
granite-4.0-h-1b-DISTILL-glm-4.7-GGUF
GGUF quantized versions of granite-4.0-h-1b-DISTILL-glm-4.7
Available Formats
Quick Start
Ollama
# Use Q4_K_M (recommended)
ollama run hf.co/glogwa68/granite-4.0-h-1b-DISTILL-glm-4.7-GGUF:Q4_K_M
# Or other quantizations
ollama run hf.co/glogwa68/granite-4.0-h-1b-DISTILL-glm-4.7-GGUF:Q8_0
ollama run hf.co/glogwa68/granite-4.0-h-1b-DISTILL-glm-4.7-GGUF:Q2_Kllama.cpp
# Download and run
llama-cli --hf-repo glogwa68/granite-4.0-h-1b-DISTILL-glm-4.7-GGUF --hf-file granite-4.0-h-1b-distill-glm-4.7-q4_k_m.gguf -p "Hello, how are you?"
# With server
llama-server --hf-repo glogwa68/granite-4.0-h-1b-DISTILL-glm-4.7-GGUF --hf-file granite-4.0-h-1b-distill-glm-4.7-q4_k_m.gguf -c 2048LM Studio / GPT4All
Download the .gguf file of your choice and load it in your application.
Quantization Details
Original Model
This is the quantized version of granite-4.0-h-1b-DISTILL-glm-4.7
- Base Model: ibm-granite/granite-4.0-h-1b
- Fine-tuning Dataset: TeichAI/glm-4.7-2000x
- Training Loss: 0.6364
