TyKaoz/gemma-4-12B-it-4bit
04
Gemma 4 12B Instruct — 4-bit MLX
4-bit MLX quantization of `google/gemma-4-12B-it`, for Apple Silicon (~6.3 GB). Vision-language model — run it with mlx-vlm, not mlx-lm.
Usage
pip install -U mlx-vlmpython -m mlx_vlm.generate \
--model TyKaoz/gemma-4-12B-it-4bit \
--prompt "Explique la quantization en une phrase." \
--max-tokens 200By [TyKaoz](https://www.tykaoz.bzh) — privacy-first native macOS LLM chat client. Apache 2.0, inherited from the base model.
