TyKaoz/gemma-4-12B-it-6bit
04
Gemma 4 12B Instruct — 6-bit MLX
6-bit MLX quantization of `google/gemma-4-12B-it`, for Apple Silicon (~9.1 GB). Vision-language model — run it with mlx-vlm, not mlx-lm.
Usage
pip install -U mlx-vlmpython -m mlx_vlm.generate \
--model TyKaoz/gemma-4-12B-it-6bit \
--prompt "Explique la quantization en une phrase." \
--max-tokens 200By [TyKaoz](https://www.tykaoz.bzh) — privacy-first native macOS LLM chat client. Apache 2.0, inherited from the base model.
