OsaurusAI/gemma-4-31B-it-qat-JANG_4M
<p align="center"><a href="https://osaurus.ai"><img src="./osaurus-x-banner.png" alt="Osaurus AI"></a></p>
OsaurusAI/gemma-4-31B-it-qat-JANG_4M
JANG4M MLX affine bundle converted from [google/gemma-4-31B-it-qat-q40-unquantized](https://huggingface.co/google/gemma-4-31B-it-qat-q4_0-unquantized).
This bundle keeps Gemma 4 bookends and media-sensitive components coherent: token embeddings/output projection, norms, media towers/embedders, and Gemma 4 per-layer embedding/gate/projector tensors are fp16 passthrough; self-attention and MoE router projections are 8-bit affine; decoder MLP/expert bulk is 4-bit affine.
Bundle
Modalities
Runtime support depends on a Gemma 4 compatible MLX/vMLX loader that understands config.json quantization overrides, jang_config.json, and Gemma 4 processor/chat-template files.
Runtime Metadata
config.jsonhas source-derivedhas_vision,has_audio,has_video,modalities, andcapabilities.tokenizer_config.jsonincludesbos_token_id,eos_token_id,pad_token_id, and the patched Gemma 4 chat template.processor_config.jsonis preserved for Gemma 4 multimodal processing.- MTP/speculative drafter weights are not present in the source checkpoint; metadata is
mtp: none/mtp_policy: none.
