CoolFace
Modelpublic

salohcin714/granite-4.2-3b-bf16-mlx

sourceHugging Faceapache-2.0updated 29d agoView on Hugging Face
0likes627downloads
Model Card

granite-4.2-3b-bf16-mlx

Provenance

Converted from `ibm-granite/granite-4.2-3b` using mlx-lm 0.31.3.

Usage

python
from mlx_lm import load, generate

model, tokenizer = load("salohcin714/granite-4.2-3b-bf16-mlx")
messages = [{"role": "user", "content": "Hello"}]
prompt = tokenizer.apply_chat_template(messages, add_generation_prompt=True)
text = generate(model, tokenizer, prompt=prompt, verbose=True)

Modifications

Weights converted to MLX safetensors layout (no quantization; weights kept in their original bfloat16 precision). Redundant tied lm_head.weight dropped where the model ties input/output embeddings. No fine-tuning; no added training data.

License and attribution

Licensed under Apache 2.0. Original weights by the Granite Team, IBM. See the upstream model card and the included LICENSE file for the full text.

Disclaimer

This repository is not affiliated with or endorsed by IBM. "Granite" is an IBM trademark, used here descriptively to identify the origin of the base model. IBM's published benchmarks describe the original weights, not this quantized/converted artifact, and must not be read as claims about this repo.