mlx-community/quantized-gemma-7b-it
6616
mlx-community/quantized-gemma-7b-it
This model was converted to MLX format from [google/gemma-7b-it](). Refer to the original model card for more details on the model.
Use with mlx
pip install mlx-lmfrom mlx_lm import load, generate
model, tokenizer = load("mlx-community/quantized-gemma-7b-it")
response = generate(model, tokenizer, prompt="hello", verbose=True)