CoolFace
Modelpublic

mlx-community/Llama3-ChatQA-1.5-8B-4bit

sourceHugging Faceotherupdated 2y agoView on Hugging Face
0likes33downloads
Model Card

mlx-community/Llama3-ChatQA-1.5-8B-4bit

This model was converted to MLX format from [mlx-community/Llama3-ChatQA-1.5-8B]() using mlx-lm version 0.12.0.

Model added by Prince Canuma.

Refer to the original model card for more details on the model.

Use with mlx

bash
pip install mlx-lm
python
from mlx_lm import load, generate

model, tokenizer = load("mlx-community/Llama3-ChatQA-1.5-8B-4bit")
response = generate(model, tokenizer, prompt="hello", verbose=True)