CoolFace
Modelpublic

dareymon/YandexGPT-5-Lite-8B-instruct-oQ4

sourceHugging Faceupdated 10d agoView on Hugging Face
0likes112downloads
Model Card

YandexGPT-5-Lite-8B-instruct-oQ4

This model was quantized using oQ (oMLX v0.6.4) mixed-precision quantization.

Quantization details

  • —Model type: llama
  • —Bits: 4
  • —Group size: 64
  • —Format: MLX safetensors