CoolFace
Modelpublic

inferencerlabs/Hy3-MLX-Q9

sourceHugging Faceupdated 24d agoView on Hugging Face
3likes39downloads
Model Card

Hy3

No longer available on HF due to storage restrictions - archived here

See Hy3 in action: demonstration videos

Tested with an M3 Ultra 512 GiB using Inferencer app

  • —Text inference: ~20 tokens/s @ 1000 tokens ~315 GiB

Screenshot