inferencerlabs/Hy3-MLX-Q9
339
Hy3
No longer available on HF due to storage restrictions - archived here
See Hy3 in action: demonstration videos
Tested with an M3 Ultra 512 GiB using Inferencer app
- Text inference: ~20 tokens/s @ 1000 tokens ~315 GiB

No longer available on HF due to storage restrictions - archived here
See Hy3 in action: demonstration videos
Tested with an M3 Ultra 512 GiB using Inferencer app
