prykin/flash-moe-weights
Flash-MoE Weights: Qwen3.5-397B-A17B (4-bit, Metal) Pre-packed weights for flash-moe — a pure C/Metal inference engine that runs the 397B-parameter Qwen3.5 MoE model on a single MacBook. Source Model mlx-community/Qwen3.5-397B-A17B-4bit File Structure File Size Description model_weights.bin 5.5 GB Non-expert weights (embeddings, attention, norms, shared expert, routing gates) model_weights.json 371 KB Tensor manifest (offsets, shapes… See the full description on the dataset page: https://huggingface.co/datasets/prykin/flash-moe-weights.
This repository belongs to prykin on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
