CoolFace
Datasetpublic

prykin/flash-moe-weights

Flash-MoE Weights: Qwen3.5-397B-A17B (4-bit, Metal) Pre-packed weights for flash-moe — a pure C/Metal inference engine that runs the 397B-parameter Qwen3.5 MoE model on a single MacBook. Source Model mlx-community/Qwen3.5-397B-A17B-4bit File Structure File Size Description model_weights.bin 5.5 GB Non-expert weights (embeddings, attention, norms, shared expert, routing gates) model_weights.json 371 KB Tensor manifest (offsets, shapes… See the full description on the dataset page: https://huggingface.co/datasets/prykin/flash-moe-weights.

sourceHugging Faceapache-2.0updated 6mo agoView on Hugging Face
0likes42downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
prykin/flash-moe-weights · CoolFace