CoolFace
Datasetpublic

prykin/flash-moe-weights

Flash-MoE Weights: Qwen3.5-397B-A17B (4-bit, Metal) Pre-packed weights for flash-moe — a pure C/Metal inference engine that runs the 397B-parameter Qwen3.5 MoE model on a single MacBook. Source Model mlx-community/Qwen3.5-397B-A17B-4bit File Structure File Size Description model_weights.bin 5.5 GB Non-expert weights (embeddings, attention, norms, shared expert, routing gates) model_weights.json 371 KB Tensor manifest (offsets, shapes… See the full description on the dataset page: https://huggingface.co/datasets/prykin/flash-moe-weights.

sourceHugging Faceapache-2.0updated 6mo agoView on Hugging Face
0likes42downloads
filemodel_weights.bin5.14 GBdownload
filetokenizer.bin7.8 MBdownload
filevocab.bin2.2 MBdownload

prykin/flash-moe-weights · main · files are served by the source, never re-hosted here