CoolFace
Datasetpublic

prykin/flash-moe-weights

Flash-MoE Weights: Qwen3.5-397B-A17B (4-bit, Metal) Pre-packed weights for flash-moe — a pure C/Metal inference engine that runs the 397B-parameter Qwen3.5 MoE model on a single MacBook. Source Model mlx-community/Qwen3.5-397B-A17B-4bit File Structure File Size Description model_weights.bin 5.5 GB Non-expert weights (embeddings, attention, norms, shared expert, routing gates) model_weights.json 371 KB Tensor manifest (offsets, shapes… See the full description on the dataset page: https://huggingface.co/datasets/prykin/flash-moe-weights.

sourceHugging Faceapache-2.0updated 6mo agoView on Hugging Face
0likes42downloads
5 commits on main
1b466066mo ago

Delete .DS_Store

prykin
42c35836mo ago

Add files using upload-large-folder tool

prykin
29b35dd6mo ago

Add files using upload-large-folder tool

prykin
be3f3266mo ago

Create README.md

prykin
8ded85c6mo ago

initial commit

prykin