majentik/Nemotron-3.5-Lightning-30B-A3B-MLX-MXFP4
0171
Nemotron-3.5-Lightning-30B-A3B-MLX-MXFP4
MLX MXFP4 (mxfp4, group size 32) quantized variant of nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 for Apple silicon via mlx-lm.
Provenance
- Source: nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 @ revision
d468880b6ad3c6e0d21377ce7242adaea4cc884d(OpenMDW v1.1 — see LICENSE in this repo). - Quantized with
mlx_lm.convert(mlx-lm 0.31.3): mxfp4, 4-bit, group size 32.
Smoke gate
Before upload this pack passed a deterministic coherence gate: greedy 48-token chat generation loaded through mlx_lm.load, judged for emptiness, repetition loops, multi-script gibberish, and special-token debris. Verdict: ok.
Usage
pip install mlx-lm
mlx_lm.generate --model majentik/Nemotron-3.5-Lightning-30B-A3B-MLX-MXFP4 --prompt "Hello"Evaluation
License
OpenMDW v1.1 (openmdw-1.1): permissive open model, distribution, and weights license; the full license text ships in this repository as LICENSE. See https://openmdw.ai/license/1-1/.
