batiai/Qwen3.5-35B-A3B-GGUF
064
Qwen 3.5 35B-A3B GGUF — Quantized by BatiAI
<p align="center"> <a href="https://flow.bati.ai"><img src="https://img.shields.io/badge/BatiFlow-macOS%20AI%20Automation-blue?style=for-the-badge&logo=apple" alt="BatiFlow"></a> <a href="https://ollama.com/batiai/qwen3.5-35b"><img src="https://img.shields.io/badge/Ollama-batiai%2Fqwen3.5--35b-green?style=for-the-badge" alt="Ollama"></a> </p>
IQ4_XS quantization of Qwen/Qwen3.5-35B-A3B for on-device AI on Mac. Built and verified by BatiAI for BatiFlow.
Quick Start
ollama pull batiai/qwen3.5-35b:iq4Available Quantizations
Why MoE Beats Dense
35B-A3B is a Mixture-of-Experts model — 35B total, only 3B active per token:
MoE activates 9x fewer parameters — same quality, much faster, less memory.
Benchmarks — M4 Max (128GB)
Full BatiAI Qwen 3.5 Lineup
Technical Details
- Original Model: Qwen/Qwen3.5-35B-A3B
- Architecture: MoE (35B total, 3B active, 256 experts, 8 routed + 1 shared)
- Context Window: 262K tokens
- License: Apache 2.0
- Quantized with: llama.cpp (build 400ac8e)
About BatiFlow
BatiFlow — free, on-device AI automation for Mac. 5MB app, 100% local, unlimited.
License
Quantized from Qwen/Qwen3.5-35B-A3B. License: Apache 2.0.
Benchmarks
<!-- BENCH-START -->
<!-- BENCH-END -->
