zenlm/zen5-nano-9B-gguf
Zen5 Nano 9B
Edge / on-device tier of the Zen5 family. Multimodal dense vision-language model (image + text in -> text out) at the 9B scale, qwen3_5 architecture, tuned for fast inference on laptops, phones, and consumer GPUs.
Repackaged from Qwen/Qwen3.5-9B (apache-2.0, Alibaba Qwen). Native HuggingFace safetensors (Qwen3_5ForConditionalGeneration), re-hosted under the Zen5 nano line. Not trained from scratch — a permissively-licensed redistribution for the OSS-clean Zen model line.
Zen5 nano ladder — pick by hardware budget:
The larger Zen5 tiers (zen5-flash, zen5, zen5-pro, zen5-max, zen5-coder) cover everything from agentic-default to frontier-quality multi-GPU.
Files
Native HuggingFace safetensors. Load with transformers following the upstream Qwen/Qwen3.5-9B recipe, or fetch directly:
hf download zenlm/zen5-nano-9B-gguf --local-dir zen5-nano-9B-ggufRun
Hosted via the Hanzo gateway (api.hanzo.ai) as zen5-nano-9B.
License
apache-2.0. Upstream: Qwen/Qwen3.5-9B by Alibaba Qwen. This repository redistributes the weights under the same license.
