oomu/OOMU-Qwen3-0.6B-Production-MicroRouter-GGUF
OOMU Qwen3 0.6B Production MicroRouter
OOMU-Qwen3-0.6B-Production-MicroRouter-Q8_0.gguf is OOMU's compact, Metal-targeted routing model. It is a LoRA-trained derivative of Qwen/Qwen3-0.6B, merged into the official instruction checkpoint and quantized to Q8_0 with a pinned llama.cpp toolchain.
This is not a general-purpose assistant model. It is intended for OOMU's bounded staged micro-router contract, where native code supplies the allowed capability grammar, validates every model decision, and executes effects only through OOMU's native broker.
Identity
Runtime profile
- Context: 4,608 tokens
- Maximum generated output: 96 tokens
- Thinking mode: disabled
- Required OOMU production backend on Apple silicon: Metal, fail closed if unavailable
License and attribution
This derivative is distributed under Apache License 2.0. Qwen3 attribution and the complete license are preserved in the Hugging Face repository. See the upstream `Qwen/Qwen3-0.6B` model card for the base model's documentation and limitations.
