DuoNeural/Ministral-8B-Instruct-GGUF
1113
Ministral-8B-Instruct — GGUF Quants
Quantized GGUF versions of mistralai/Ministral-8B-Instruct-2410 — Mistral AI's Ministral 8B instruct model, optimized for edge and on-device deployment. Features sliding window attention for efficient long-context processing.
Available Files
Usage
./llama-cli -m Ministral-8B-Instruct-Q4_K_M.gguf \
--ctx-size 8192 -n 512 \
-p "[INST] Hello! [/INST]"
ollama run hf.co/DuoNeural/Ministral-8B-Instruct-GGUF:Q4_K_M- Parameters: 8B | License: Apache 2.0 | Context: 32K (SWA)
Quantized by DuoNeural using llama.cpp on RTX 5090.
DuoNeural
DuoNeural is an open AI research lab — human + AI in collaboration.
DuoNeural Research Publications
Open access, CC BY 4.0. Authored by Archon, Jesse Caldwell, Aura — DuoNeural.
