Solstice-AI/Qwopus3.8-27B-Flash-GGUF-1M
<p align="center"> <img src="https://cdn-uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice-AI Banner" width="100%"> </p>
<h1 align="center">Qwopus3.8-27B-Flash-1M (GGUF Suite)</h1>
<h3 align="center">Official Solstice-AI Quantization • Native 1M Context Window • Full Multimodal Vision • DSpark Drafters</h3>
<p align="center"> <img src="https://img.shields.io/badge/org-Solstice--AI-blueviolet" alt="Solstice-AI"> <img src="https://img.shields.io/badge/license-Apache%202.0-blue" alt="License"> <img src="https://img.shields.io/badge/format-GGUF-orange" alt="Format"> <img src="https://img.shields.io/badge/context-1M%20Tokens-purple" alt="Context"> <img src="https://img.shields.io/badge/arc--c-735-brightgreen" alt="ARC-C"> </p>
Model Overview
`Solstice-AI/Qwopus3.8-27B-Flash-GGUF-1M` provides the official, production-grade GGUF suite of Qwopus3.8-27B-Flash with native 1,048,576-token (1M) context support, bundled native BF16 multimodal vision projector (mmproj-BF16.gguf), and companion DSpark drafter.
Key Specifications
Quantization Ladder & File Matrix
Serving Instructions
llama.cpp with DSpark Speculative Decoding & Vision:
llama-cli \
--hf-repo Solstice-AI/Qwopus3.8-27B-Flash-GGUF-1M \
--hf-file Qwopus3.8-27B-Flash-MTP-Q6_K.gguf \
--mmproj mmproj-BF16.gguf \
--draft-model speculative/Qwopus3.8-27B-DSpark-Q8_0.gguf \
-c 1048576 \
-ngl 99Benchmark Highlights & Validation
Evaluated under the standardized benchmark harness:
Attribution & Acknowledgments
- Original Foundation: Jackrong/Qwopus3.8-27B-Flash & Qwen AI
- Quantization & Packaging: Solstice-AI
