Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4-1M
<p align="center"> <img src="https://cdn-uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice-AI Banner" width="100%"> </p>
<h1 align="center">Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored (NVFP4-1M)</h1> <h3 align="center">Official Solstice-AI Native 1M Context • NVIDIA FP4 / Blackwell</h3>
<p align="center"> <b>Original Model & GAIN Merge by <a href="https://huggingface.co/DavidAU">DavidAU</a> • Native 1M YaRN Scaling & Packaging by <a href="https://huggingface.co/Solstice-AI">Solstice-AI</a></b> </p>
Overview
This model is pre-configured with native-esque 1,048,576 token (1M) YaRN scaling baked directly into config.json. Users do not need to supply command-line flags or rope overrides: vLLM and SGLang automatically initialize 1M rotary positional frequencies on load.
Serving with vLLM (Plug & Play 1M)
vllm serve Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4-1M \
--tensor-parallel-size 1