CoolFace
Modelpublic

Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4-1M

sourceHugging Faceapache-2.0updated 16d agoView on Hugging Face
1likes956downloads
Model Card

<p align="center"> <img src="https://cdn-uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice-AI Banner" width="100%"> </p>

<h1 align="center">Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored (NVFP4-1M)</h1> <h3 align="center">Official Solstice-AI Native 1M Context &bull; NVIDIA FP4 / Blackwell</h3>

<p align="center"> <b>Original Model & GAIN Merge by <a href="https://huggingface.co/DavidAU">DavidAU</a> &bull; Native 1M YaRN Scaling & Packaging by <a href="https://huggingface.co/Solstice-AI">Solstice-AI</a></b> </p>

Overview

This model is pre-configured with native-esque 1,048,576 token (1M) YaRN scaling baked directly into config.json. Users do not need to supply command-line flags or rope overrides: vLLM and SGLang automatically initialize 1M rotary positional frequencies on load.

Serving with vLLM (Plug & Play 1M)

bash
vllm serve Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4-1M \
  --tensor-parallel-size 1