Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored (NVFP4-1M)
Official Solstice-AI Native 1M Context • NVIDIA FP4 / Blackwell
Original Model & GAIN Merge by DavidAU • Native 1M YaRN Scaling & Packaging by Solstice-AI
Overview
This model is pre-configured with native-esque 1,048,576 token (1M) YaRN scaling baked directly into config.json. Users do not need to supply command-line flags or rope overrides: vLLM and SGLang automatically initialize 1M rotary positional frequencies on load.
Serving with vLLM (Plug & Play 1M)
vllm serve Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4-1M \
--tensor-parallel-size 1