GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Solstice-AI/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-GGUF overview

<p align="center" <img src="https://cdn uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice AI Banner"…

ggufsolstice-aiqwen3.8qwen3_827bneo-codercoderswe-benchllama.cppollamavisionmultimodalmmprojspeculative-decodingmtpuncensoredabliteratedimage-text-to-textenzhbase_model:DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAUbase_model:quantized:DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAUlicense:apache-2.0endpoints_compatible

Runs locally from ~888.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
image-text-to-text

Repository Files & Downloads

11 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q4_K_M.ggufGGUFQ4_K_M16.81 GBDownload
Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q5_K_M.ggufGGUFQ5_K_M19.31 GBDownload
Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q6_K.ggufGGUFQ6_K21.96 GBDownload
Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q8_0.ggufGGUFQ8_027.74 GBDownload
mmproj-BF16.ggufGGUFBF16888.0 MBDownload
speculative-mtp/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-MTP-Q4_K_M.ggufGGUFQ4_K_M17.23 GBDownload
speculative-mtp/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-MTP-Q5_K_M.ggufGGUFQ5_K_M19.73 GBDownload
speculative-mtp/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-MTP-Q6_K.ggufGGUFQ6_K22.38 GBDownload
speculative-mtp/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-MTP-Q8_0.ggufGGUFQ8_028.16 GBDownload
speculative/Qwen3.8-27B-DSpark-Q4_K_M.ggufGGUFQ4_K_M1.03 GBDownload
speculative/Qwen3.8-27B-DSpark-Q8_0.ggufGGUFQ8_01.85 GBDownload

Model Details

Model IDSolstice-AI/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-GGUF
AuthorSolstice-AI
Pipelineimage-text-to-text
Licenseapache-2.0
Base modelDavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU
Last modified2026-09-08T02:26:29.000Z

Model README

---

language:

  • en
  • zh

license: apache-2.0

base_model: DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU

tags:

  • solstice-ai
  • qwen3.8
  • qwen3_8
  • 27b
  • neo-coder
  • coder
  • swe-bench
  • gguf
  • llama.cpp
  • ollama
  • vision
  • multimodal
  • mmproj
  • speculative-decoding
  • mtp
  • uncensored
  • abliterated

pipeline_tag: image-text-to-text

---

<p align="center">

<img src="https://cdn-uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice-AI Banner" width="100%">

</p>

<h1 align="center">Qwen3.8-27B-TURBO NEO-CODER (Official Clean GGUF Suite)</h1>

<h3 align="center">Official Solstice-AI Release &bull; Standard Clean UD 3.0 Matrix &bull; Multi-Token Prediction (MTP) Speculative Tiers &bull; Pure BF16 Multimodal Vision Projector</h3>

<p align="center">

<b>Original Architecture by <a href="https://huggingface.co/Qwen">Qwen / Alibaba Cloud</a> &bull; Uncensored Weights by <a href="https://huggingface.co/DavidAU">DavidAU</a> &bull; Curated &amp; Packaged by <a href="https://huggingface.co/Solstice-AI">Solstice-AI</a></b>

</p>

---

Model Summary

Solstice-AI/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-GGUF contains the official clean GGUF suite of Qwen3.8-27B NEO-CODER, bringing DavidAU's latest coding and agentic prompt engineering optimizations into standard, clean UD 3.0 GGUF binaries.

Key NEO-CODER Capabilities:

  1. Dynamic Reasoning Effort Controls (reasoning_effort):

- medium: Suppresses default system prompt injection for direct, unrestricted coding execution and SWE-bench compatibility.

- xhigh: Injects deep-reasoning verification tags (<thought>) for complex algorithmic design and proofs.

  1. Deterministic XML Tool Calling: Pre-configured for <tool_call><function=...><parameter=...></function></tool_call> execution.
  2. Pure BF16 Vision Transformer (mmproj-BF16.gguf): Standalone 16-bit multimodal vision projector with zero FP16 underflow risks.
  3. Multi-Token Prediction (MTP) Speculative Tiers: Bundles specialized MTP models (speculative-mtp/) and DSpark drafters (speculative/).

---

File Catalog

| Filename | Precision / Quant | Size | Recommended Use Case |

| :--- | :--- | :---: | :--- |

| Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q4_K_M.gguf | Q4_K_M (UD-Q4_K_XL) | 16.81 GB | Recommended: Best balance of speed, RAM footprint &amp; accuracy |

| Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q5_K_M.gguf | Q5_K_M | 19.31 GB | High-accuracy coding and mathematical reasoning |

| Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q6_K.gguf | Q6_K | 21.96 GB | Near-lossless weights for complex multi-file refactoring |

| Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q8_0.gguf | Q8_0 | 27.74 GB | Pure lossless 8-bit precision |

| *speculative-mtp/-MTP-Q4_K_M.gguf** | MTP Q4_K_M | 17.23 GB | Multi-Token Prediction enabled speculative decoding |

| *speculative-mtp/-MTP-Q8_0.gguf** | MTP Q8_0 | 28.16 GB | Lossless MTP speculative decoding |

| mmproj-BF16.gguf | Pure BF16 | 0.87 GB | Official standalone Multimodal Vision Projector |

| speculative/Qwen3.8-27B-DSpark-Q4_K_M.gguf | DSpark Drafter | 1.03 GB | Ultra-fast pre-aligned speculative draft model |

---

Quickstart with llama.cpp

Standard Multimodal Inference:

llama-server \
  -m Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q4_K_M.gguf \
  --mmproj mmproj-BF16.gguf \
  -c 131072 \
  --port 8080

Speculative Decoding (1.8x Speedup):

llama-cli \
  -m Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-Q4_K_M.gguf \
  -md speculative/Qwen3.8-27B-DSpark-Q4_K_M.gguf \
  --mmproj mmproj-BF16.gguf \
  -p "Write a high-performance async actor pool in Rust using Tokio."

Run Solstice-AI/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-NEO-CODER-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models