GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

barozp/Qwen3.8-27B-MTP-GGUF overview

Qwen3.8 27B MTP GGUF NOTE Placeholder — not yet available. This repo is reserved ahead of the Qwen/Qwen3.8 27B release and will be filled in once that model sh…

ggufllama.cppqwenlicense:apache-2.0region:us
Downloads
0
Likes
2
Pipeline
Author

Repository Files & Downloads

0 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Browse files on Hugging Face

Model Details

Model IDbarozp/Qwen3.8-27B-MTP-GGUF
Authorbarozp
Pipeline
Licenseapache-2.0
Base model
Last modified2026-08-12T07:10:28.000Z

Model README

---

license: apache-2.0

tags:

  • gguf
  • llama.cpp
  • qwen

---

Qwen3.8-27B-MTP-GGUF

> [!NOTE]

> Placeholder — not yet available. This repo is reserved ahead of the

> Qwen/Qwen3.8-27B release and will be filled in once that model ships.

> Watch this repo or check back after the Qwen3.8 announcement.

GGUF of Qwen/Qwen3.8-27B with its Multi-Token Prediction (MTP) head preserved for self-speculative decoding in llama.cpp. Dense architecture (not MoE), so this is a straightforward tensor extraction, no expert remapping involved.

Plain + MTP — base model with self-speculative decoding enabled.

Related releases in this line

Prior work from this account

This follows the same recipe (REAP-style efficiency work, Opus reasoning

distillation, MTP grafting, full GGUF quant ladders with measured

speed/quality benchmarks) used for the Qwen3.6 line:

Qwen3.6-29B-REAP-Opus-Reasoning-Distill-MTP-GGUF.

Run barozp/Qwen3.8-27B-MTP-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models