GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

grapeV-ai/Qwen3.8-27B-GGUF overview

What is this? Qwen3.8 27B https://huggingface.co/Qwen/Qwen3.8 27B をMTP(=マルチトークン予測)レイヤーを含めてGGUFフォーマットに変換したものです。 Note mm mmproj Qwen3.8 27b BF16.gguf でビジョンエンコーダー…

gguflicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~13.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

9 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.8-27B-BF16.ggufGGUFBF1650.90 GBDownload
Qwen3.8-27B-IQ4_XS.ggufGGUFIQ4_XS14.26 GBDownload
Qwen3.8-27B-Q4_K_M.ggufGGUFQ4_K_M15.66 GBDownload
Qwen3.8-27B-Q5_K_M.ggufGGUFQ5_K_M18.19 GBDownload
Qwen3.8-27B-Q6_K.ggufGGUFQ6_K20.89 GBDownload
imatrix.ggufGGUFGGUF13.0 MBDownload
mmproj-Qwen3.8-27b-BF16.ggufGGUFBF16888.0 MBDownload
mmproj-Qwen3.8-27b-Q8_0.ggufGGUFQ8_0600.1 MBDownload
mtp-Qwen3.8-27b-BF16.ggufGGUFBF165.54 GBDownload

Model Details

Model IDgrapeV-ai/Qwen3.8-27B-GGUF
AuthorgrapeV-ai
Pipeline
Licenseapache-2.0
Base model
Last modified2026-08-25T14:01:19.000Z

Model README

---

license: apache-2.0

---

What is this?

Qwen3.8-27BをMTP(=マルチトークン予測)レイヤーを含めてGGUFフォーマットに変換したものです。

Note

-mm mmproj-Qwen3.8-27b-BF16.ggufでビジョンエンコーダーをロードし、Vision対応モデルとして使用することができます。

推論努力(Reasoning effort)はlow / medium / xhighから選択可能で、デフォルトはxhighです。<br>

llama.cppのserverの場合は起動時に以下の引数を追加することで変更可能です。

--chat-template-kwargs '{\"reasoning_effort\":\"xhigh\"}'

引数に--spec-type draft-mtp --spec-draft-n-max 2と追加することでMTPが有効化されます。<br>

なお、--spec-draft-n-maxの値については、日本語では2がスイートスポットのようです。

License

Apache 2.0

Developer

Alibaba Cloud

Run grapeV-ai/Qwen3.8-27B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models