Model Intelligence Sheet
canhdu/Qwen3.6-27B-DFlash-GGUF overview
This repository stores Z Lab's Qwen3.6 27B DFlash drafter quantized to Q4 K M and Q8 0 using upstream llama.cpp and following spiritbuun's guide https://huggin…
Runs locally from ~985.2 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
license: mit
base_model:
- z-lab/Qwen3.6-27B-DFlash
tags:
- dflash
- speculative-decoding
- diffusion
- efficiency
- flash-decoding
- qwen
- diffusion-language-model
---
This repository stores Z Lab's Qwen3.6 27B DFlash drafter quantized to Q4_K_M and Q8_0 using upstream llama.cpp and following spiritbuun's guide.
Run canhdu/Qwen3.6-27B-DFlash-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models