GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ridd1er1/Qwen3.6-35B-A3B-DFlash-GGUF overview

Qwen 3.6 35B A3B DFlash GGUF llama.cpp quantizations of z lab DFlash draft model https://huggingface.co/z lab/Qwen3.6 35B A3B DFlash for Qwen 3.6 35B A3B https…

ggufqwen3custom_codebase_model:z-lab/Qwen3.6-35B-A3B-DFlashbase_model:quantized:z-lab/Qwen3.6-35B-A3B-DFlashendpoints_compatibleregion:usconversational

Runs locally from ~253.8 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
qwen36-35b-a3b-dflash-IQ4_XS.ggufGGUFIQ4_XS253.8 MBDownload
qwen36-35b-a3b-dflash-Q4_K_M.ggufGGUFQ4_K_M278.2 MBDownload
qwen36-35b-a3b-dflash-Q5_K_M.ggufGGUFQ5_K_M328.2 MBDownload
qwen36-35b-a3b-dflash-Q6_K.ggufGGUFQ6_K381.4 MBDownload
qwen36-35b-a3b-dflash-Q8_0.ggufGGUFQ8_0490.8 MBDownload
qwen36-35b-a3b-dflash-bf16.ggufGGUFBF16914.6 MBDownload

Model Details

Model IDridd1er1/Qwen3.6-35B-A3B-DFlash-GGUF
Authorridd1er1
Pipeline
License
Base modelz-lab/Qwen3.6-35B-A3B-DFlash
Last modified2026-07-01T20:44:06.000Z

Model README

---

base_model: z-lab/Qwen3.6-35B-A3B-DFlash

---

Qwen 3.6 35B A3B DFlash GGUF

llama.cpp quantizations of z-lab DFlash draft model for Qwen 3.6 35B A3B.

Use with BeeLlama.cpp — a llama.cpp fork with advanced DFlash support that enables using these draft models to their full potential.

Run ridd1er1/Qwen3.6-35B-A3B-DFlash-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models