GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

magnitudedev/Qwen3.6-35B-A3B-DFlash-GGUF overview

Qwen3.6 35B A3B DFlash GGUF Q8 0 GGUF conversion of z lab/Qwen3.6 35B A3B DFlash https://huggingface.co/z lab/Qwen3.6 35B A3B DFlash for use with Magnitude htt…

llama.cppggufdflashspeculative-decodingmagnitudebase_model:z-lab/Qwen3.6-35B-A3B-DFlashbase_model:quantized:z-lab/Qwen3.6-35B-A3B-DFlashlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~401.6 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.6-35B-A3B-DFlash-Q8_0.ggufGGUFQ8_0401.6 MBDownload

Model Details

Model IDmagnitudedev/Qwen3.6-35B-A3B-DFlash-GGUF
Authormagnitudedev
Pipeline
Licenseapache-2.0
Base modelz-lab/Qwen3.6-35B-A3B-DFlash
Last modified2026-08-13T01:59:23.000Z

Model README

---

base_model:

- z-lab/Qwen3.6-35B-A3B-DFlash

license: apache-2.0

library_name: llama.cpp

tags:

- gguf

- dflash

- speculative-decoding

- magnitude

---

Qwen3.6-35B-A3B-DFlash GGUF

Q8_0 GGUF conversion of

z-lab/Qwen3.6-35B-A3B-DFlash

for use with Magnitude.

Converted with

llama.cpp

at revision 8e7f22b67ef4667b4ddd50230771287f328cfb3f.

Artifact

  • Quantization: Q8_0
  • SHA-256: db1e23ddc5b68be4183af6fd2414f51c3f2ba19db59e4af74becb5291b4801a1

Run magnitudedev/Qwen3.6-35B-A3B-DFlash-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models