Model Intelligence Sheet
giocom/Qwen3.6-35B-A3B-DFlash-GGUF overview
Qwen3.6 35B A3B DFlash
Runs locally from ~224.8 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
| Model ID | giocom/Qwen3.6-35B-A3B-DFlash-GGUF |
|---|---|
| Author | giocom |
| Pipeline | text-generation |
| License | apache-2.0 |
| Base model | z-lab/Qwen3.6-35B-A3B-DFlash |
| Last modified | 2026-07-20T00:12:54.000Z |
Model README
---
pipeline_tag: text-generation
library_name: transformers
base_model:
- z-lab/Qwen3.6-35B-A3B-DFlash
license: apache-2.0
inference: false
tags:
- dflash
- speculative-decoding
- speculative-decoding-draft
- block-diffusion
- draft-model
- diffusion-language-model
- efficiency
- qwen
- qwen3
- qwen3.6
- sglang
---
Qwen3.6-35B-A3B-DFlash
Run giocom/Qwen3.6-35B-A3B-DFlash-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models