GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Anbeeld/gemma-4-31B-it-DFlash-GGUF overview

Gemma 4 31B IT DFlash GGUF llama.cpp quantizations of z lab DFlash draft model https://huggingface.co/z lab/gemma 4 31B it DFlash for Gemma 4 31B IT https://hu…

transformersggufsafetensorsqwen3dflashspeculative-decodingblock-diffusiondraft-modelefficiencyqwengemmadiffusion-language-modeltext-generationarxiv:2602.06036license:apache-2.0text-generation-inferenceendpoints_compatibleregion:usbase_model:z-lab/gemma-4-31B-it-DFlashbase_model:quantized:z-lab/gemma-4-31B-it-DFlashconversational

Runs locally from ~551.4 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
5,133
Likes
6
Pipeline
text-generation
Author

Repository Files & Downloads

7 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
gemma4-31b-it-dflash-Q2_K.ggufGGUFQ2_K551.4 MBDownload
gemma4-31b-it-dflash-Q3_K_M.ggufGGUFQ3_K_M714.0 MBDownload
gemma4-31b-it-dflash-Q4_K_M.ggufGGUFQ4_K_M870.4 MBDownload
gemma4-31b-it-dflash-Q5_K_M.ggufGGUFQ5_K_M1.01 GBDownload
gemma4-31b-it-dflash-Q6_K.ggufGGUFQ6_K1.19 GBDownload
gemma4-31b-it-dflash-Q8_0.ggufGGUFQ8_01.53 GBDownload
gemma4-31b-it-dflash-bf16.ggufGGUFBF162.88 GBDownload

Model Details

Model IDAnbeeld/gemma-4-31B-it-DFlash-GGUF
AuthorAnbeeld
Pipelinetext-generation
License
Base modelz-lab/gemma-4-31B-it-DFlash
Last modified2026-07-19T02:30:05.000Z

Model README

---

base_model: z-lab/gemma-4-31B-it-DFlash

tags:

  • transformers
  • safetensors
  • qwen3
  • dflash
  • speculative-decoding
  • block-diffusion
  • draft-model
  • efficiency
  • qwen
  • gemma
  • diffusion-language-model
  • text-generation
  • arxiv:2602.06036
  • license:apache-2.0
  • text-generation-inference
  • endpoints_compatible
  • region:us

---

Gemma 4 31B IT DFlash GGUF

llama.cpp quantizations of z-lab DFlash draft model for Gemma 4 31B IT.

Run Anbeeld/gemma-4-31B-it-DFlash-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models