GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

0xSero/DeepSeek-V4-Flash-Spark-Mini-GGUF overview

DeepSeek V4 Flash Spark Mini GGUF Q3 Dynamic REAP GGUF quantization of deepseek ai/DeepSeek V4 Flash 0731 https://huggingface.co/deepseek ai/DeepSeek V4 Flash …

ggufdeepseek-v4llama.cppreapquantizedtext-generationbase_model:deepseek-ai/DeepSeek-V4-Flash-0731base_model:quantized:deepseek-ai/DeepSeek-V4-Flash-0731license:mitendpoints_compatibleregion:usconversational

Runs locally from ~69.86 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).

Downloads
223
Likes
7
Pipeline
text-generation
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.ggufGGUFQ369.86 GBDownload

Model Details

Model ID0xSero/DeepSeek-V4-Flash-Spark-Mini-GGUF
Author0xSero
Pipelinetext-generation
Licensemit
Base modeldeepseek-ai/DeepSeek-V4-Flash-0731
Last modified2026-09-01T11:26:11.000Z

Model README

---

license: mit

base_model:

  • deepseek-ai/DeepSeek-V4-Flash-0731

base_model_relation: quantized

pipeline_tag: text-generation

library_name: gguf

tags:

  • deepseek-v4
  • gguf
  • llama.cpp
  • reap
  • quantized

---

DeepSeek-V4-Flash-Spark-Mini GGUF (Q3-Dynamic REAP)

GGUF quantization of deepseek-ai/DeepSeek-V4-Flash-0731 (MIT),

Mini variant of the REAP-pruned Spark lineage with dynamic per-layer Q3 allocation.

Files

  • DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf — single-file GGUF
  • DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf.sha256 — checksum sidecar

Usage

Compatible with llama.cpp and other GGUF loaders:

llama-server -m DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf

License

MIT, inherited from the upstream DeepSeek-V4-Flash-0731 repository.

See the upstream model card for full terms.

Acknowledgements

  • DeepSeek for the upstream DeepSeek-V4-Flash-0731 (MIT) this quantization derives from.
  • The llama.cpp project and community for the GGUF format and runtime support.

Run 0xSero/DeepSeek-V4-Flash-Spark-Mini-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models