0xSero/DeepSeek-V4-Flash-Spark-Mini-GGUF overview
DeepSeek V4 Flash Spark Mini GGUF Q3 Dynamic REAP GGUF quantization of deepseek ai/DeepSeek V4 Flash 0731 https://huggingface.co/deepseek ai/DeepSeek V4 Flash …
Runs locally from ~69.86 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf | GGUF | Q3 | 69.86 GB | Download |
Model Details
| Model ID | 0xSero/DeepSeek-V4-Flash-Spark-Mini-GGUF |
|---|---|
| Author | 0xSero |
| Pipeline | text-generation |
| License | mit |
| Base model | deepseek-ai/DeepSeek-V4-Flash-0731 |
| Last modified | 2026-09-01T11:26:11.000Z |
Model README
---
license: mit
base_model:
- deepseek-ai/DeepSeek-V4-Flash-0731
base_model_relation: quantized
pipeline_tag: text-generation
library_name: gguf
tags:
- deepseek-v4
- gguf
- llama.cpp
- reap
- quantized
---
DeepSeek-V4-Flash-Spark-Mini GGUF (Q3-Dynamic REAP)
GGUF quantization of deepseek-ai/DeepSeek-V4-Flash-0731 (MIT),
Mini variant of the REAP-pruned Spark lineage with dynamic per-layer Q3 allocation.
Files
DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf— single-file GGUFDeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf.sha256— checksum sidecar
Usage
Compatible with llama.cpp and other GGUF loaders:
llama-server -m DeepSeek-V4-Flash-Spark-Mini-Q3-Dynamic-REAP-ds4.gguf
License
MIT, inherited from the upstream DeepSeek-V4-Flash-0731 repository.
See the upstream model card for full terms.
Acknowledgements
- DeepSeek for the upstream DeepSeek-V4-Flash-0731 (MIT) this quantization derives from.
- The llama.cpp project and community for the GGUF format and runtime support.
Run 0xSero/DeepSeek-V4-Flash-Spark-Mini-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models