Model Intelligence Sheet
ddh0/DeepSeek-V4-Flash-GGUF overview
GGUF quantizations of DeepSeek V4 Flash. Using MTP requires am17an/llama.cpp:dsv4 mtp https://github.com/am17an/llama.cpp/tree/dsv4 mtp until llama.cpp 25784 h…
Runs locally from ~3.92 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
4 GGUF files detected
Direct downloads for local inference
Model Details
Model README
---
base_model:
- deepseek-ai/DeepSeek-V4-Flash
---
GGUF quantizations of DeepSeek-V4-Flash.
Using MTP requires am17an/llama.cpp:dsv4-mtp until llama.cpp#25784 is merged.
Run ddh0/DeepSeek-V4-Flash-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models