Model Intelligence Sheet
sm54/deepseek-v4-flash-0731-gguf overview
Built to run with antirez / ds4 inference engine. Not tested or built for llama cpp. Q4K available with dspark drafter.
Runs locally from ~5.58 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
base_model:
- deepseek-ai/DeepSeek-V4-Flash-0731
---
Built to run with antirez / ds4 inference engine. Not tested or built for llama cpp. Q4K available with dspark drafter.
Run sm54/deepseek-v4-flash-0731-gguf with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models