GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/DeepSeek-V4-Flash-Vision-Exp-GGUF overview

DeepSeek V4 Flash Vision Exp Run with https://llama.app bash llama serve hf ggml org/DeepSeek V4 Flash Vision Exp GGUF Source models https://huggingface.co/dee…

ggufquantizedimage-text-to-textbase_model:deepseek-ai/DeepSeek-V4-Flash-Vision-Expbase_model:quantized:deepseek-ai/DeepSeek-V4-Flash-Vision-Explicense:mitendpoints_compatibleregion:usconversational

Runs locally from ~5.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
3
Pipeline
image-text-to-text
Author

Repository Files & Downloads

10 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
DeepSeek-V4-Flash-Vision-Exp-MXFP4-00001-of-00002.ggufGGUFGGUF5.0 MBDownload
DeepSeek-V4-Flash-Vision-Exp-MXFP4-00002-of-00002.ggufGGUFGGUF144.34 GBDownload
DeepSeek-V4-Flash-Vision-Exp-Q2_K-00001-of-00002.ggufGGUFQ2_K5.0 MBDownload
DeepSeek-V4-Flash-Vision-Exp-Q2_K-00002-of-00002.ggufGGUFQ2_K109.29 GBDownload
DeepSeek-V4-Flash-Vision-Exp-Q2_K_S-00001-of-00002.ggufGGUFQ2_K_S5.0 MBDownload
DeepSeek-V4-Flash-Vision-Exp-Q2_K_S-00002-of-00002.ggufGGUFQ2_K_S91.82 GBDownload
dspark-DeepSeek-V4-Flash-Vision-Exp-BF16.ggufGGUFBF1610.15 GBDownload
dspark-DeepSeek-V4-Flash-Vision-Exp-MXFP4.ggufGGUFGGUF10.08 GBDownload
mmproj-DeepSeek-V4-Flash-Vision-Exp-BF16.ggufGGUFBF16891.2 MBDownload
mmproj-DeepSeek-V4-Flash-Vision-Exp-Q8_0.ggufGGUFQ8_0474.9 MBDownload

Model Details

Model IDggml-org/DeepSeek-V4-Flash-Vision-Exp-GGUF
Authorggml-org
Pipelineimage-text-to-text
Licensemit
Base modeldeepseek-ai/DeepSeek-V4-Flash-Vision-Exp
Last modified2026-09-02T21:00:42.000Z

Model README

---

license: mit

pipeline_tag: image-text-to-text

tags:

  • gguf
  • quantized

base_model:

  • deepseek-ai/DeepSeek-V4-Flash-Vision-Exp

---

DeepSeek-V4-Flash-Vision-Exp

Run with https://llama.app

llama serve -hf ggml-org/DeepSeek-V4-Flash-Vision-Exp-GGUF

Source models

  • https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-Vision-Exp

Notes

  • Includes a Q8_0 mmproj for the vision encoder.
  • Currently, the Q2 models do not use an imatrix calibration due to lack of one.

TODOs

  • add info

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/DeepSeek-V4-Flash-Vision-Exp-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models