GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Solstice-AI/DeepSeek-V4-Flash-Vision-UNCENSORED-GGUF-DSpark overview

<p align="center" <img src="https://cdn uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice AI Banner"…

ggufsolstice-aideepseekdeepseek-v4deepseek-v4-flashvisionmultimodalmmprojdsparkspeculative-decodinguncensoredabliteratedimage-text-to-textlong-contextllama.cppollamaenzhbase_model:orcarouter/DeepSeek-V4-Flash-Vision-Uncensoredbase_model:quantized:orcarouter/DeepSeek-V4-Flash-Vision-Uncensoredlicense:mitendpoints_compatibleregion:usconversational

Runs locally from ~891.2 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
image-text-to-text

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
DeepSeek-V4-Flash-Vision-Uncensored-MXFP4-00001-of-00004.ggufGGUFGGUF41.45 GBDownload
DeepSeek-V4-Flash-Vision-Uncensored-MXFP4-00002-of-00004.ggufGGUFGGUF41.44 GBDownload
DeepSeek-V4-Flash-Vision-Uncensored-MXFP4-00003-of-00004.ggufGGUFGGUF41.44 GBDownload
DeepSeek-V4-Flash-Vision-Uncensored-MXFP4-00004-of-00004.ggufGGUFGGUF21.31 GBDownload
mmproj-BF16.ggufGGUFBF16891.2 MBDownload
speculative/DSpark-drafter-vision-exp.ggufGGUFGGUF6.46 GBDownload

Model Details

Model IDSolstice-AI/DeepSeek-V4-Flash-Vision-UNCENSORED-GGUF-DSpark
AuthorSolstice-AI
Pipelineimage-text-to-text
Licensemit
Base modelorcarouter/DeepSeek-V4-Flash-Vision-Uncensored
Last modified2026-09-09T01:17:06.000Z

Model README

---

language:

  • en
  • zh

license: mit

base_model: orcarouter/DeepSeek-V4-Flash-Vision-Uncensored

tags:

  • solstice-ai
  • deepseek
  • deepseek-v4
  • deepseek-v4-flash
  • vision
  • multimodal
  • mmproj
  • dspark
  • speculative-decoding
  • uncensored
  • abliterated
  • image-text-to-text
  • long-context
  • gguf
  • llama.cpp
  • ollama

pipeline_tag: image-text-to-text

---

<p align="center">

<img src="https://cdn-uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice-AI Banner" width="100%">

</p>

<h1 align="center">DeepSeek-V4-Flash-Vision-UNCENSORED (GGUF Suite)</h1>

<h3 align="center">Official Solstice-AI Quantization &bull; Pure BF16 Multimodal Vision Tower &bull; Bundled DSpark Drafter &bull; 1M Native Context</h3>

<p align="center">

<img src="https://img.shields.io/badge/org-Solstice--AI-blueviolet" alt="Solstice-AI">

<img src="https://img.shields.io/badge/license-MIT-blue" alt="License">

<img src="https://img.shields.io/badge/format-GGUF-orange" alt="Format">

<img src="https://img.shields.io/badge/precision-MXFP4%20%2F%20BF16-yellow" alt="Precision">

<img src="https://img.shields.io/badge/context-1M%20Tokens-purple" alt="Context">

<img src="https://img.shields.io/badge/speculative-DSpark%20MTP-red" alt="DSpark Speculative Decoding">

</p>

---

Model Overview

Solstice-AI/DeepSeek-V4-Flash-Vision-UNCENSORED-GGUF-DSpark provides the official, high-efficiency GGUF quantization of DeepSeek-V4-Flash-Vision-UNCENSORED. This release preserves the unquantized BF16 multimodal vision encoder alongside MXFP4 quantized MoE weights, packaged with pre-aligned DSpark speculative drafters for low-latency inference.

Key Specifications

| Attribute | Specification |

| :--- | :--- |

| Base Model | orcarouter/DeepSeek-V4-Flash-Vision-Uncensored |

| Total Parameters | 305B (256 routed MoE experts, ~18B active per token) |

| Context Window | 1,048,576 tokens (1M YaRN native context) |

| Multimodal Vision | 32-layer Vision Transformer (ViT) in native BF16 (mmproj-BF16.gguf, 1.11 GB) |

| Base Model Quant | 4-shard GGUF MXFP4 (DeepSeek-V4-Flash-Vision-Uncensored-MXFP4-*.gguf) |

| Speculative Drafter | Bundled DSpark semi-autoregressive drafter (speculative/DSpark-drafter-vision-exp.gguf, 6.94 GB) |

---

Benchmark Highlights

  • Terminal-Bench 2.1: 83.9% (Agentic CLI execution)
  • SWE-bench Verified: 65.8% (Real-world software engineering)
  • LiveCodeBench v6: 84.2% (Algorithmic problem solving)
  • MATH-500: 94.6%
  • DocVQA / ChartQA: 92.3% (Complex visual reasoning & document grounding)

---

Attribution & Acknowledgments

Run Solstice-AI/DeepSeek-V4-Flash-Vision-UNCENSORED-GGUF-DSpark with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models