GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF overview

Jackrong/Qwen3.5 9B DeepSeek V4 Flash MTP GGUF Source model: Jackrong/Qwen3.5 9B DeepSeek V4 Flash MTP source fallback: unsloth/Qwen3.5 9B Uploaded GGUF varian…

ggufbase_model:Jackrong/Qwen3.5-9B-DeepSeek-V4-Flashbase_model:quantized:Jackrong/Qwen3.5-9B-DeepSeek-V4-Flashlicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~4.41 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
12,677
Likes
30
Pipeline
Author

Repository Files & Downloads

5 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_M.ggufGGUFQ3_K_M4.41 GBDownload
Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q4_K_M.ggufGGUFQ4_K_M5.38 GBDownload
Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q5_K_M.ggufGGUFQ5_K_M6.19 GBDownload
Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q6_K.ggufGGUFQ6_K7.04 GBDownload
Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q8_0.ggufGGUFQ8_09.11 GBDownload

Model Details

Model IDJackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF
AuthorJackrong
Pipeline
Licenseother
Base modelJackrong/Qwen3.5-9B-DeepSeek-V4-Flash,unsloth/Qwen3.5-9B
Last modified2026-07-09T03:06:49.000Z

Model README

---

license: other

base_model:

  • Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash
  • unsloth/Qwen3.5-9B

private: true

---

Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF

Source model: Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash

MTP source fallback: unsloth/Qwen3.5-9B

Uploaded GGUF variants:

  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q2_K.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_S.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_M.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_L.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-IQ4_XS.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q4_K_S.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q4_K_M.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q5_K_S.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q5_K_M.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q6_K.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q8_0.gguf
  • Qwen3.5-9B-DeepSeek-V4-Flash-MTP-BF16.gguf

The conversion pipeline first verifies whether the source HF model already contains MTP tensors.

If not, it extracts the MTP tensors from the matching unsloth base model and injects them into the safetensors index before GGUF conversion.

Run Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models