Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF overview
Jackrong/Qwen3.5 9B DeepSeek V4 Flash MTP GGUF Source model: Jackrong/Qwen3.5 9B DeepSeek V4 Flash MTP source fallback: unsloth/Qwen3.5 9B Uploaded GGUF varian…
Runs locally from ~4.41 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_M.gguf | GGUF | Q3_K_M | 4.41 GB | Download |
| Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q4_K_M.gguf | GGUF | Q4_K_M | 5.38 GB | Download |
| Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q5_K_M.gguf | GGUF | Q5_K_M | 6.19 GB | Download |
| Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q6_K.gguf | GGUF | Q6_K | 7.04 GB | Download |
| Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q8_0.gguf | GGUF | Q8_0 | 9.11 GB | Download |
Model Details
Model README
---
license: other
base_model:
- Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash
- unsloth/Qwen3.5-9B
private: true
---
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF
Source model: Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash
MTP source fallback: unsloth/Qwen3.5-9B
Uploaded GGUF variants:
Qwen3.5-9B-DeepSeek-V4-Flash-MTP-Q2_K.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_S.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_M.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q3_K_L.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-IQ4_XS.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q4_K_S.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q4_K_M.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q5_K_S.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q5_K_M.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q6_K.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-Q8_0.ggufQwen3.5-9B-DeepSeek-V4-Flash-MTP-BF16.gguf
The conversion pipeline first verifies whether the source HF model already contains MTP tensors.
If not, it extracts the MTP tensors from the matching unsloth base model and injects them into the safetensors index before GGUF conversion.
Run Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models