GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Dsg2/Qwen3.5-4B-Pruned-Coder-UD-Q6_K_XL.gguf overview

Qwen3.5 4B Pruned Coder UD Q6 K XL: A pruned + quantization aware fine tuned coding agent Trimmed 17% of neurons. Light bench: | model | size | prefill t/s | d…

ggufcodetext-generationendataset:Dsg2/CodeMixbase_model:unsloth/Qwen3.5-4B-GGUFbase_model:quantized:unsloth/Qwen3.5-4B-GGUFlicense:apache-2.0endpoints_compatibleregion:usimatrixconversational

Runs locally from ~3.65 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.5-4B-Pruned-UD-Q6_K_XL.ggufGGUFQ6_K_XL3.65 GBDownload

Model Details

Model IDDsg2/Qwen3.5-4B-Pruned-Coder-UD-Q6_K_XL.gguf
AuthorDsg2
Pipelinetext-generation
Licenseapache-2.0
Base modelunsloth/Qwen3.5-4B-GGUF
Last modified2026-06-21T08:40:54.000Z

Model README

---

license: apache-2.0

datasets:

  • Dsg2/CodeMix

language:

  • en

base_model:

  • unsloth/Qwen3.5-4B-GGUF

pipeline_tag: text-generation

tags:

  • code

---

Qwen3.5-4B-Pruned-Coder-UD-Q6_K_XL: A pruned + quantization aware fine tuned coding agent

Trimmed 17% of neurons.

Light bench:

| model | size | prefill t/s | decode t/s | code | instruct | prose | tool |

|---|---|---|---|---|---|---|---|

| 4B base | 3.86 GiB | 11.1 | 4.77 | 3.13 | 2.52 | 1.73 | 8.16 |

| 4B-Pruned | 3.65 GiB | 12.2 | 5.51 | 3.51 | 2.81 | 9.32 | 5.82 |

Calibrated on ~63k code/instruct/tool call tokens for 2 CPU hours.

Run Dsg2/Qwen3.5-4B-Pruned-Coder-UD-Q6_K_XL.gguf with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models