GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Auguments/Josiefied-Qwen3.5-2B-gabliterated-v1-MTP-GGUF-BF16 overview

Josiefied Qwen3.5 2B gabliterated v1 Highest quality GGUF BF16. In this repo: BF16 conversion, and Q6 K converted from BF16 MTP block included For mmproj, use …

ggufquantizedbase_model:Goekdeniz-Guelmez/Josiefied-Qwen3.5-2B-gabliterated-v1base_model:quantized:Goekdeniz-Guelmez/Josiefied-Qwen3.5-2B-gabliterated-v1endpoints_compatibleregion:usconversational

Runs locally from ~1.50 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Josiefied-Qwen3.5-2B-gabliterated-v1-MTP-BF16.ggufGGUFBF163.63 GBDownload
Josiefied-Qwen3.5-2B-gabliterated-v1-MTP-Q6_K.ggufGGUFQ6_K1.50 GBDownload

Model Details

Model IDAuguments/Josiefied-Qwen3.5-2B-gabliterated-v1-MTP-GGUF-BF16
AuthorAuguments
Pipeline
License
Base modelGoekdeniz-Guelmez/Josiefied-Qwen3.5-2B-gabliterated-v1
Last modified2026-08-05T12:47:06.000Z

Model README

---

base_model: Goekdeniz-Guelmez/Josiefied-Qwen3.5-2B-gabliterated-v1

base_model_relation: quantized

library_name: gguf

tags:

- gguf

- quantized

---

Josiefied-Qwen3.5-2B-gabliterated-v1

Highest quality GGUF - BF16.

In this repo:

  • BF16 conversion, and Q6_K converted from BF16
  • MTP block included

For mmproj, use unsloth's or bartowski's.

Converted using base llama.cpp b9294. Never versions can't convert this particular repo to GGUF due to regressed Transformers.

Run Auguments/Josiefied-Qwen3.5-2B-gabliterated-v1-MTP-GGUF-BF16 with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models