GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ecwu/qwen35-4b-423333-lora-gguf-Q4_K_M overview

Qwen3.5 4B 423333 LoRA merged GGUF This repository contains a GGUF export of ecwu/qwen35 4b unsloth 423333 , merged into unsloth/Qwen3.5 4B . Conversion notes:…

llama.cppggufqwen3.5loramergedzhenbase_model:unsloth/Qwen3.5-4Bbase_model:adapter:unsloth/Qwen3.5-4Bendpoints_compatibleregion:usconversational

Runs locally from ~2.52 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
qwen35-4b-Q4_K_M.ggufGGUFQ4_K_M2.52 GBDownload

Model Details

Model IDecwu/qwen35-4b-423333-lora-gguf-Q4_K_M
Authorecwu
Pipeline
License
Base modelunsloth/Qwen3.5-4B
Last modified2026-07-02T13:46:43.000Z

Model README

---

language:

  • zh
  • en

base_model: unsloth/Qwen3.5-4B

tags:

  • gguf
  • qwen3.5
  • lora
  • merged

library_name: llama.cpp

---

Qwen3.5 4B 423333 LoRA merged GGUF

This repository contains a GGUF export of ecwu/qwen35-4b-unsloth-423333, merged into unsloth/Qwen3.5-4B.

Conversion notes:

  • LoRA was merged with the Qwen3.5 conditional generation architecture, matching adapter keys under model.language_model.*.
  • GGUF conversion used llama.cpp commit 4fc4ec554.
  • GGUF metadata was corrected to qwen35.block_count = 32 and qwen35.nextn_predict_layers = 0, because the exported language model does not include MTP tensors.

Source adapter: https://huggingface.co/ecwu/qwen35-4b-unsloth-423333

Base model: https://huggingface.co/unsloth/Qwen3.5-4B

File: qwen35-4b-Q4_K_M.gguf

Quantization: Q4_K_M

Run ecwu/qwen35-4b-423333-lora-gguf-Q4_K_M with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models