GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Latentiq/Qwen3-4B-GGUF overview

Qwen3 4B GGUF This repository provides a GGUF quantized version of Qwen3 4B. The original model was developed by Alibaba Cloud through the Qwen team. The GGUF …

ggufbase_model:Qwen/Qwen3-4Bbase_model:quantized:Qwen/Qwen3-4Blicense:apache-2.0endpoints_compatibleregion:usimatrixconversational

Runs locally from ~2.33 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
qwen3-4b-q4_k_m.ggufGGUFQ4_K_M2.33 GBDownload

Model Details

Model IDLatentiq/Qwen3-4B-GGUF
AuthorLatentiq
Pipeline
Licenseapache-2.0
Base modelQwen/Qwen3-4B,unsloth/Qwen3-4B-GGUF
Last modified2026-06-25T12:59:03.000Z

Model README

---

license: apache-2.0

base_model:

- Qwen/Qwen3-4B

- unsloth/Qwen3-4B-GGUF

---

Qwen3 4B GGUF

This repository provides a GGUF quantized version of Qwen3 4B.

The original model was developed by Alibaba Cloud through the Qwen team.

The GGUF quantization was created by Unsloth.

This repository redistributes the model for convenient use in GGUF compatible inference frameworks.

Attribution

Qwen3 4B © Alibaba Cloud, Qwen team

GGUF quantization by Unsloth

Sources

Original model: https://huggingface.co/Qwen/Qwen3-4B

GGUF quantization: https://huggingface.co/unsloth/Qwen3-4B-GGUF

License

This repository is distributed under the Apache 2.0 License, subject to the terms and conditions of the original model and any applicable upstream notices.

Notes

This repository does not claim ownership of the original model.

Please refer to the upstream repositories for full documentation, usage details, and any additional requirements.

Run Latentiq/Qwen3-4B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models