GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

RockMan256/Qwen3.5-2B-home16k-GGUF overview

Qwen3.5 2B home16k GGUF GGUF quantizations of RockMan256/Qwen3.5 2B home16k https://huggingface.co/RockMan256/Qwen3.5 2B home16k . Converted using llama.cpp ht…

llama-cppggufbase_model:RockMan256/Qwen3.5-2B-home16kbase_model:quantized:RockMan256/Qwen3.5-2B-home16kendpoints_compatibleregion:usconversational

Runs locally from ~637.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
mmproj-model-f16.ggufGGUFF16637.3 MBDownload
qwen3.5-2B-home16k-Q4_K_M.ggufGGUFQ4_K_M1.22 GBDownload

Model Details

Model IDRockMan256/Qwen3.5-2B-home16k-GGUF
AuthorRockMan256
Pipeline
License
Base modelRockMan256/Qwen3.5-2B-home16k
Last modified2026-07-11T09:52:41.000Z

Model README

---

base_model: RockMan256/Qwen3.5-2B-home16k

library_name: llama-cpp

---

Qwen3.5-2B-home16k-GGUF

GGUF quantizations of RockMan256/Qwen3.5-2B-home16k.

Converted using llama.cpp.

Files

| File | Type | Size |

|------|------|------|

| qwen3.5-2B-home16k-Q4_K_M.gguf | Text model (Q4_K_M) | 1.3 GB |

| mmproj-model-f16.gguf | Vision projector (F16) | 638 MB |

Usage

llama.cpp CLI

./llama-server \
  -m qwen3.5-2B-home16k-Q4_K_M.gguf \
  --mmproj mmproj-model-f16.gguf \
  --host 0.0.0.0 --port 8080

Open WebUI

  1. Go to Settings > Connections > Model Ollama
  2. Set the API URL to your llama.cpp server endpoint
  3. Upload both qwen3.5-2B-home16k-Q4_K_M.gguf and mmproj-model-f16.gguf

Source Model

Run RockMan256/Qwen3.5-2B-home16k-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models