GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

nuofang/Qwen3.5-4B-SOMPOA-heresy-v2-GGUF overview

Auto Quantized GGUF Model This repository contains automated GGUF quantization files for MuXodious/Qwen3.5 4B SOMPOA heresy v2 https://huggingface.co/MuXodious…

ggufllama.cppquantizedimatrixbase_model:MuXodious/Qwen3.5-4B-SOMPOA-heresy-v2base_model:quantized:MuXodious/Qwen3.5-4B-SOMPOA-heresy-v2endpoints_compatibleregion:usconversational

Runs locally from ~3.5 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

4 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.5-4B-SOMPOA-heresy-v2-Q4_K_M.ggufGGUFQ4_K_M2.52 GBDownload
Qwen3.5-4B-SOMPOA-heresy-v2-Q5_K_M.ggufGGUFQ5_K_M2.86 GBDownload
imatrix.ggufGGUFGGUF3.5 MBDownload
mmproj-Qwen3.5-4B-SOMPOA-heresy-v2-f16.ggufGGUFF16644.3 MBDownload

Model Details

Model IDnuofang/Qwen3.5-4B-SOMPOA-heresy-v2-GGUF
Authornuofang
Pipeline
License
Base modelMuXodious/Qwen3.5-4B-SOMPOA-heresy-v2
Last modified2026-07-07T12:39:10.000Z

Model README

---

base_model: MuXodious/Qwen3.5-4B-SOMPOA-heresy-v2

tags:

  • llama.cpp
  • quantized
  • imatrix

---

Auto-Quantized GGUF Model

This repository contains automated GGUF quantization files for MuXodious/Qwen3.5-4B-SOMPOA-heresy-v2.

The calibration data for the imatrix is targeted at Chinese novels and role-playing (RP), while preserving logic and common sense.

imatrix 的校准数据以中文的小说、角色扮演为目标,同时保留逻辑和常识。

📊 Perplexity Evaluation

(Tested against the provided calibration dataset)

  • Base (F16/BF16): PPL = 16.6021 +/- 0.14031
  • Q4_K_M: PPL = 13.9404 +/- 0.11399
  • Q5_K_M: PPL = 13.8567 +/- 0.11368

Run nuofang/Qwen3.5-4B-SOMPOA-heresy-v2-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models