GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

nuofang/Qwen3.5-9B-SOMPOA-heresy-GGUF overview

Auto Quantized GGUF Model This repository contains automated GGUF quantization files for MuXodious/Qwen3.5 9B SOMPOA heresy https://huggingface.co/MuXodious/Qw…

ggufllama.cppquantizedimatrixbase_model:MuXodious/Qwen3.5-9B-SOMPOA-heresybase_model:quantized:MuXodious/Qwen3.5-9B-SOMPOA-heresyendpoints_compatibleregion:usconversational

Runs locally from ~4.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.5-9B-SOMPOA-heresy-IQ4_XS.ggufGGUFIQ4_XS4.84 GBDownload
imatrix.ggufGGUFGGUF4.9 MBDownload
mmproj-Qwen3.5-9B-SOMPOA-heresy-f16.ggufGGUFF16879.0 MBDownload

Model Details

Model IDnuofang/Qwen3.5-9B-SOMPOA-heresy-GGUF
Authornuofang
Pipeline
License
Base modelMuXodious/Qwen3.5-9B-SOMPOA-heresy
Last modified2026-07-07T14:52:13.000Z

Model README

---

base_model: MuXodious/Qwen3.5-9B-SOMPOA-heresy

tags:

  • llama.cpp
  • quantized
  • imatrix

---

Auto-Quantized GGUF Model

This repository contains automated GGUF quantization files for MuXodious/Qwen3.5-9B-SOMPOA-heresy.

The calibration data for the imatrix is targeted at Chinese novels and role-playing (RP), while preserving logic and common sense.

imatrix 的校准数据以中文的小说、角色扮演为目标,同时保留逻辑和常识。

📊 Perplexity Evaluation

(Tested against the provided calibration dataset)

  • Base (F16/BF16): PPL = 14.1755 +/- 0.11558
  • IQ4_XS: PPL = 12.0099 +/- 0.09481

Run nuofang/Qwen3.5-9B-SOMPOA-heresy-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models