Model Intelligence Sheet
nuofang/Qwen3.5-4B-SOMPOA-heresy-v2-GGUF overview
Auto Quantized GGUF Model This repository contains automated GGUF quantization files for MuXodious/Qwen3.5 4B SOMPOA heresy v2 https://huggingface.co/MuXodious…
Runs locally from ~3.5 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
4 GGUF files detected
Direct downloads for local inference
Model Details
Model README
---
base_model: MuXodious/Qwen3.5-4B-SOMPOA-heresy-v2
tags:
- llama.cpp
- quantized
- imatrix
---
Auto-Quantized GGUF Model
This repository contains automated GGUF quantization files for MuXodious/Qwen3.5-4B-SOMPOA-heresy-v2.
The calibration data for the imatrix is targeted at Chinese novels and role-playing (RP), while preserving logic and common sense.
imatrix 的校准数据以中文的小说、角色扮演为目标,同时保留逻辑和常识。
📊 Perplexity Evaluation
(Tested against the provided calibration dataset)
- Base (F16/BF16): PPL = 16.6021 +/- 0.14031
- Q4_K_M: PPL = 13.9404 +/- 0.11399
- Q5_K_M: PPL = 13.8567 +/- 0.11368
Run nuofang/Qwen3.5-4B-SOMPOA-heresy-v2-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models