nuofang/Qwen3.5-9B-ultra-uncensored-heretic-v2-GGUF overview
Auto Quantized GGUF Model This repository contains automated GGUF quantization files for llmfan46/Qwen3.5 9B ultra uncensored heretic v2 https://huggingface.co…
Runs locally from ~0.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3.5-9B-ultra-uncensored-heretic-v2-IQ4_NL.gguf | GGUF | IQ4_NL | 5.05 GB | Download |
| Qwen3.5-9B-ultra-uncensored-heretic-v2-IQ4_XS.gguf | GGUF | IQ4_XS | 4.84 GB | Download |
| Qwen3.5-9B-ultra-uncensored-heretic-v2-Q4_K_M.gguf | GGUF | Q4_K_M | 5.24 GB | Download |
| Qwen3.5-9B-ultra-uncensored-heretic-v2-Q4_K_S.gguf | GGUF | Q4_K_S | 4.98 GB | Download |
| imatrix.gguf | GGUF | GGUF | 4.9 MB | Download |
| mmproj-Qwen3.5-9B-ultra-uncensored-heretic-v2-f16.gguf | GGUF | F16 | 0.0 MB | Download |
Model Details
Model README
---
base_model: llmfan46/Qwen3.5-9B-ultra-uncensored-heretic-v2
tags:
- llama.cpp
- quantized
- imatrix
---
Auto-Quantized GGUF Model
This repository contains automated GGUF quantization files for llmfan46/Qwen3.5-9B-ultra-uncensored-heretic-v2.
The calibration data for the imatrix is targeted at Chinese novels and role-playing (RP), while preserving logic and common sense.
imatrix 的校准数据以中文的小说、角色扮演为目标,同时保留逻辑和常识。
📊 Perplexity Evaluation
(Tested against the provided calibration dataset)
- Base (F16/BF16): PPL = 14.2066 +/- 0.11553
- IQ4_XS: PPL = 12.0006 +/- 0.09439
- IQ4_NL: PPL = 12.0064 +/- 0.09447
- Q4_K_M: PPL = 11.9559 +/- 0.09373
Run nuofang/Qwen3.5-9B-ultra-uncensored-heretic-v2-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models