Model Intelligence Sheet
nif0/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUF overview
Pruned layers after 49 for H3. You need only the model for T2V and model + mmproj for I2V. place both the model and the mmproj file in text encoder folder. Can…
Runs locally from ~736.6 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
8 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3-VL-32B-Ultra-Heretic-H3-L0-49-IQ2_XS.gguf | GGUF | IQ2_XS | 6.91 GB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-L0-49-IQ2_XXS.gguf | GGUF | IQ2_XXS | 6.24 GB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-L0-49-IQ3_XS.gguf | GGUF | IQ3_XS | 9.52 GB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-L0-49-IQ3_XXS.gguf | GGUF | IQ3_XXS | 9.01 GB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-L0-49-IQ4_XS.gguf | GGUF | IQ4_XS | 12.49 GB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-L0-49-Q4_K_M.gguf | GGUF | Q4_K_M | 13.91 GB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-mmproj-Q8_0.gguf | GGUF | Q8_0 | 736.6 MB | Download |
| Qwen3-VL-32B-Ultra-Heretic-H3-mmproj-f16.gguf | GGUF | F16 | 1.11 GB | Download |
Model Details
Model README
---
license: apache-2.0
base_model:
- mradermacher/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-i1-GGUF
---
Pruned layers after 49 for H3.
You need only the model for T2V and model + mmproj for I2V. place both the model and the mmproj file in text_encoder folder.
Can be used in ComfyUI with my fork of City96's ComfyUI-GGUF.
https://github.com/Nif00/ComfyUI-GGUF
Run nif0/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models