oolfBER/Qwen3.6-27B-UD-Q4_K_XL-single-GGUF overview
Qwen3.6 27B UD Q4 K XL single file GGUF This repository intentionally contains only the exact GGUF used by the Hermes RunPod Serverless worker. Keeping one qua…
Runs locally from ~16.68 GB disk (24 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3.6-27B-UD-Q4_K_XL.gguf | GGUF | Q4_K_XL | 16.68 GB | Download |
Model Details
Model README
---
license: apache-2.0
base_model: unsloth/Qwen3.6-27B-MTP-GGUF
tags:
- gguf
- llama.cpp
- qwen3.6
---
Qwen3.6-27B UD-Q4_K_XL single-file GGUF
This repository intentionally contains only the exact GGUF used by the Hermes RunPod Serverless worker. Keeping one quantization here prevents RunPod cached-model provisioning from downloading every quantization in the upstream repository.
- Upstream: unsloth/Qwen3.6-27B-MTP-GGUF
- Upstream revision:
5cb35eb3dcbf52dbce5f87dbc64df6aaffadcace - File:
Qwen3.6-27B-UD-Q4_K_XL.gguf - Size:
17,909,097,600bytes - SHA-256:
4085665ee36d82a672a238a43f0e5643f2f0e39f2d7bd5d373f0ef10ecf53095
The model and quantization remain under the upstream Apache-2.0 license and attribution. This repository does not modify the model weights.
Run oolfBER/Qwen3.6-27B-UD-Q4_K_XL-single-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models