tooltd/Qwen3.6-27B-AutoRound-GGUF-NoMTP overview
This is a repository of GGUF files without MTP , It is old version before updating MTP re uploaded from sphaela/Qwen3.6 27B AutoRound GGUF https://huggingface.…
Runs locally from ~11.24 GB disk (12 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
license: apache-2.0
language:
- multilingual
base_model: Qwen/Qwen3.6-27B
tags:
- auto-round
- intel
- gguf
- quantization
---
This is a repository of GGUF files (without MTP , It is old version before updating MTP) re-uploaded from sphaela/Qwen3.6-27B-AutoRound-GGUF
Qwen3.6-27B GGUF (AutoRound Quantized)
This repository contains GGUF quantized versions of Qwen/Qwen3.6-27B created using Intel's AutoRound quantization method.
About AutoRound
AutoRound is an advanced quantization technique from Intel that aims to minimize accuracy loss through automated rounding optimization. The iterative calibration mode (--enable_alg_ext) runs gradient-based optimization for 200 iterations per block, finding optimal rounding thresholds that minimize reconstruction error.
---
Run tooltd/Qwen3.6-27B-AutoRound-GGUF-NoMTP with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models