Model Intelligence Sheet
yomaytk/Qwen3-Next-80B-A3B-Instruct-MTP-HEAD-GGUF overview
Qwen3 Next 80B A3B Instruct MTP Head GGUF This repository contains an unofficial GGUF conversion of the MTP head weights extracted from Qwen/Qwen3 Next 80B A3B…
Runs locally from ~2.26 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
license: apache-2.0
base_model:
- Qwen/Qwen3-Next-80B-A3B-Instruct
library_name: gguf
tags:
- gguf
- llama.cpp
- qwen3-next
- mtp
---
Qwen3-Next-80B-A3B-Instruct MTP Head GGUF
This repository contains an unofficial GGUF conversion of the MTP head weights extracted from
Qwen/Qwen3-Next-80B-A3B-Instruct.
It includes:
- BF16 GGUF
- Q8_0 quantized GGUF
Run yomaytk/Qwen3-Next-80B-A3B-Instruct-MTP-HEAD-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models