Model Intelligence Sheet
lm-kit/qwq-32b-gguf overview
Model Summary This repository hosts quantized versions of the QWQ 32B reasoning model. Format: GGUF Converter: llama.cpp ba7654380a3c7c1b5ae154bea19134a3a9417a…
Runs locally from ~18.49 GB disk (24 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
license: apache-2.0
---
Model Summary
This repository hosts quantized versions of the QWQ-32B reasoning model.
Format: GGUF
Converter: llama.cpp ba7654380a3c7c1b5ae154bea19134a3a9417a1e
Quantizer: LM-Kit.NET 2025.3.3
For more detailed information on the base model, please visit the following link
Run lm-kit/qwq-32b-gguf with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models