Model Intelligence Sheet
giocom/GLM-4.7-Flash-MTP-GGUF overview
GLM 4.7 Flash MTP ONLY GGUF
Runs locally from ~1.14 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
| Model ID | giocom/GLM-4.7-Flash-MTP-GGUF |
|---|---|
| Author | giocom |
| Pipeline | text-generation |
| License | apache-2.0 |
| Base model | zai-org/GLM-4.7-Flash |
| Last modified | 2026-08-05T12:51:54.000Z |
Model README
---
license: apache-2.0
base_model: zai-org/GLM-4.7-Flash
tags:
- moe
- quantized
- llama.cpp
- glm
- mtp
- speculative-decoding
- gguf
pipeline_tag: text-generation
---
GLM-4.7-Flash-MTP-ONLY-GGUF
Run giocom/GLM-4.7-Flash-MTP-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models