GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

giocom/GLM-4.7-Flash-MTP-GGUF overview

GLM 4.7 Flash MTP ONLY GGUF

ggufmoequantizedllama.cppglmmtpspeculative-decodingtext-generationbase_model:zai-org/GLM-4.7-Flashbase_model:quantized:zai-org/GLM-4.7-Flashlicense:apache-2.0endpoints_compatibleregion:us

Runs locally from ~1.14 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
GLM-4.7-Flash-MTP-BF16.ggufGGUFBF163.57 GBDownload
GLM-4.7-Flash-MTP-Q4_K_M.ggufGGUFQ4_K_M1.14 GBDownload
GLM-4.7-Flash-MTP-Q8_0.ggufGGUFQ8_01.90 GBDownload

Model Details

Model IDgiocom/GLM-4.7-Flash-MTP-GGUF
Authorgiocom
Pipelinetext-generation
Licenseapache-2.0
Base modelzai-org/GLM-4.7-Flash
Last modified2026-08-05T12:51:54.000Z

Model README

---

license: apache-2.0

base_model: zai-org/GLM-4.7-Flash

tags:

  • moe
  • quantized
  • llama.cpp
  • glm
  • mtp
  • speculative-decoding
  • gguf

pipeline_tag: text-generation

---

GLM-4.7-Flash-MTP-ONLY-GGUF

Run giocom/GLM-4.7-Flash-MTP-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models