aessedai/glm-5-gguf Q4_K_M GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
aessedai/glm-5-gguf overview
MoE-quants of GLM-5 (Q80 quantization default with routed experts quantized further) Note: running this GGUF requires pulling and compiling this llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/19460 More quants to come soon. | Quant | Size | Mixture | PPL | KLD | | :--------- | :--------- | :------- | :------- | :--------- | | Q4KM | 432.80 GiB (4.93 BPW) | Q80-Q4K-Q4K-Q5_K | 8.7486 ± 0.17123 | TBD |
Downloads
180
Likes
8
Pipeline
—
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
11 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| GLM-5-Q4_K_M-00001-of-00011.gguf | GGUF | Q4_K_M | 8.98 MB | Download |
| GLM-5-Q4_K_M-00002-of-00011.gguf | GGUF | Q4_K_M | 44.89 GB | Download |
| GLM-5-Q4_K_M-00003-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00004-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00005-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00006-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00007-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00008-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00009-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00010-of-00011.gguf | GGUF | Q4_K_M | 45.23 GB | Download |
| GLM-5-Q4_K_M-00011-of-00011.gguf | GGUF | Q4_K_M | 26.10 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"base_model": [
"zai-org/GLM-5"
],
"frontmatter": {
"base_model": [
"zai-org/GLM-5"
]
},
"hero_image_url": "",
"summary": "MoE-quants of GLM-5 (Q8_0 quantization default with routed experts quantized further) Note: running this GGUF requires pulling and compiling this llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/19460 More quants to come soon. | Quant | Size | Mixture | PPL | KLD | | :--------- | :--------- | :------- | :------- | :--------- | | Q4_K_M | 432.80 GiB (4.93 BPW) | Q8_0-Q4_K-Q4_K-Q5_K | 8.7486 ± 0.17123 | TBD |",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nbase_model:\n- zai-org/GLM-5\n---\n\nMoE-quants of GLM-5 (Q8_0 quantization default with routed experts quantized further)\n\nNote: running this GGUF requires pulling and compiling this llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/19460\n\nMore quants to come soon.\n\n| Quant | Size | Mixture | PPL | KLD | \n| :--------- | :--------- | :------- | :------- | :--------- |\n| Q4_K_M | 432.80 GiB (4.93 BPW) | Q8_0-Q4_K-Q4_K-Q5_K | 8.7486 ± 0.17123 | TBD |",
"related_quantizations": []
},
"tags": [
"gguf",
"base_model:zai-org/GLM-5",
"base_model:quantized:zai-org/GLM-5",
"endpoints_compatible",
"region:us",
"conversational"
],
"likes": 8,
"downloads": 180,
"gated": false,
"private": false,
"last_modified": "2026-02-12T19:52:19.000Z",
"created_at": "2026-02-12T06:28:16.000Z",
"pipeline_tag": "",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "698d73008ee9e6c9c63f8a7c",
"id": "AesSedai/GLM-5-GGUF",
"modelId": "AesSedai/GLM-5-GGUF",
"sha": "b1aea06f06fd0f807d84cbabe19da38cb891200b",
"createdAt": "2026-02-12T06:28:16.000Z",
"lastModified": "2026-02-12T19:52:19.000Z",
"author": "AesSedai",
"downloads": 180,
"likes": 8,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "",
"siblings_count": 13
}